Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Actions Projects Releases Wiki Activity
Files
a686f2962a5d8c8587e841c8c8b2ac2c866c1ce8
xc-llm-ascend/tests/ut/quantization
History
wangxiyuan a2e4c3fe78 Revert "[cherry-pick][refactor]support gatingtopk operator generalization (#4050)" (#4352)
This reverts commit c87a77e8b4.

it breaks ops e2e test

Signed-off-by: wangxiyuan <wangxiyuan1007@gmail.com>
2025-11-21 23:03:20 +08:00
..
test_quant_config.py
[Feat] Unquantized Linear to nz and control all nz-cast (#3356)
2025-10-14 17:39:26 +08:00
test_utils.py
[1/N][Refactor][Quantization] remove redundant quantizer class (#2680)
2025-09-04 11:35:14 +08:00
test_w4a4_flatquant_dynamic.py
[Refactor] Clean up w4a4_flatquant_dynamic implementation (#3440)
2025-10-17 23:53:19 +08:00
test_w4a8_dynamic.py
[0.11.0] [Cherry-pick #4058] Fixes Qwen3-Next enable nz accuracy problem (#4056)
2025-11-10 20:56:39 +08:00
test_w8a8_dynamic.py
[feat]: oproj tensor parallelism in pure DP and graph-mode scenarios. (#2167)
2025-09-07 10:31:32 +08:00
test_w8a8.py
Revert "[cherry-pick][refactor]support gatingtopk operator generalization (#4050)" (#4352)
2025-11-21 23:03:20 +08:00
Powered by Gitea Version: 1.24.3 Page: 72ms Template: 12ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API