Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Actions Projects Releases Wiki Activity
Files
389030a8f81f9d95e1b82bb3bef95c82d09d8ca3
xc-llm-ascend/tests/ut/quantization
History
1092626063 ceadc2788d Revert "[refactor]support gatingtopk operator generalization (#4356)" (#4873)
This reverts commit c4a11a745a.

ops npu_gating_top_k caused Qwen3-30B precision problem, so revert it.

Signed-off-by: 1092626063 <1092626063@qq.com>
2025-12-10 15:45:20 +08:00
..
test_quant_config.py
[Feat] Unquantized Linear to nz and control all nz-cast (#3356)
2025-10-14 17:39:26 +08:00
test_utils.py
[1/N][Refactor][Quantization] remove redundant quantizer class (#2680)
2025-09-04 11:35:14 +08:00
test_w4a4_flatquant_dynamic.py
[Refactor] Clean up w4a4_flatquant_dynamic implementation (#3440)
2025-10-17 23:53:19 +08:00
test_w4a8_dynamic.py
[0.11.0] [Cherry-pick #4058] Fixes Qwen3-Next enable nz accuracy problem (#4056)
2025-11-10 20:56:39 +08:00
test_w8a8_dynamic.py
[feat]: oproj tensor parallelism in pure DP and graph-mode scenarios. (#2167)
2025-09-07 10:31:32 +08:00
test_w8a8.py
Revert "[refactor]support gatingtopk operator generalization (#4356)" (#4873)
2025-12-10 15:45:20 +08:00
Powered by Gitea Version: 1.24.3 Page: 66ms Template: 6ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API