Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Projects Releases Wiki Activity
Files
96b2cdf6d8d8f577f85f4a5c96e1c8fde532fb8b
xc-llm-ascend/tests/ut/quantization
History
wangxiyuan 400af665e6 [CI] Drop ascend scheduler from test (#4613)
Drop ascend scheduler from test

- vLLM version: v0.11.2

Signed-off-by: wangxiyuan <wangxiyuan1007@gmail.com>
2025-12-02 13:18:17 +08:00
..
test_quant_config.py
[Quantization] Support compressed tensors w8a8 static and w8a8 dynamic weight (#4036)
2025-11-28 14:09:39 +08:00
test_utils.py
[1/N][Refactor][Quantization] remove redundant quantizer class (#2680)
2025-09-04 11:35:14 +08:00
test_w4a4_flatquant_dynamic.py
[Refactor] Clean up w4a4_flatquant_dynamic implementation (#3440)
2025-10-17 23:53:19 +08:00
test_w4a8_dynamic.py
[Feat][BugFix]Support the Qwen3-Next-80B-A3B-Instruct quantization model&Fix the NZ issue (#4245)
2025-11-21 10:42:56 +08:00
test_w8a8_dynamic.py
[CI] Drop ascend scheduler from test (#4613)
2025-12-02 13:18:17 +08:00
test_w8a8.py
[refact] unified soc_version code (#4359)
2025-11-26 14:28:55 +08:00
Powered by Gitea Version: 1.24.3 Page: 584ms Template: 91ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API