Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Projects Releases Wiki Activity
Files
400af665e6b031f4b731b904e8bd0bb878ceb42f
xc-llm-ascend/tests/ut/quantization
History
wangxiyuan 400af665e6 [CI] Drop ascend scheduler from test (#4613)
Drop ascend scheduler from test

- vLLM version: v0.11.2

Signed-off-by: wangxiyuan <wangxiyuan1007@gmail.com>
2025-12-02 13:18:17 +08:00
..
test_quant_config.py
[Quantization] Support compressed tensors w8a8 static and w8a8 dynamic weight (#4036)
2025-11-28 14:09:39 +08:00
test_utils.py
[1/N][Refactor][Quantization] remove redundant quantizer class (#2680)
2025-09-04 11:35:14 +08:00
test_w4a4_flatquant_dynamic.py
[Refactor] Clean up w4a4_flatquant_dynamic implementation (#3440)
2025-10-17 23:53:19 +08:00
test_w4a8_dynamic.py
[Feat][BugFix]Support the Qwen3-Next-80B-A3B-Instruct quantization model&Fix the NZ issue (#4245)
2025-11-21 10:42:56 +08:00
test_w8a8_dynamic.py
[CI] Drop ascend scheduler from test (#4613)
2025-12-02 13:18:17 +08:00
test_w8a8.py
[refact] unified soc_version code (#4359)
2025-11-26 14:28:55 +08:00
Powered by Gitea Version: 1.24.3 Page: 294ms Template: 9ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API