Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Projects Releases Wiki Activity
Files
203b4e67773d74ff196cb42d33a716dab939f83c
xc-llm-ascend/tests/ut/quantization
History
Mengqing Cao 517fd9272d Revert "drop ascend scheduler" (#4580)
Reverts vllm-project/vllm-ascend#4498
- vLLM version: v0.11.2
- vLLM main: https://github.com/vllm-project/vllm/commit/v0.11.2
2025-11-29 22:20:48 +08:00
..
test_quant_config.py
[Quantization] Support compressed tensors w8a8 static and w8a8 dynamic weight (#4036)
2025-11-28 14:09:39 +08:00
test_utils.py
[1/N][Refactor][Quantization] remove redundant quantizer class (#2680)
2025-09-04 11:35:14 +08:00
test_w4a4_flatquant_dynamic.py
[Refactor] Clean up w4a4_flatquant_dynamic implementation (#3440)
2025-10-17 23:53:19 +08:00
test_w4a8_dynamic.py
[Feat][BugFix]Support the Qwen3-Next-80B-A3B-Instruct quantization model&Fix the NZ issue (#4245)
2025-11-21 10:42:56 +08:00
test_w8a8_dynamic.py
Revert "drop ascend scheduler" (#4580)
2025-11-29 22:20:48 +08:00
test_w8a8.py
[refact] unified soc_version code (#4359)
2025-11-26 14:28:55 +08:00
Powered by Gitea Version: 1.24.3 Page: 556ms Template: 86ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API