Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Projects Releases Wiki Activity
Files
2b3bfe432e886b4773ef5cfa33a0e69b2c7d5b6d
xc-llm-ascend/tests/ut/quantization
History
Mengqing Cao 517fd9272d Revert "drop ascend scheduler" (#4580)
Reverts vllm-project/vllm-ascend#4498
- vLLM version: v0.11.2
- vLLM main: https://github.com/vllm-project/vllm/commit/v0.11.2
2025-11-29 22:20:48 +08:00
..
test_quant_config.py
[Quantization] Support compressed tensors w8a8 static and w8a8 dynamic weight (#4036)
2025-11-28 14:09:39 +08:00
test_utils.py
[1/N][Refactor][Quantization] remove redundant quantizer class (#2680)
2025-09-04 11:35:14 +08:00
test_w4a4_flatquant_dynamic.py
[Refactor] Clean up w4a4_flatquant_dynamic implementation (#3440)
2025-10-17 23:53:19 +08:00
test_w4a8_dynamic.py
[Feat][BugFix]Support the Qwen3-Next-80B-A3B-Instruct quantization model&Fix the NZ issue (#4245)
2025-11-21 10:42:56 +08:00
test_w8a8_dynamic.py
Revert "drop ascend scheduler" (#4580)
2025-11-29 22:20:48 +08:00
test_w8a8.py
[refact] unified soc_version code (#4359)
2025-11-26 14:28:55 +08:00
Powered by Gitea Version: 1.24.3 Page: 141ms Template: 23ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API