Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Projects Releases Wiki Activity
Files
6360eb1deab4e118116f563fc3d39167c3d6e09e
xc-llm-ascend/tests/ut/quantization
History
Mengqing Cao 517fd9272d Revert "drop ascend scheduler" (#4580)
Reverts vllm-project/vllm-ascend#4498
- vLLM version: v0.11.2
- vLLM main: https://github.com/vllm-project/vllm/commit/v0.11.2
2025-11-29 22:20:48 +08:00
..
test_quant_config.py
[Quantization] Support compressed tensors w8a8 static and w8a8 dynamic weight (#4036)
2025-11-28 14:09:39 +08:00
test_utils.py
[1/N][Refactor][Quantization] remove redundant quantizer class (#2680)
2025-09-04 11:35:14 +08:00
test_w4a4_flatquant_dynamic.py
[Refactor] Clean up w4a4_flatquant_dynamic implementation (#3440)
2025-10-17 23:53:19 +08:00
test_w4a8_dynamic.py
[Feat][BugFix]Support the Qwen3-Next-80B-A3B-Instruct quantization model&Fix the NZ issue (#4245)
2025-11-21 10:42:56 +08:00
test_w8a8_dynamic.py
Revert "drop ascend scheduler" (#4580)
2025-11-29 22:20:48 +08:00
test_w8a8.py
[refact] unified soc_version code (#4359)
2025-11-26 14:28:55 +08:00
Powered by Gitea Version: 1.24.3 Page: 396ms Template: 7ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API