Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Actions Projects Releases Wiki Activity
Files
2a9d02e08039749cf811c5fb190d4c8a950d792d
xc-llm-ascend/vllm_ascend/torchair/models
History
liziyu 464270e4ca Remove useless PD check in deepseek (#3161)
### What this PR does / why we need it?
Remove useless PD check in deepseek

### How was this patch tested?


- vLLM version: v0.10.2
- vLLM main:
f225ea7dd9

Signed-off-by: wangxiaoteng <wangxiaoteng@huawei.com>
Co-authored-by: wangxiaoteng <wangxiaoteng@huawei.com>
2025-09-24 23:25:47 +08:00
..
__init__.py
[1/N][refactor] torchair deepseek modeling refactor (#2384)
2025-08-18 15:00:37 +08:00
qwen2.py
[KVCache][Bugfix] Fix kv cache initialization error of attention layer (#3113)
2025-09-24 11:32:34 +08:00
qwen3_moe.py
[Refactor] [SP]The sequence parallelism characteristics in the MoE and Dense models are integrated into a single solution. (#3085)
2025-09-24 11:29:59 +08:00
torchair_deepseek_mtp.py
[KVCache][Bugfix] Fix kv cache initialization error of attention layer (#3113)
2025-09-24 11:32:34 +08:00
torchair_deepseek_v2.py
Remove useless PD check in deepseek (#3161)
2025-09-24 23:25:47 +08:00
torchair_deepseek_v3.py
[1/N][refactor] torchair deepseek modeling refactor (#2384)
2025-08-18 15:00:37 +08:00
torchair_pangu_moe.py
[KVCache][Bugfix] Fix kv cache initialization error of attention layer (#3113)
2025-09-24 11:32:34 +08:00
Powered by Gitea Version: 1.24.3 Page: 119ms Template: 6ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API