Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Actions Projects Releases Wiki Activity
Files
5d13bbe796d9bbc15e7db5de87f2b9da72f7b4a5
xc-llm-ascend/vllm_ascend/torchair/models
History
liziyu 464270e4ca Remove useless PD check in deepseek (#3161)
### What this PR does / why we need it?
Remove useless PD check in deepseek

### How was this patch tested?


- vLLM version: v0.10.2
- vLLM main:
f225ea7dd9

Signed-off-by: wangxiaoteng <wangxiaoteng@huawei.com>
Co-authored-by: wangxiaoteng <wangxiaoteng@huawei.com>
2025-09-24 23:25:47 +08:00
..
__init__.py
[1/N][refactor] torchair deepseek modeling refactor (#2384)
2025-08-18 15:00:37 +08:00
qwen2.py
[KVCache][Bugfix] Fix kv cache initialization error of attention layer (#3113)
2025-09-24 11:32:34 +08:00
qwen3_moe.py
[Refactor] [SP]The sequence parallelism characteristics in the MoE and Dense models are integrated into a single solution. (#3085)
2025-09-24 11:29:59 +08:00
torchair_deepseek_mtp.py
[KVCache][Bugfix] Fix kv cache initialization error of attention layer (#3113)
2025-09-24 11:32:34 +08:00
torchair_deepseek_v2.py
Remove useless PD check in deepseek (#3161)
2025-09-24 23:25:47 +08:00
torchair_deepseek_v3.py
[1/N][refactor] torchair deepseek modeling refactor (#2384)
2025-08-18 15:00:37 +08:00
torchair_pangu_moe.py
[KVCache][Bugfix] Fix kv cache initialization error of attention layer (#3113)
2025-09-24 11:32:34 +08:00
Powered by Gitea Version: 1.24.3 Page: 190ms Template: 8ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API