wangxiyuan
c556038ef0
[New model] Qwen3-next support ( #2917 )
...
### What this PR does / why we need it?
Add Qwen3-next support.
### Does this PR introduce _any_ user-facing change?
Yes, users can use Qwen3 next.
Related doc: https://github.com/vllm-project/vllm-ascend/pull/2916 the
tutorial will be ready in
[here](https://vllm-ascend.readthedocs.io/en/latest/tutorials/multi_npu_qwen3_next.html )
### How was this patch tested?
Doc CI passed
Related: https://github.com/vllm-project/vllm-ascend/issues/2884
Co-Authored-By: Angazenn <supperccell@163.com >
Co-Authored-By: zzzzwwjj <1183291235@qq.com >
Co-Authored-By: MengqingCao <cmq0113@163.com >
Co-Authored-By: linfeng-yuan <1102311262@qq.com >
Co-Authored-By: hust17yixuan <303660421@qq.com >
Co-Authored-By: SunnyLee219 <3294305115@qq.com >
Co-Authored-By: maoxx241 <maoxx241@umn.edu >
- vLLM version: v0.10.2
- vLLM main:
b834b4cbf1
---------
Signed-off-by: MengqingCao <cmq0113@163.com >
Signed-off-by: wangxiyuan <wangxiyuan1007@gmail.com >
Signed-off-by: Angazenn <supperccell@163.com >
Signed-off-by: Your Name <you@example.com >
Signed-off-by: zzzzwwjj <1183291235@qq.com >
Signed-off-by: linfeng-yuan <1102311262@qq.com >
Signed-off-by: hust17yixuan <303660421@qq.com >
Co-authored-by: MengqingCao <cmq0113@163.com >
Co-authored-by: Angazenn <supperccell@163.com >
Co-authored-by: Your Name <you@example.com >
Co-authored-by: zzzzwwjj <1183291235@qq.com >
Co-authored-by: linfeng-yuan <1102311262@qq.com >
Co-authored-by: hust17yixuan <303660421@qq.com >
2025-09-16 01:17:42 +08:00
1092626063
5b3646ab21
[FEATURE][MTP] Support MTP > 1 ( #2708 )
...
### What this PR does / why we need it?
[RFC:Support MTP > 1 for
DeepSeek](https://github.com/vllm-project/vllm-ascend/issues/2745 )
- [x] dp1 tp16
- [x] dp4 tp4
- [x] dp2 tp 8
- [x] torchair graph
- vLLM version: v0.10.1.1
- vLLM main:
c9f7081f9c
Signed-off-by: 1092626063 <1092626063@qq.com >
2025-09-05 09:11:22 +08:00
Icey
d4370ebc42
[Refactor] Refactor Spec Decode ( #2668 )
...
### What this PR does / why we need it?
Refactor spec decode
### Does this PR introduce _any_ user-facing change?
N/A
### How was this patch tested?
CI passed with new added/existing test.
- vLLM version: v0.10.1.1
- vLLM main:
6997a25ac6
---------
Signed-off-by: wangxiyuan <wangxiyuan1007@gmail.com >
Signed-off-by: Icey <1790571317@qq.com >
Co-authored-by: wangxiyuan <wangxiyuan1007@gmail.com >
2025-09-04 11:34:47 +08:00