ldh2020
|
8261a09e2a
|
[Kernel] Optimize the selection and update OP of ssm state
|
2025-12-21 15:45:32 +08:00 |
|
ldh2020
|
b97c781300
|
[Kernel] Optimize the recurrent op
|
2025-12-21 11:22:06 +08:00 |
|
Xinyu Dong
|
5a75795ade
|
[Model] Update llama.py
Remove redundancy
|
2025-12-15 21:28:56 +08:00 |
|
Xinyu Dong
|
7c7d0326c5
|
[Model] registry llama.py to vLLM
|
2025-12-15 21:21:28 +08:00 |
|
Xinyu Dong
|
ca059110b3
|
[Model] Supporet llama3 on v0.11.0
FULL AND PIECEWISE GRAPH ENBALE
|
2025-12-15 21:20:44 +08:00 |
|
chenyili
|
7c22d621fb
|
提交vllm0.11.0开发分支
|
2025-12-10 17:51:24 +08:00 |
|
zhaoyingzhuo
|
b614823125
|
[chore] Remove obsolete comments
|
2025-12-10 15:52:23 +08:00 |
|
dongxinyu03
|
c728e52505
|
Initial commit for vLLM-Kunlun Plugin
|
2025-12-10 12:05:39 +08:00 |
|