Commit Graph

8 Commits

Author SHA1 Message Date
ldh2020
8261a09e2a [Kernel] Optimize the selection and update OP of ssm state 2025-12-21 15:45:32 +08:00
ldh2020
b97c781300 [Kernel] Optimize the recurrent op 2025-12-21 11:22:06 +08:00
Xinyu Dong
5a75795ade [Model] Update llama.py
Remove redundancy
2025-12-15 21:28:56 +08:00
Xinyu Dong
7c7d0326c5 [Model] registry llama.py to vLLM 2025-12-15 21:21:28 +08:00
Xinyu Dong
ca059110b3 [Model] Supporet llama3 on v0.11.0
FULL AND PIECEWISE GRAPH ENBALE
2025-12-15 21:20:44 +08:00
chenyili
7c22d621fb 提交vllm0.11.0开发分支 2025-12-10 17:51:24 +08:00
zhaoyingzhuo
b614823125 [chore] Remove obsolete comments 2025-12-10 15:52:23 +08:00
dongxinyu03
c728e52505 Initial commit for vLLM-Kunlun Plugin 2025-12-10 12:05:39 +08:00