sglang

Author	SHA1	Message	Date
lambert0312	471650dee0	Fix broadcast use cuda device lead to memory capacity unbalanced (#5416 )	2025-04-15 02:47:26 -07:00
tianlian yi	bc92107b03	Support server based rollout in Verlengine (#4848 ) Co-authored-by: Jin Pan <jpan236@wisc.edu> Co-authored-by: Chayenne <zhaochen20@outlook.com> Co-authored-by: Jinn <47354855+jhinpan@users.noreply.github.com>	2025-04-12 10:07:52 -07:00
XinyuanTong	d09a51f1f6	[feat&refactor] Enhance multimodal input support with refactor io_struct (#4938 ) Signed-off-by: Xinyuan Tong <justinning0323@outlook.com>	2025-04-08 14:48:07 -07:00
fzyzcjy	92bb49a7f9	Patch PyTorch's bug that cross-process tensor transfer will lead to wrong device (#4565 )	2025-03-27 00:22:33 -07:00
Lianmin Zheng	ac2387279e	Support penalty in overlap mode; return logprob with chunked prefill; improve benchmark scripts (#3988 ) Co-authored-by: SangBin Cho <rkooo567@gmail.com> Co-authored-by: dhou-xai <dhou@x.ai> Co-authored-by: Hanming Lu <hanming_lu@berkeley.edu>	2025-03-03 00:12:04 -08:00
fzyzcjy	e3e0bc50a9	[Feature] SPMD for SGLang + Verl (#3852 )	2025-02-28 09:53:10 -08:00