Logo
Explore Help
Register Sign In
EngineX-Hygon/sglang
5
0
Fork 0
You've already forked sglang
Code Issues Pull Requests Actions 7 Projects Releases Wiki Activity
Files
eefcbdd3533b065b950276ce23c8ab7a4f69bd99
sglang/sgl-kernel/tests
History
Xiaoyu Zhang bb418ced80 optimize per token group quant fp8 (#3490)
2025-02-11 22:19:05 +08:00
..
test_activation.py
sync flashinfer and update sgl-kernel tests (#3081)
2025-01-23 21:13:55 +08:00
test_bmm_fp8.py
feat: integrate bmm_fp8 kernel into sgl-kernel (#3056)
2025-01-23 00:39:38 +08:00
test_fp8_gemm.py
support w8a8 fp8 kernel with CUTLASS (#3047)
2025-01-26 15:46:51 +08:00
test_int8_gemm.py
Support sm90 Int8 gemm (#3035)
2025-01-21 22:21:54 +08:00
test_lightning_attention_decode.py
sync flashinfer and update sgl-kernel tests (#3081)
2025-01-23 21:13:55 +08:00
test_moe_align.py
clean moe align block kernel code and add acc test (#3332)
2025-02-06 16:42:36 +08:00
test_norm.py
cleanup sgl-kernel kernels (#3175)
2025-01-27 19:11:01 +08:00
test_per_token_group_quant_fp8.py
optimize per token group quant fp8 (#3490)
2025-02-11 22:19:05 +08:00
test_rotary_embedding.py
[kernel] Fix position ids in rope (#3173)
2025-01-27 17:09:51 +08:00
test_sampling.py
feat: integrate sampling kernels into sgl-kernel (#3086)
2025-01-24 01:54:47 +08:00
test_speculative_sampling.py
support speculative decoding kernel in sgl-kernel (#3373)
2025-02-07 20:29:51 +08:00
test_trt_reduce.py
optimize custom allreduce kernel (#2904)
2025-01-16 03:04:25 +08:00
Powered by Gitea Version: 1.24.3 Page: 77ms Template: 5ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API