This website requires JavaScript.
Explore
Help
Register
Sign In
dylanyunlong
/
project_6
Watch
1
Star
0
Fork
0
You've already forked project_6
Code
Issues
Pull Requests
Actions
Projects
Releases
Wiki
Activity
Files
3bee73207e616b1ce83db35e5f42fb6add1f427a
project_6
/
ex_engine
/
xllm_kernels
History
Claude
3bee73207e
fix: add cuda_runtime.h to hgemm_bind.cpp for cudaStream_t
2026-08-14 16:33:29 +00:00
..
cuda
fix: add cuda_runtime.h to hgemm_bind.cpp for cudaStream_t
2026-08-14 16:33:29 +00:00
ilu
feat: import CUDA kernels from xllm/CCCL/FLA upstream repos
2026-08-14 07:48:52 +00:00
npu
fix(critical): fold max_completion_tokens + max_num_seqs=2 + max_model_len=80000 + xllm_latest layer import
2026-08-13 03:19:39 +00:00
build_test_hgemm.sh
feat: hgemm_blocktiling.cu — FP16 GEMM kernel for MoE expert dispatch on BI-V100
2026-08-14 16:22:00 +00:00