Logo
Explore Help
Register Sign In
EngineX-Hygon/sglang
5
0
Fork 0
You've already forked sglang
Code Issues Pull Requests Actions 7 Projects Releases Wiki Activity
Files
229d2b95f19573ece9c1c5d6b357df9874e04f59
sglang/python/sglang/srt/lora
History
Lifu Huang 9241f4fd20 Move cached kernel to srt.utils (#10776)
2025-09-22 23:00:36 -07:00
..
backend
[3/4] Speed up CSGMV backend perf by 10% through dynamic chunking + kernel optimization (#10592)
2025-09-20 22:47:48 -07:00
triton_ops
Move cached kernel to srt.utils (#10776)
2025-09-22 23:00:36 -07:00
layers.py
[1/2] Refactor LoRA to support backend-specific batch preprocessing. (#10251)
2025-09-10 09:58:37 -07:00
lora_config.py
[Fix] Fix bugs and refactor codes in lora for better scalability. (#3652)
2025-02-20 11:51:57 -08:00
lora_manager.py
[3/4] Speed up CSGMV backend perf by 10% through dynamic chunking + kernel optimization (#10592)
2025-09-20 22:47:48 -07:00
lora_registry.py
Support pinning adapter via server args. (#9249)
2025-08-20 16:25:01 -07:00
lora.py
[2/2] Introduce Chunked-SGMV kernels and corresponding LoRA backend for improved performance (#10286)
2025-09-15 16:04:03 -07:00
mem_pool.py
support Llama4 with non uniformed intermediate size across layers for… (#10047)
2025-09-05 17:28:15 -07:00
utils.py
[3/4] Speed up CSGMV backend perf by 10% through dynamic chunking + kernel optimization (#10592)
2025-09-20 22:47:48 -07:00
Powered by Gitea Version: 1.24.3 Page: 169ms Template: 6ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API