Logo
Explore Help
Register Sign In
EngineX/xc-llm-ascend
3
0
Fork 0
You've already forked xc-llm-ascend
Code Issues Pull Requests Projects Releases Wiki Activity
Files
18eefc23c3fd7275c6d2a9f540de80f2bf5e7f5f
xc-llm-ascend/vllm_ascend/patch/worker
History
Shanshan Shen 2a19215e5f [MM][Model] Remove Qwen2-VL modeling files (#4534)
### What this PR does / why we need it?

Following https://github.com/vllm-project/vllm-ascend/pull/4349, remove
Qwen2-VL modeling files.


- vLLM version: v0.11.2
- vLLM main: https://github.com/vllm-project/vllm/commit/v0.11.2

---------

Signed-off-by: shen-shanshan <467638484@qq.com>
2025-11-29 18:07:01 +08:00
..
__init__.py
[MM][Model][Perf] Remove Qwen2.5-VL modeling files and add patch for VisionAttention (#4349)
2025-11-28 14:23:00 +08:00
patch_distributed.py
[Refactor] refactor patch module (#3555)
2025-10-21 20:19:46 +08:00
patch_minicpm.py
[Refactor] refactor patch module (#3555)
2025-10-21 20:19:46 +08:00
patch_multimodal_merge.py
[Refactor] refactor patch module (#3555)
2025-10-21 20:19:46 +08:00
patch_qwen2_5_vl.py
[MM][Model] Remove Qwen2-VL modeling files (#4534)
2025-11-29 18:07:01 +08:00
patch_roberta.py
[1/N][Refactor] Refactor code to adapt with vllm main (#3612)
2025-10-24 16:55:08 +08:00
patch_rope.py
[MM][Model][Perf] Remove Qwen2.5-VL modeling files and add patch for VisionAttention (#4349)
2025-11-28 14:23:00 +08:00
patch_triton.py
【OPS】qwen3-next support triton chunk_gated_delta_rule ops (#4070)
2025-11-28 20:55:43 +08:00
patch_weight_loader.py
Drop 0.11.0 support (#4377)
2025-11-24 17:08:20 +08:00
Powered by Gitea Version: 1.24.3 Page: 135ms Template: 7ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API