xc-llm-ascend

Files

Icey 378e92a2a2 [Cherry-pick][0.11.0] Adapted to torch_npu.npu_fused_infer_attention_score (#4202 )

### What this PR does / why we need it?
Fixes a compatible bug with torch_npu.npu_fused_infer_attention_score
which is discribed in
https://github.com/vllm-project/vllm-ascend/issues/4020.
@momo609 tells us this solution.
cherry-pick: https://github.com/vllm-project/vllm-ascend/pull/4025

### Does this PR introduce _any_ user-facing change?
N/A

### How was this patch tested?
CI passed with new added/existing test.

Signed-off-by: Icey <1790571317@qq.com>

2025-11-17 10:56:23 +08:00

platform

[Cherry-pick][0.11.0] Adapted to torch_npu.npu_fused_infer_attention_score (#4202 )

2025-11-17 10:56:23 +08:00

worker

[0.11.0][MTP][Aclgraph] Fix the support aclgraph with MTP (#3912 )

2025-11-03 14:25:37 +08:00

__init__.py

[v0.11.0][Perf] Eliminating the zerolike operator through patch (#3632 )

2025-10-23 14:49:28 +08:00