xc-llm-ascend/configs at e17006077ab1e1cf52a659952e3e187647e4be32 - xc-llm-ascend - Gitea: Git with a cup of tea

EngineX/xc-llm-ascend

Files

History

Nagisa125 2cb9195ff0 [Releases/v0.18.0][CI] Updated the parameters for the single-node test to fix the OOM issue for DeepSeek-V3.2 (#7862 )

### What this PR does / why we need it?
Fix the OOM (Out-of-Memory) error in the single-node-deepseek-v3-2-w8a8
nightly test of vllm-ascend:

- Reduced the value of HCCL_BUFFSIZE

- Lowered the gpu-memory-utilization

Optimize service-side performance:
Updated service-oriented configuration parameters (e.g., max-num-seqs,
cudagraph_capture_sizes, batch_size) to improve the inference
performance,so that the performance is closer to the optimal performance
of the current mainline.
Align performance baseline with main branch:
Updated the performance baseline according to the latest performance
data

### Does this PR introduce _any_ user-facing change?
No.

### How was this patch tested?
The test has passed.

https://github.com/vllm-project/vllm-ascend/actions/runs/23734079080/job/69134387320?pr=7793

---------

Signed-off-by: wyh145 <1987244901@qq.com>

2026-04-01 10:28:46 +08:00

..

DeepSeek-R1-0528-W8A8.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

DeepSeek-R1-W8A8-HBM.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

DeepSeek-V3.2-W8A8.yaml

[Releases/v0.18.0][CI] Updated the parameters for the single-node test to fix the OOM issue for DeepSeek-V3.2 (#7862 )

2026-04-01 10:28:46 +08:00

GLM-4.5.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

GLM-4.7.yaml

[CI] Add nightly CI test cases for the GLM-4.7 model. (#7391 )

2026-03-19 16:43:29 +08:00

GLM-5.yaml

[CI] Add nightly CI test cases for the GLM-5 (#7429 )

2026-03-23 19:14:19 +08:00

Kimi-K2-Thinking.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Kimi-K2.5.yaml

[CI] Add nightly CI test cases for the Kimi-K2.5 (#7416 )

2026-03-19 11:02:29 +08:00

MiniMax-M2.5-A3.yaml

[DOC] MiniMax-M2.5 model intro (#7296 )

2026-03-18 20:14:36 +08:00

MTPX-DeepSeek-R1-0528-W8A8.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Prefix-Cache-DeepSeek-R1-0528-W8A8.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Prefix-Cache-Qwen3-32B-Int8.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen2.5-VL-7B-Instruct-EPD.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen2.5-VL-7B-Instruct.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen2.5-VL-32B-Instruct.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen3-30B-A3B-W8A8.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen3-32B-Int8-A2.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen3-32B-Int8-A3-Feature-Stack3.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen3-32B-Int8.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen3-32B.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen3-235B-A22B-W8A8.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00

Qwen3-Next-80B-A3B-Instruct-A2.yaml

Fix Qwen3Next CI Config (#7561 )

2026-03-24 17:08:17 +08:00

Qwen3-Next-80B-A3B-Instruct-W8A8.yaml

Fix Qwen3Next CI Config (#7561 )

2026-03-24 17:08:17 +08:00

Qwen3-Next-80B-A3B-Instruct.yaml

Fix Qwen3Next CI Config (#7561 )

2026-03-24 17:08:17 +08:00

QwQ-32B.yaml

[Nightly][Refactor]Migrate nightly single-node model tests from .py to .yaml (#6503 )

2026-03-03 20:13:43 +08:00