初始化项目,由ModelHub XC社区提供模型

Model: Phoebe13/Video-MTR
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-09-08 10:45:17 +08:00
commit 5bd9f1db39
18 changed files with 1168 additions and 0 deletions

20
README.md Normal file
View File

@@ -0,0 +1,20 @@
---
license: apache-2.0
language:
- en
base_model:
- Qwen/Qwen2.5-VL-7B-Instruct
pipeline_tag: visual-question-answering
---
# Video-MTR
Video-MTR: Reinforced Multi-Turn Reasoning for Long Video Understanding.
This checkpoint extends Qwen2.5-VL-7B-Instruct with a multi-turn frame-retrieval
policy trained via PPO, with an 80-frame video input budget.
## References
* [Paper](https://arxiv.org/abs/2508.20478)
* [Code](https://github.com/Xyuan13/Video-MTR)