ModelHub XC 26b2b4c01e 初始化项目,由ModelHub XC社区提供模型
Model: miniHui/Geo-R1
Source: Original Platform
2026-09-05 03:36:18 +08:00

library_name, base_model, pipeline_tag, license
library_name base_model pipeline_tag license
transformers
Qwen/Qwen2.5-VL-7B-Instruct
image-text-to-text mit

Geo-R1: Unlocking VLM Geospatial Reasoning with Cross-View Reinforcement Learning

This repository contains the Geo-R1 model, a reasoning-centric post-training framework that unlocks geospatial reasoning in vision-language models, as introduced in the paper:

Geo-R1: Unlocking VLM Geospatial Reasoning with Cross-View Reinforcement Learning

Geo-R1 combines "thinking scaffolding" (supervised fine-tuning on synthetic chain-of-thought exemplars) and an "elevating" stage using GRPO-based reinforcement learning on a weakly-supervised cross-view pairing proxy. This approach enables models to connect visual cues with geographic priors and harness reasoning for accurate prediction, achieving state-of-the-art performance across various geospatial reasoning benchmarks.

Description
Model synced from source: miniHui/Geo-R1
Readme 2 MiB