RewardMap: Tackling Sparse Rewards in Fine-grained Visual Reasoning via Multi-Stage Reinforcement Learning
Published in International Conference on Learning Representations (ICLR), 2025
RewardMap combines difficulty-aware dense rewards with a multi-stage reinforcement learning curriculum that progresses from visual perception to complex reasoning. The framework improves spatial, fine-grained visual, and general reasoning performance across six benchmarks.
