CURV reframes chart question answering as multi-step visually grounded reasoning, coordinating each logical step with dynamic spatial attention. The authors also introduce CCQA, a synthetically scalable three-level curriculum that progresses from single-operation problems to complex multi-chart compositional tasks. According to the abstract, CURV improves over baselines by up to 20.50%, transfers to real-world benchmarks with gains up to 12.30%, and improves out-of-domain multimodal reasoning tasks by up to 10.20%. These are reported peak gains; the supplied material does not identify the exact benchmarks, metrics, baseline strengths, or absolute scores.
No heat snapshots are available in the last 24 hours.