TerraScope: Pixel-Grounded Visual Reasoning for Earth Observation
Yan Shu, Nicu Sebe, Bin Ren et al. · 19 Mar 2026
Vision-language models (VLMs) have shown promise in earth observation (EO), yet they struggle with tasks that require grounding complex spatial reasoning in precise pixel-level visual representations.