Urban Socio-Semantic Segmentation with Vision-Language Reasoning Paper • 2601.10477 • Published 4 days ago • 150
Taming Hallucinations: Boosting MLLMs' Video Understanding via Counterfactual Video Generation Paper • 2512.24271 • Published 20 days ago • 59
Thinking with Map: Reinforced Parallel Map-Augmented Agent for Geolocalization Paper • 2601.05432 • Published 10 days ago • 159