Skip to content

Author

Xumeng Han

We have 2 of 20 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

GeoBridge: Decoupled Semantic Conditioning for Generative Image Geolocalization

Multimodal large language models (MLLMs) have advanced image geolocalization mainly by improving how they reason about geographic cues. How that reasoning isdecoded into coordinates, however, has lagged behind. Predicting a place name for a geocoding API is discrete and lossy: it ignores image evidence and collapses mu...

Zhi-Yang Dou, Xumeng Han, Feng-De Peng et al. · 0 citations
#artificial intelligence Open access Jun 2025

Evidence-Grounded Trustworthy Multimodal Reasoning and Evaluation Benchmark in Complex Urban Scenes

Evidence-grounded visual reasoning (EGVOR) is proposed, departing from implicit inference, EGVOR reformulates reasoning as the explicit generation of Evidence Atoms–structured triplets that enforce strict spatial-semantic alignment.

Zhao-Yang Wei, Bo-Wen Jiang, Xumeng Han et al. · 6 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.