Skip to content

Author

Xiaodan Liang

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

When Prompts Become Pixels: Prompt-Region Grounding for Multimodal Reasoning

This work introduces Visualized Task Semantics (VTS), a controlled intervention that moves the question into the image while keeping the source problem and answer fixed, and requires no OCR or region metadata at inference.

Yongxin Wang, Ruizhe Zhou, Yueling Tang et al. · 0 citations