Skip to content

Author

Yaokang Wu

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#computer vision Preprint Sep 2026

InfoEdit: Probing Global Layout Reasoning in Infographic Editing

Multimodal foundation models edit natural photographs at production quality, yet the same models struggle with structured visual content such as infographics. Unlike photographs, infographics encode information through logical relations; editing one element often requires surrounding elements to be adapted. We refer to...

Cheng Yang, Chu-Fan Shi, Hui-Juan Wang et al. · 0 citations
#computer vision Preprint Feb 2026

UReason: Benchmarking Reasoning-to-Generation Alignment in Unified Multimodal Models

UReason is introduced, a benchmark for evaluating reasoning-to-generation alignment in this paradigm, consisting of 2,000 human-curated and human-verified instances spanning five reasoning-intensive tasks: Code, Arithmetic, Spatial, Attribute, and Text, and it is found that decontextualized generation consistently outp...

Cheng Yang, Chufan Shi, Bo Shui et al. · 5 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.