Book
Open access
Aug 2026
Don't Just Encode But See: A Data-Centric Paradigm for Visual Molecular Understanding in Large Language Models
MolGlass is proposed, a data-centric paradigm for visual molecular understanding in vision-language models (VLMs) that injects chemical priors directly into the visual input through chemical-aware visual augmentations, without modifying model architectures or training molecule-specific encoders.
Runqing Xu, Xiaotang Wang, Chunfeng Gao et al.
· Proceedings of the 32nd ACM... · 0 citations