This work presents a diagnostic and optimization framework grounded in a key empirical finding: the value of tokens within a CoT reasoning sequence is highly non-uniform, and this non-uniformity can be effectively characterized by token-level log probability signals.
Run-Jia Zeng, Hang Hua, Yiyang Liu et al.· 0 citations
The visual modality, i.e., images, plays a key role in multi-modal entity alignment (MMEA). Existing approaches often directly fuse the image with other modalities to align different entities. Although simple, such strategies overlook the potential noise in the images and their semantic misalignment with corresponding...
Chen-Xiao Li, Yun-He Feng, Dongfang Liu et al.· 0 citations
VLA-Scope is introduced, a two-stage framework that combines input-shift characterization with execution history to predict failure during OOD rollouts and achieves a higher ROC-AUC than the evaluated ActProbe and SAFE-MLP baselines.