It is demonstrated that CoT encodes recoverable, token-level problem-solving information, offering new insight into how reasoning is represented and where it breaks down, suggesting complete reasoning chains are not always necessary.
Houman Mehrafarin, Amit Parekh, Ioannis Konstas· arXiv.org· 2 citations
This work derives a closed-form expression for this adversarial perturbation, bypassing the iterative inner optimization of adversarial training entirely and enabling linear-time evaluation in the state dimension, and shows that this expression approximates the exact minimizer of the value function over the modeled uncertainty set with second-order accuracy.
Alex Zongo, Filippos Fotiadis, U. Topcu et al.· arXiv.org· 1 citation
PeopleSearchBench, an open-source benchmark comprising 119 multilingual queries across four scenarios: corporate recruiting, B2B sales prospecting, expert search, and influencer discovery, finds that multi-source search agents significantly outperform single-domain systems, particularly in influencer discovery where the performance gap is largest.
Tianyu Shi, Wei Wang, Zequn Xie et al.· 0 citations
This paper replicates and extends the system used in the AuTexTification shared task for authorship attribution of machine-generated texts, and tested newer multilingual language models and added 26 document-level stylometric features, using ablation, permutation importance, and SHAP analysis to assess feature influence.
Adam Skurla, D. Macko, Jakub Simko· arXiv.org· 0 citations
Reach audiences
Advertise in front of researchers, engineers, and readers.
This work proposes a multi-stage alignment method that teaches models to recall and apply relevant business policies during chain-of-thought reasoning at inference time, without including the full business policy in-context.
Shubhashis Roy Dipta, Daniel Bis, Kun Zhou et al.· arXiv.org· 6 citations
This work proposes DesignAsCode, a novel framework that reimagines graphic design as a programmatic synthesis task using HTML/CSS, incorporating a Plan-Implement-Reflect pipeline, incorporating a Semantic Planner to construct dynamic, variable-depth element hierarchies and a Visual-Aware Reflection mechanism that optimizes the code to rectify rendering artifacts.
LSTR (Latent Sparse Transcoder Reasoning), a framework that turns sparse transcoders from post-hoc diagnostic tools into in-loop, intervenable transition components for latent reasoning, and suggests that sparse latent transitions can preserve the compression benefits of latent reasoning while making the resulting trajectories more inspectable and intervenable.
Yadong Wang, Hao-Dong Chen, Yu Tian et al.· 0 citations
SPADE (Soil moisture Pattern and Anomaly DEtection), which is the first LLM-based framework specifically developed for soil moisture time-series analysis, is proposed, which is the first LLM-based framework specifically developed for soil moisture time-series analysis.
Yeonju Lee, Rui-Qi Chen, Joseph Oboamah et al.· arXiv.org· 0 citations
Rank-One Safety Injection (ROSI), a white-box method that amplifies a model's safety alignment by permanently steering its activations toward the refusal-mediating subspace, is proposed, suggesting that targeted, interpretable weight steering is a cheap and potent mechanism to improve LLM safety, complementing more resource-intensive fine-tuning paradigms.
H. Shairah, Hasan Abed Al Kader Hammoud, G. Turkiyyah et al.· arXiv.org· 7 citations
Language-Guided Tuning is introduced, a framework that employs multi-agent Large Language Models to automatically optimize configurations through natural language reasoning to demonstrate substantial improvements over traditional optimization methods while maintaining high interpretability.
Yuxing Lu, Yucheng Hu, Nan Sun et al.· 0 citations
RNop represents a shift in mRNA optimization methodology: by infusing explicit and interpretable knowledge, the"black-box"mRNA design can be transformed into a predictable, explainable engineering problem.
This work develops a new SpFT framework, based on ideas from neural network pruning, that improves SpFT's memory efficiency by 20-50\% while matching the accuracy of state-of-the-art methods like LoRA's variants.