Skip to content

Category

artificial intelligence

6,499 papers

Var-JEPA: A Variational Formulation of the Joint-Embedding Predictive Architecture - Bridging Predictive and Generative Self-Supervised Learning

The Variational JEPA (Var-JEPA), which makes the latent generative structure explicit by optimizing a single Evidence Lower Bound (ELBO) and yields meaningful representations without ad-hoc anti-collapse regularizers and allows principled uncertainty quantification in the latent space.

Moritz Gögl, Christopher Yau · 3 citations · ⚡1

The Autonomy Tax: Defense Training Breaks LLM Agents

These findings demonstrate that current defense paradigms optimize for single-turn refusal benchmarks while rendering multi-step agents fundamentally unreliable, necessitating new approaches that preserve tool execution competence under adversarial conditions.

Li Li, Yue Zhao · 8 citations

InfoMamba: An Attention-Free Hybrid Mamba-Transformer Model

A consistency boundary analysis is presented that characterizes when diagonal short-memory SSMs can approximate causal attention and identifies structural gaps that remain and proposes InfoMamba, an attention-free hybrid architecture that consistently outperforms strong Transformer and SSM baselines.

Youjin Wang, Jiaqi Zhao, Rong Fu et al. · 0 citations

Large Reasoning Models Struggle to Transfer Parametric Knowledge Across Scripts

There is potential to improve cross-lingual parametric knowledge transfer during post-training by providing the LLMs with the key entities of the questions in their source language and finding that this disproportionately improves cross-script questions.

Lucas Bandarkar, Alan Ansell, Trevor Cohn · 3 citations
#artificial intelligence Preprint Feb 2026

From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves

The results show that improving IF in LRMs can significantly enhance privacy, suggesting a promising direction for future privacy-aware LRMs, and introduces an SFT dataset that teaches models to follow general instructions throughout their reasoning process.

Haritz Puerto, Haonan Li, Xudong Han et al. · 0 citations

FENCE: A Financial and Multimodal Jailbreak Detection Dataset

FENCE, a bilingual (Korean-English) multimodal dataset for training and evaluating jailbreak detectors in financial applications, provides a focused resource for advancing multimodal jailbreak detection in finance and for supporting safer, more reliable AI systems in sensitive domains.

Mirae Kim, Seonghun Jeong, Youngjun Kwak · 0 citations
#artificial intelligence Preprint Feb 2026

ASA: Backbone-Training-Free Representation Engineering for Tool-Calling Agents

Activation Steering Adapter (ASA), a training-free, inference-time controller that performs a single-shot mid-layer intervention and targets tool domains via a router-conditioned mixture of steering vectors with a probe-guided signed gate to amplify true intent while suppressing spurious triggers is proposed.

Youjin Wang, Run Zhou, Rong Fu et al. · 4 citations · ⚡2

CoFrGeNet: Continued Fraction Architectures for Language Generation

The architecture family implementing this function class is named CoFrGeNets - Continued Fraction Generative Networks, and novel architectural components based on this function class that can replace Multi-head Attention and Feed-Forward Networks in Transformer blocks while requiring much fewer parameters are designed.

Amit Dhurandhar, Vijil Chenthamarakshan, Dennis Wei et al. · 0 citations

Aligning Agentic World Models via Knowledgeable Experience Learning

WorldMind is introduced, a framework that autonomously constructs a symbolic World Knowledge Repository by synthesizing environmental feedback that unifies Process Experience to enforce physical feasibility via prediction errors and Goal Experience to guide task optimality through successful trajectories.

Baochang Ren, Yunzhi Yao, Rui Sun et al. · 3 citations · ⚡1

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.