Skip to content

Category

artificial intelligence

4,637 papers

Semantic Flow Regularization: Teaching LLMs to Generate Diverse Yet Coherent Responses

Semantic Flow Regularization (SFR), a lightweight auxiliary objective that supervises the backbone with continuous sentence-encoder embeddings of future segments via conditional flow matching, improves output diversity, style fidelity, and response quality over SFT on a large-scale industrial dialogue dataset.

Ke Peng, Feifei Li, Xing Fan et al. · 0 citations

QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents

Evaluating three frontier VLMs in both homogeneous and cross-model adversarial settings, it is found that even the strongest agent hallucinates 15.1% of its verifiable spatial claims and 11.5% of accusations are strictly unsupported.

Ye Yuan, Ruiqi Song, Wei-En Li et al. · 2 citations

DRIP-R: A Benchmark for Decision-Making and Reasoning Under Real-World Policy Ambiguity in the Retail Domain

DRIP-R is introduced, a benchmark that systematically exploits real-world retail policy ambiguities to construct scenarios in which no single correct resolution exists, and shows that frontier models fundamentally disagree on identical policy-ambiguous scenarios, confirming that ambiguity poses a genuine and systematic challenge to LLM decision-making.

Hsuvas Borkakoty, Sebastian Pohl, Cheng Wang et al. · 0 citations

KoALa-Bench: Evaluating Large Audio Language Models on Korean Speech Understanding and Faithfulness

This paper introduces KoALa-Bench, a comprehensive benchmark for evaluating Korean speech understanding and speech faithfulness of LALMs, and incorporates listening questions from the Korean college scholastic ability test as well as content covering Korean cultural domains.

Jinyoung Kim, Hyeongsoo Lim, Eunseo Seo et al. · 0 citations
#artificial intelligence Preprint Apr 2026

Do We Still Need Humans in the Loop? Human vs. LLM Annotation in Active Learning for TikTok Hate Speech Detection

LLM annotation at scale outperforms human-supervised classifiers at roughly one-tenth the cost, for both a closed-source and an open-weight LLM, and the advantage is robust under soft-label evaluation.

Ahmad Dawar Hakimi, Lea Hirlimann, Isabelle Augenstein et al. · 0 citations

Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation

A systematic analysis of expert routing patterns in MoE models reveals Language Routing Isolation, in which high- and low-resource languages tend to activate largely disjoint expert sets, and proposes RISE, a framework that exploits routing isolation to identify and adapt language-specific expert subnetworks.

Kening Zheng, Wei-Chieh Huang, Jiahao Huo et al. · 4 citations · ⚡2
#artificial intelligence Preprint Apr 2026

Where Does Robustness Live? Neuron-Guided Adaptation for Retrieval-Augmented Language Models

NeuRIT is proposed, a Neuron-guided Robust Instruction-Tuning framework built on a localization-first perspective that mines context-aware neurons associated with relevant and irrelevant context processing, and uses them as anchors to selectively adapt both the identified neuron groups and the layers in which they concentrate.

Jae Lee, Jaemin Kim, Sumyeong Ahn et al. · 0 citations

Do Large Language Models Possess a Theory of Mind? A Comparative Evaluation Using the Strange Stories Paradigm

This study tested five Large Language Models and compared their performance to that of human controls using an adapted version of a text-based tool widely used in human ToM research, revealing a performance gap between the models.

Anna Babarczy, András Lukács, Péter Vedres et al. · 1 citation

Small Updates, Big Doubts: Does Parameter-Efficient Fine-tuning Enhance Hallucination Detection ?

Experimental results show that PEFT consistently strengthens hallucination detection ability, substantially improving AUROC across a wide range of hallucination detectors, and indicates that PEFT methods primarily reshapes how uncertainty is encoded and surfaced, comparing with injecting new factual knowledge into the models.

Xuehai Hu, Yifan Zhang, Song-Tao Wei et al. · 2 citations
#artificial intelligence Review Jan 2026

The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models

It is taken that TLMs encode a non-trivial amount of syntactic knowledge, which shows strong performance on formal syntactic phenomena, but weaker and more variable performance on phenomena at the syntax-semantics interface.

Nora Graichen, Iria de-Dios-Flores, Gemma Boleda · 2 citations

Beyond Factual QA: Mentorship-Oriented Question Answering over Long-Form Multilingual Content

This work introduces MentorQA, the first multilingual dataset and evaluation framework for mentorship-focused question answering from long-form videos, and defines mentorship-focused evaluation dimensions that go beyond factual accuracy, capturing clarity, alignment, and learning value.

Parth Bhalerao, D. D’souza, Rui Guan et al. · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.