Skip to content

Category

artificial intelligence

4,653 papers

#artificial intelligence Preprint Aug 2026

A Unifying Perspective on Language Model Representations: From Filler-Role Structure to Mechanistic Interpretability

This work proposes using Tensor Product Representations (TPRs) as a unifying hypothesis, and shows that TPRs can unify several prior interpretability methods: additive analogies, linear probing, sparse autoencoders, and activation patching.

Enshang Zhang, R. Thomas McCoy · 0 citations
#artificial intelligence Preprint Aug 2026

No Detectable Change in Side-Level WER from Prompt-Level Context: A Preregistered Ablation on a Production Oral-History Corpus

Evaluating context mechanisms requires sequence-aligned term-level, insertion, and speaker-label measures alongside aggregate accuracy, and sequence-alignment analysis found a small improvement on complete context-listed phrases, too small to materially change side-level WER, and for Gemini coexisting with worsened unlisted-token error.

Theodore O. Cochran, Stephanie Dodson, Keith Nore · 0 citations
#artificial intelligence Preprint Aug 2026

GreenBench: Benchmarking Energy Efficiency and Carbon Footprint of Open-Source LLM Inference on Apple Silicon

GreenBench, a benchmarking framework that evaluates the energy efficiency, throughput, and carbon footprint of five open-source LLMs across three NLP tasks on an Apple M4 Pro with 48 GB unified memory, is presented.

R. Kannan, Rajendra P. Firke, Shreya Bengle et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Can Large Language Models Identify Meaningful Touchpoints in Conversion Attribution?

This evaluation shows that while LLMs effectively uncover a substantial portion of implicitly-related touchpoints, significant room for improvement remains in their selection performance, and offers a new roadmap for transitioning conversion attribution from mechanical rule-matching to human-aligned semantic reasoning.

Jinqi Wu, Sishuo Chen, Zhangming Chan et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Redesigning and Auditing Deep Research Writing for Faithful Reports

CLAIMPROBE is introduced, a claim-level audit that decomposes DR reports into claims and measures hallucination, misattribution, citation hygiene, and necessary-fact recall against retrieved evidence and proposes CLAIMWRITER, a hierarchical claim-based writer that extracts source facts, maps them to a query-derived outline, and drafts each section from a source-linked claim representation.

Hiroaki Hayashi, P. Venkit, Prafulla Kumar Choubey et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Terminal-Bench-LILT: Multilingual Agentic Coding Benchmark Grounded in Language, Region, and Culture

Evaluation of six frontier models reveals that even the strongest model reaches only 63.1\% pass rate, with many tasks unsolved by any model, highlighting that multilingual coding competence is a distinct and underexplored capability axis.

Yunsu Kim, Kaden Uhlig, Ashwin Purohit et al. · 0 citations
#artificial intelligence Preprint Aug 2026

PromptKWS: A Novel Prompt-Guided Open-Vocabulary Keyword Spotting Framework

The Prompt Phrases Prediction Network (PPN) is introduced, an encoder-decoder architecture designed to effectively extract keyword prompts embeddings and infuse the prompt embedding into the Prompt-guided KWS encoder by utilizing a Prompt-acoustic Multi-head Cross-attention (MHCA).

G. Xu, Cheng-Fei Li, Xian-Liang Wang et al. · 0 citations
#artificial intelligence Preprint Aug 2026

PAUSE: Editable Strategy Artifacts for Long-Form Cultural Story Adaptation

PAUSE (Pause-And-Update Strategy Editing) is an intervention that exposes an editable adaptation strategy as a human control surface for cultural decisions in long-form story adaptation, a structured artifact that can be inspected, edited, and then projected through downstream character, entity, and chapter-localization stages.

Taaha Kazi, Vasu Sharma, M. Saifullah et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Enabling Proactive Spoken Turns via a Generalized Style-Aware Full-Duplex Framework

This work proposes LPS-TC, a Lightweight Proactive Speech Turn Controller for plug-and-play integration, and introduces a two-tier evaluation scheme that assesses both chunk-level timing precision and turn-level interaction quality under realistic streaming constraints.

Tianrui Pan, Qinglin Zhang, Chong Deng et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Intelligent Identification and Repair of Design Defects in BIM via Domain-Specific Large Language Models

This study establishes an end-to-end prototype from raw BIM data input, through defect identification, to repair suggestion generation, and establishes an end-to-end prototype to identify and repair various defects in BIM via domain-specific LLMs.

Jia-Rui Lin, Yunzhen Cai, Xiang Ni et al. · 0 citations

From tech blogs

See all →

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.