Skip to content
Book Open access

Towards AI-Powered Localization of Software Bugs: From Semantic Understanding to Agentic Cognition

Jul 2026 · SIGSOFT FSE Companion · pp. 45-48 · 0 citations · 35 references
Computer Science

TL;DR

The studies suggest that bug localization can be improved significantly by leveraging program semantics to bridge gaps between bug report and source code, and capturing Intelligent Relevance Feedback through contextual reasoning, and replicating developers' cognitive debugging practices.

Abstract

Software bugs cost billions annually and consume nearly 50% of developers' time. Despite decades of research, automated bug localization remains challenging according to software practitioners. Traditional approaches (e.g., Information Retrieval) rely on surface-level textual matching, while deep learning methods require extensive training data, limiting their effectiveness and applicability. Recent Large Language Models (LLMs) offer unprecedented capabilities in understanding both natural language and source code, yet their potential for bug localization remains underexplored. In this dissertation work, we hypothesize that through program semantics understanding, contextual reasoning, and developer-inspired debugging practices, bug localization systems can better overcome the limitations of existing approaches. We ask three research questions targeting the hypothesis, conduct three studies leveraging different forms of intelligence, and use them to test our hypothesis. Our studies suggest that bug localization can be improved significantly by (1) leveraging program semantics to bridge gaps between bug report and source code, (2) capturing Intelligent Relevance Feedback through contextual reasoning, and (3) replicating developers' cognitive debugging practices.

Read PDF

Similar papers

Open access Sep 2026

DuaLoc: Dual-Encoder Bug Localization with Bug-Report-Conditioned Attention and Contrastive Learning

DuaLoc combines two pre-trained language models: UniXcoder for the semantic understanding of source code and GraphCodeBERT for awareness of data-flow structure and fine-tuned with a contrastive objective that shapes the embedding space around the localization task.

Amany AlBatlaa, M. Abdullah-Al-Wadud · 0 citations
Preprint Aug 2026

DeepRepoQA: Code Repository Question Answering with Deep Agent Exploration

DeepRepoQA is proposed, a novel question answering (QA) framework for repository-level code understanding that builds on an agentic framework where LLM agents find answers through a systematic tree search over the repository structure.

Wei Peng, Yu-Ling Shi, Yingwei Ma et al. · 0 citations
Jul 2026

Semantic Drift in Bug Resolution: How Behavioral Signals Propagate from Reports to Tests and Patches

This analysis covers 2,857 report-test-patch triplets from Defects4J and SWT-Bench using two widely adopted instruction-tuned LLMs from distinct model families and finds that both models exhibit systematic optimism relative to humans and only modest rank agreement, motivating bias-aware evaluation.

Wendkûuni C. Ouédraogo, Yinghua Li, Xueqi Dang et al. · 0 citations
Open access

Evaluation and Distillation of Source Code Generation Tasks by Large Language Models

Two novel contributions are introduced: CodeEval and CodeQual, an open-source execution framework that provides researchers with a ready-to-use evaluation pipeline for evaluating and improving LLMs in software engineering contexts, encompassing both functional correctness assessment and subjective code quality evaluati...

Danny Brahman · 0 citations
#software testing Review Sep 2026

Debugging Functionality-Twisting Translations by LLMs via Differential Testing with Bayesian Prior

This work proposes tHinter, an automated approach that frames translation error localization as a differential testing task, and integrates mixed-factorial user studies, expert validation, and SWOT-based strategic analysis to assess the perceived helpfulness and resilience within the rapidly evolving LLM landscape.

Shengnan Wu, Xin-Yu Sun, Xin Wang et al. · 0 citations
Preprint Aug 2026

ConFL: Explainable Concurrent Fault Localization via Hierarchy-Guided LLM Reasoning

Experiments on real-world concurrent bugs from eight large-scale Java projects show that ConFL significantly outperforms state-of-the-art IR-based and LLM-based baselines, achieving an MRR of 0.503 and a MAP of 0.486.

Shuai Shao, Dingbang Wang, Yiming Zeng et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.