Skip to content
Preprint

Ontology-Grounded Project Memory for Coding Agents

Aug 2026 · 0 citations · 12 references
Computer Science

TL;DR

MOOSEDev, a system designed to give coding agents structured, ontology-grounded project memory, is introduced and a temporal commit-history bootstrap of the author's own codebase, a pre-registered live trial, and lessons learned are described.

Abstract

Coding agents have become the primary means of generating new code in many software projects, and the resulting velocity of changes makes keeping track of the reasons behind those changes challenging. This paper introduces MOOSEDev, a system designed to give coding agents structured, ontology-grounded project memory. The system captures architectural decisions, lessons, constraints, and rationales in a knowledge graph exposed to agents via a Model Context Protocol (MCP) interface. Records carry lifecycle status, provenance, and supersession links, queryable via MOOSE, a proprietary neurosymbolic engine that treats the symbolic layer as the primary reasoning substrate. We compared MOOSEDev against a production vector-memory tool on a neutral public corpus of 835 typed records. MOOSEDev returned the expected answer set essentially in full (0.98-1.00) on supersession, set-completeness, and negation questions, whereas the baseline's top-k retrieval surfaced between 6% and 27%. Conversely, relevance recall and token cost were largely equivalent between the two systems. We also describe a temporal commit-history bootstrap of our own codebase, a pre-registered live trial, and lessons learned.

View source

Similar papers

Preprint Aug 2026

GrOIL: Graph-Grounded Domain Ontology Induction with Constrained LLM Mediation

A seven-stage graph-grounded pipeline that converts domain documents into a complete, auditable Web Ontology Language (OWL) Terminological Box (TBox) without any unconstrained generation step is presented, demonstrating that the pipeline produces stable, reusable domain representations from large document corpora.

Maruf Ahmed Mridul, A. Talukder, O. Seneviratne · 0 citations
Book Open access Aug 2026

Structure Shapes the Future of DataxLLM Systems: Retrieval, Structuring, and Reasoning

This tutorial presents a unified vision in which structuring serves as the enabling foundation for three pillars of next-generation LLM systems, highlighting how the cooperative interplay between classical KDD techniques and modern LLMs-where KDD defines structural schemas and quality constraints while LLMs execute fle...

Peng-Cheng Jiang, Jiashuo Sun, Wonbin Kweon et al. · 0 citations
Preprint Aug 2026

Towards Researcher Agents for Knowledge-Graph Question Answering

This work presents an agentic text-to-SPARQL system that goes one step beyond static tool-using agents: a researcher agent that, after each round of inference on a validation set, proposes and tests changes to its own prompts, rules, and tool-orchestration code.

Tommaso Soru, Abdulsobur Oyewale · 0 citations
Preprint Aug 2026

SodaMem: Evidence-Grounded Temporal Graph Memory for LLM Agents

SodaMem is presented, an evidence-grounded temporal graph memory that extracts typed FactEvents with mandatory provenance spans, persists mention time, occurrence time, and validity with SUPERSEDES/CONTRADICTS/UPDATES edges under hybrid lexical-dense indexing and answers via a planner-reader loop that gathers citable e...

Fengrong Wan, Chengcan Wu, Ning Lyu · 1 citation
#artificial intelligence Preprint Sep 2026

Surprising Effectiveness of Self-Demonstrations in Enhancing Schema-Ontology Mapping with LLMs

This paper presents a self-demonstration-driven approach that combines a neuro-symbolic task decomposition with a novel mechanism for automatically generating pattern-guided, dependency-aware demonstrations to address the integration challenge of heterogeneous relational databases into a centralized ontology.

Siddhesh Thombre, Manasi S. Patwardhan, Sunita Sarawagi · 0 citations
Open access Aug 2026

Architecting Reliable Knowledge Retrieval Systems Using Large Language Models

A literature-based architectural framework for reliable knowledge retrieval systems that separates external knowledge management from LLM-based reasoning and generation is developed and indicates that reliable LLM deployment should be treated as an end-to-end architectural problem rather than solely a model-performance...

Bharat Kumar Reddy Karumuri · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.