Skip to content
Open access

Intent-Aware Adaptive Trust Retrieval-Augmented Generation (IAT-RAG): A Framework for Trustworthy Document Question Answering Using Open-Source Large Language Models

Jul 2026 · International Journal of Creative and Open Research in Engineering and Management · Vol 02, pp. 1-9 · 0 citations

TL;DR

The paper summarizes the evolution of the conversational AI, Transformer-based LLMs, and RAG architectures, provides an illustrative evaluation protocol and literature-based comparison, and summarizes comparative results which suggest that the IAT-RAG system should outperform all seven baselines on the tasks of citation accuracy and retrieval F1-score, particularly on multi-hop and comparative queries.

Abstract

Large Language Models (LLMs) show impressive natural-language understanding and generation skills but are limited to knowledge that is pre-stored and trained, and can hallucinate when asked about private, domain-specific or newly generated documents. While conventional RAG and its recent variants (Hybrid RAG, Corrective RAG (CRAG), Adaptive-RAG, FLARE, RAPTOR) each make improvements to only a single stage of the retrieval pipeline, they do not effectively condition retrieval on the intent of the query and quantify the trustworthiness of evidence before generation. We introduce the Intent-Aware Adaptive Trust Retrieval-Augmented Generation (IAT-RAG) framework that integrates (i) query-intent classification to guide each query to a suitable retrieval strategy, (ii) an Adaptive Trust Score (ATS) to adjust the retrieval confidence based on the model's intent classification and (iii) an Evidence Quality Score (EQS) to filter the credibility and internal consistency of each retrieved passage before it is used for generation. Conventional RAG passes all top-k retrieved passages to the generator, whereas the passages supplied to the generator are only those that meet Combined Trust threshold from an intent perspective. The paper summarizes the evolution of the conversational AI, Transformer-based LLMs, and RAG architectures, provides an illustrative evaluation protocol and literature-based comparison, and summarizes comparative results which suggest that the IAT-RAG system should outperform all seven baselines on the tasks of citation accuracy and retrieval F1-score, particularly on multi-hop and comparative queries. The current work is a design and protocol stage contribution, whereas the actual implementation is fully open-source (Sentence-Transformer embeddings, a FAISS vector index, and an open-source instruction-tuned LLM), while complete empirical validation on real data, including statistical-significance testing, is identified as the next immediate step. Keywords: Retrieval-Augmented Generation, Large Language Models, Intent Classification, Adaptive Trust Score, Evidence Quality Score, Hallucination Mitigation, Open-Source LLMs.

Read PDF

Similar papers

Open access 2026

A New Metadata-Aware Retrieval-Augmented Generation (RAG) Architecture for Trustworthy Legal Question Answering

Large Language Models (LLMs) offer strong capabilities for Natural Language Processing, yet their inherent uncertainty often produces hallucinations, confident but incorrect statements, which is critical in domains requiring precise knowledge representation. Retrieval-Augmented Generation (RAG) reduces this risk throug...

Alexandra V. Jove-Ticona, Luis J. Duarte-Coaquera, Israel N. Chaparro-Cruz et al. · 0 citations
Preprint Aug 2026

Why RAGs Hallucinate: Penalty-Aware Evaluation of Retrieval-Augmented Generation Systems with Knowledge-Gap Canaries

Volume-based accuracy rewards retrieval-augmented generation (RAG) systems for guessing: a system that answers everything outscores one that declines when its knowledge base cannot support an answer. Building on the confidence-target analysis of Kalai et al. (2025), we present a penalty-aware evaluation framework for d...

Alden Do Rosario, Hussein Younes, Felipe Pires · 0 citations
Book Open access Jul 2026

SCORE-RAG: Self-Correcting Exploration-Exploitation Retrieval for Multi-hop Question Answering

SCORE-RAG reformulates multi-hop RAG as a two-phase adaptive process: exploration for dynamic query understanding, followed by exploitation for precise evidence gathering, which enables adaptive query comprehension, reduces error accumulation via self-verification, and produces interpretable reasoning chains for accura...

Shuran Zhou, Rui Ling, Junan Chen et al. · 0 citations
Review Open access 2026

Bridging Generative AI and External Knowledge: A Review of Retrieval-Augmented Generation (RAG) and Vector Database Integration

The evidence indicates that no single RAG or vector-database configuration dominates across retrieval quality, faithfulness, latency, throughput, storage, cost, and scalability, and the review positions RAG–vector database integration as a joint retrieval-and-systems optimization problem rather than a database-selectio...

Muhammad Fuad Bin Abdullah, Safwan Abd Razak, Noorrezam Yusop et al. · 0 citations
Open access Sep 2026

INTENT-GUIDED RETRIEVAL-AUGMENTED GENERATION WITH LORA-BASED INTENT ROUTING FOR INTELLIGENT CUSTOMER SUPPORT

Automated customer support technologies based on large-scale language models (LLMs) suffer from three classic failure modes: false creation of non-existing policies, inconsistent answers to the same questions, and misleading interpretations of user intentions due to confusing, joking, or vague inquiries. Large-scale la...

Unknown authors · 0 citations
Preprint Aug 2026

When Context Misleads: Intent-Guided Decoding for Robust Retrieval-Augmented Generation

Retrieval-augmented generation (RAG) improves large language models by grounding generation in external evidence, but it also introduces a source trust problem: retrieved context may be useful, irrelevant, or even misleading. Existing RAG systems often apply a fixed trust policy toward retrieved evidence, which can eit...

Haolin Jin, Pengyue Yang, Hua-Min Chen · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.