Skip to content
Preprint

Who Speaks Matters: Authority-Aware Multi-View RAG over Italian Parliamentary Proceedings

Aug 2026 · 0 citations · 28 references
Computer Science

TL;DR

ParliamentRAG is a topic-dependent authority model that estimates each speaker's authority as a function of the current query, combining interpretable components such as profession, education, and previous interventions that addresses risks of dominance of the most frequent speakers, inability to weight speakers according to topical expertise, and citation misattribution in politically sensitive text.

Abstract

Parliamentary proceedings are a primary record of democratic deliberation, yet their volume and fragmentation make multi-perspective access difficult for citizens, journalists, and researchers. Applying Retrieval-Augmented Generation (RAG) to parliamentary transcripts introduces three specific risks: dominance of the most frequent speakers, inability to weight speakers according to topical expertise, and citation misattribution in politically sensitive text. We present ParliamentRAG, a RAG system for the Italian Chamber of Deputies that addresses these risks jointly. Its core contribution is a topic-dependent authority model that estimates each speaker's authority as a function of the current query, combining interpretable components such as profession, education, and previous interventions. Given a user query, the system retrieves relevant speech chunks, identifies topic-relevant experts across parliamentary groups, and generates a summary synthesizing their perspectives, accompanied by supporting quotations. ParliamentRAG is evaluated against Google NotebookLM on 15 policy topics via a two-level protocol combining automated metrics and blind A/B human evaluation by six domain experts. The system achieves higher coverage across political groups (0.97 vs. 0.95), perfect quotation faithfulness (1.00 vs. 0.95), and stronger expert preferences on source-related dimensions, while NotebookLM remains stronger on prose-oriented dimensions.

View source

Similar papers

Preprint Sep 2026

The Wisdom of the Loudest: A Large-Scale Audit of Generative Search on Reddit

Online communities are valued not only for answers, but for the diversity of experiences and perspectives they contain. Generative search increasingly mediates access to this discourse, yet little is known about which community voices survive retrieval and synthesis. We audit Reddit Answers using 10,000 queries from 20...

Agam Goyal, Wang Claire, Eshwar Chandrasekharan · 0 citations
#natural language process... Preprint Sep 2026

From Echo Chambers to Epistemic Monoculture: Large Language Models Present Temporally Contingent Partisan Alignments as Knowledge

Large language models (LLMs) are rapidly becoming an interface between citizens and political information. They are often regarded as"a better Google."While this analogy might work for some instances, it is unintuitively problematic for democratic politics. A search engine retrieves human-authored documents, while a la...

W. Tam · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.