paper-with-me

홈 › Papers

Who Speaks Matters: Authority-Aware Multi-View RAG over Italian Parliamentary Proceedings

2026-08-13 · Mirko Tritella, Riccardo Pozzi, Matteo Palmonari arxiv

Parliamentary proceedings are a primary record of democratic deliberation, yet their volume and fragmentation make multi-perspective access difficult for citizens, journalists, and researchers. Applying Retrieval-Augmented Generation (RAG) to parliamentary transcripts introduces three specific risks: dominance of the most frequent speakers, inability to weight speakers according to topical expertise, and citation misattribution in politically sensitive text. We present ParliamentRAG, a RAG system for the Italian Chamber of Deputies that addresses these risks jointly. Its core contribution is a topic-dependent authority model that estimates each speaker's authority as a function of the current query, combining interpretable components such as profession, education, and previous interventions. Given a user query, the system retrieves relevant speech chunks, identifies topic-relevant experts across parliamentary groups, and generates a summary synthesizing their perspectives, accompanied by supporting quotations. ParliamentRAG is evaluated against Google NotebookLM on 15 policy topics via a two-level protocol combining automated metrics and blind A/B human evaluation by six domain experts. The system achieves higher coverage across political groups (0.97 vs. 0.95), perfect quotation faithfulness (1.00 vs. 0.95), and stronger expert preferences on source-related dimensions, while NotebookLM remains stronger on prose-oriented dimensions.

📄 PDF Abstract BibTeX arXiv:2608.13410

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DrugClaw and DrugAudit: A Primary-Source-Grounded Agent and Authority-Aware Benchmark for Drug-Information Question Answering

2026-05-31 · Qing Wang, Bo Li, Jialu Liang, Daling Shi 외 arxiv

Drug-information question answering is a high-stakes setting where hallucinated facts can mislead clinical decision-making and the provenance of each cited fact matters as much as the fact itself. We present DrugClaw, a …

Question Answering

From Relevance to Authority: Authority-aware Generative Retrieval in Web Search Engines

2026-04-15 · Sunkyung Lee, Jihye Back, Donghyeon Jeon, Soonhwan Kwon 외 arxiv

Generative information retrieval (GenIR) formulates the retrieval process as a text-to-text generation task, leveraging the vast knowledge of large language models. However, existing works primarily optimize for relevanc…

Information RetrievalText Generation

Federated Single-Agent Robotics: Multi-Robot Coordination Without Intra-Robot Multi-Agent Fragmentation

2026-04-13 · Xue Qin, Simin Luan, John See, Cong Yang 외 arxiv

As embodied robots move toward fleet-scale operation, multi-robot coordination is becoming a central systems challenge. Existing approaches often treat this as motivation for increasing internal multi-agent decomposition…

An Extreme Multi-label Text Classification (XMTC) Library Dataset: What if we took "Use of Practical AI in Digital Libraries" seriously?

2026-03-11 · Jennifer D'Souza, Sameer Sadruddin, Maximilian Kähler, Andrea Salfinger 외 arxiv

Subject indexing is vital for discovery but hard to sustain at scale and across languages. We release a large bilingual (English/German) corpus of catalog records annotated with the Integrated Authority File (GND), plus …

Multi-Label Text ClassificationMulti-Label Classification

Bridging What the Model Thinks and How It Speaks: Expressive Speech Generation via Self-Aware Intent-Realization Alignment

2026-04-13 · Kuang Wang, Lai Wei, Ping Lin, Qibing Bai 외 arxiv

Speech Language Models (SLMs) exhibit strong semantic understanding, yet often fail to translate this capacity into expressive acoustic realization, producing speech with flattened prosody and misaligned emotion. We iden…