paper-with-me

Papers

Topic Is Not Agenda: A Citation-Community Audit of Text Embeddings

2026-05-08 · Junseon Yoo arxiv

Vector search and retrieval-augmented generation (RAG) rest on the assumption that cosine similarity between text embeddings reflects conceptual relatedness. We measure where this assumption breaks. We build an augmented citation graph over 3.58M scientific papers and partition it via Leiden CPM at two granularities: sub-field (L1) and research-agenda (L2, hierarchical inside each L1). Four state-of-the-art embeddings (Gemini, Qwen3-8B, Qwen3-0.6B, SPECTER2) clear the L1 bar reasonably (45-52% top-10 same-rate) but stop working at L2: only 15-21% of top-10 neighbors share the query's research agenda. In absolute terms, 8 of every 10 retrieved papers are off-agenda. The failure is universal across eight scientific domains and all four models; SPECTER2, despite its citation-based contrastive training, is the weakest. As a diagnostic probe, we test whether the same augmented graph also functions as a retrieval signal: a deliberately simple citation-count rerank reaches 57.7% top-1 L2 on top of LLM-expanded Boolean retrieval and 59.6% on top of plain BM25, on 80 curated agenda queries -- about 9 points above the best cosine retriever (Gemini, 50.6%) and 20 points above BM25 alone (39.3%). The probe isolates a slice of the agenda-matching signal the graph carries but the embeddings miss, connecting recent theoretical limits on single-vector retrieval to a concrete failure mode of scientific RAG.

📄 PDF Abstract BibTeX arXiv:2605.07158

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Toward Culturally Grounded Natural Language Processing

2026-03-27 · Sina Bagheri Nezhad arxiv

Multilingual NLP is often treated as a route to global inclusion, but linguistic coverage and cultural competence frequently diverge. This paper synthesizes over 50 papers spanning multilingual performance inequality, cr…

Cross-Lingual Transfer

FrameASt: A Framework for Second-level Agenda Setting in Parliamentary Debates through the Lense of Comparative Agenda Topics

2022-06-01 · ParlaCLARIN (LREC) 2022 6 · Christopher Klamm, Ines Rehbein, Simone Paolo Ponzetto

This paper presents a framework for studying second-level political agenda setting in parliamentary debates, based on the selection of policy topics used by political actors to discuss a specific issue on the parliamenta…

Topic Classification

Basic Research, Lethal Effects: Military AI Research Funding as Enlistment

2024-11-26 · David Gray Widder, Sireesh Gururaja, Lucy Suchman

In the context of unprecedented U.S. Department of Defense (DoD) budgets, this paper examines the recent history of DoD funding for academic research in algorithmically based warfighting. We draw from a corpus of DoD gra…

DeepTRACE: Auditing Deep Research AI Systems for Tracking Reliability Across Citations and Evidence

2025-09-02 · Pranav Narayanan Venkit, Philippe Laban, Yilun Zhou, Kung-Hsiang Huang 외 arxiv

Generative search engines and deep research LLM agents promise trustworthy, source-grounded synthesis, yet users regularly encounter overconfidence, weak sourcing, and confusing citation practices. We introduce DeepTRACE…

Flexible categorization for auditing using formal concept analysis and Dempster-Shafer theory

2022-10-31 · Marcel Boersma, Krishna Manoorkar, Alessandra Palmigiano, Mattia Panettiere 외

Categorization of business processes is an important part of auditing. Large amounts of transnational data in auditing can be represented as transactions between financial accounts using weighted bipartite graphs. We vie…