paper-with-me

홈 › Papers

UniRank: A Multi-Agent Calibration Pipeline for Estimating University Rankings from Anonymized Bibliometric Signals

2026-02-21 · Pedram Riyazimehr, Seyyed Ehsan Mahmoudi arxiv

We present UniRank, a multi-agent LLM pipeline that estimates university positions across global ranking systems using only publicly available bibliometric data from OpenAlex and Semantic Scholar. The system employs a three-stage architecture: (a) zero-shot estimation from anonymized institutional metrics, (b) per-system tool-augmented calibration against real ranked universities, and (c) final synthesis. Critically, institutions are anonymized -- names, countries, DOIs, paper titles, and collaboration countries are all redacted -- and their actual ranks are hidden from the calibration tools during evaluation, preventing LLM memorization from confounding results. On the Times Higher Education (THE) World University Rankings ($n=352$), the system achieves MAE = 251.5 rank positions, Median AE = 131.5, PNMAE = 12.03%, Spearman $ρ= 0.769$, Kendall $τ= 0.591$, hit rate @50 = 20.7%, hit rate @100 = 39.8%, and a Memorization Index of exactly zero (no exact-match zero-width predictions among all 352 universities). The systematic positive-signed error (+190.1 positions, indicating the system consistently predicts worse ranks than actual) and monotonic performance degradation from elite tier (MAE = 60.5, hit@100 = 90.5%) to tail tier (MAE = 328.2, hit@100 = 20.8%) provide strong evidence that the pipeline performs genuine analytical reasoning rather than recalling memorized rankings. A live demo is available at https://unirank.scinito.ai .

📄 PDF Abstract BibTeX arXiv:2602.18824

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

UniRank: End-to-End Domain-Specific Reranking of Hybrid Text-Image Candidates

2026-02-08 · Yupei Yang, Lin Yang, Wanxi Deng, Lin Qu 외 arxiv

Reranking is a critical component in many information retrieval pipelines. Despite remarkable progress in text-only settings, multimodal reranking remains challenging, particularly when the candidate set contains hybrid …

Reinforcement LearningInformation RetrievalDomain Adaptation

Bayesian Uncertainty Propagation for Agentic RAG Pipelines: A Proof-of-Concept Study on Multi-Hop Question Answering

2026-07-01 · Louis Donaldson, Connor Walker, Koorosh Aslansefat, Yiannis Papadopoulos arxiv

Trustworthy deployment of Agentic Retrieval-Augmented Generation (RAG) systems requires mechanisms for estimating when multi-stage reasoning pipelines may fail. This paper presents an uncertainty-aware Agentic Retrieval-…

Multi-hop Question Answering

UniRank: Unimodal Bandit Algorithm for Online Ranking

2022-08-02 · Camille-Sovanneary Gauthier, Romaric Gaudel, Elisa Fromont

We tackle a new emerging problem, which is finding an optimal monopartite matching in a weighted graph. The semi-bandit version, where a full matching is sampled at each iteration, has been addressed by \cite{ADMA}, crea…

Confidence Calibration and Rationalization for LLMs via Multi-Agent Deliberation

2024-04-14 · Ruixin Yang, Dheeraj Rajagopal, Shirley Anugrah Hayati, Bin Hu 외

Uncertainty estimation is a significant issue for current large language models (LLMs) that are generally poorly calibrated and over-confident, especially with reinforcement learning from human feedback (RLHF). Unlike hu…

Deep reinforcement learning for smart calibration of radio telescopes

2021-02-05 · Sarod Yatawatta, Ian M. Avruch

Modern radio telescopes produce unprecedented amounts of data, which are passed through many processing pipelines before the delivery of scientific results. Hyperparameters of these pipelines need to be tuned by hand to …

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)