paper-with-me

Papers

MIRB: Mathematical Information Retrieval Benchmark

2025-05-21 · Haocheng Ju, Bin Dong

Mathematical Information Retrieval (MIR) is the task of retrieving information from mathematical documents and plays a key role in various applications, including theorem search in mathematical libraries, answer retrieval on math forums, and premise selection in automated theorem proving. However, a unified benchmark for evaluating these diverse retrieval tasks has been lacking. In this paper, we introduce MIRB (Mathematical Information Retrieval Benchmark) to assess the MIR capabilities of retrieval models. MIRB includes four tasks: semantic statement retrieval, question-answer retrieval, premise retrieval, and formula retrieval, spanning a total of 12 datasets. We evaluate 13 retrieval models on this benchmark and analyze the challenges inherent to MIR. We hope that MIRB provides a comprehensive framework for evaluating MIR systems and helps advance the development of more effective retrieval models tailored to the mathematical domain.

📄 PDF Abstract BibTeX arXiv:2505.15585

Code (1)

j991222/mirb 공식 구현

Tasks

Automated Theorem ProvingInformation RetrievalMathRetrieval

Similar Papers 제목 키워드 기반

AutoMIR: Effective Zero-Shot Medical Information Retrieval without Relevance Labels

2024-10-26 · Lei LI, Xiangxu Zhang, Xiao Zhou, Zheng Liu

Medical information retrieval (MIR) is essential for retrieving relevant medical knowledge from diverse sources, including electronic health records, scientific literature, and medical databases. However, achieving effec…

BenchmarkingInformation RetrievalRetrievalSelf-Learning

MirBot: A collaborative object recognition system for smartphones using convolutional neural networks

2017-06-09 · Antonio Pertusa, Antonio-Javier Gallego, Marisa Bernabeu

MirBot is a collaborative application for smartphones that allows users to perform object recognition. This app can be used to take a photograph of an object, select the region of interest and obtain the most likely clas…

image-classificationImage ClassificationObjectObject Recognition+1

Benchmarking Multi-Image Understanding in Vision and Language Models: Perception, Knowledge, Reasoning, and Multi-Hop Reasoning

2024-06-18 · Bingchen Zhao, Yongshuo Zong, Letian Zhang, Timothy Hospedales

The advancement of large language models (LLMs) has significantly broadened the scope of applications in natural language processing, with multi-modal LLMs extending these capabilities to integrate and interpret visual d…

BenchmarkingWorld Knowledge

Implementing result-based agri-environmental payments by means of modelling

2019-08-22

From a theoretical point of view, result-based agri-environmental payments are clearly preferable to action-based payments. However, they suffer from two major practical disadvantages: costs of measuring the results and …

Management

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

2026-06-29 · Nikolay Georgiev, Maria Drencheva, Kseniia Ibragimova, Ivo Petrov 외 arxiv

As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theorem libraries, and educational resources. However, choosing the right re…

Information Retrieval