paper-with-me

홈 › Papers

Automated Mining of Leaderboards for Empirical AI Research

2021-08-31 · Salomon Kabongo, Jennifer D'Souza, Sören Auer

With the rapid growth of research publications, empowering scientists to keep oversight over the scientific progress is of paramount importance. In this regard, the Leaderboards facet of information organization provides an overview on the state-of-the-art by aggregating empirical results from various studies addressing the same research challenge. Crowdsourcing efforts like PapersWithCode among others are devoted to the construction of Leaderboards predominantly for various subdomains in Artificial Intelligence. Leaderboards provide machine-readable scholarly knowledge that has proven to be directly useful for scientists to keep track of research progress. The construction of Leaderboards could be greatly expedited with automated text mining. This study presents a comprehensive approach for generating Leaderboards for knowledge-graph-based scholarly information organization. Specifically, we investigate the problem of automated Leaderboard construction using state-of-the-art transformer models, viz. Bert, SciBert, and XLNet. Our analysis reveals an optimal approach that significantly outperforms existing baselines for the task with evaluation scores above 90% in F1. This, in turn, offers new state-of-the-art results for Leaderboard extraction. As a result, a vast share of empirical AI research can be organized in the next-generation digital libraries as knowledge graphs.

📄 PDF Abstract BibTeX arXiv:2109.13089

Code (1)

kabongosalomon/task-dataset-metric-nli-extraction 공식 구현 pytorch

Tasks

Knowledge GraphsScientific Results Extraction

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
BPE Byte Pair Encoding, or BPE, is a subword segmentation algorithm that encodes rare and unknown words as sequences of subword units. The intuition is that various word…
Residual Connection 설명 없음
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
SentencePiece 설명 없음
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…

Similar Papers 제목 키워드 기반

Zero-shot Entailment of Leaderboards for Empirical AI Research

2023-03-29 · Salomon Kabongo, Jennifer D'Souza, Sören Auer

We present a large-scale empirical investigation of the zero-shot learning phenomena in a specific recognizing textual entailment (RTE) task category, i.e. the automated mining of leaderboards for Empirical AI Research. …

Natural Language InferenceRTEZero-Shot Learning

ORKG-Leaderboards: A Systematic Workflow for Mining Leaderboards as a Knowledge Graph

2023-05-10 · Salomon Kabongo, Jennifer D'Souza, Sören Auer

The purpose of this work is to describe the Orkg-Leaderboard software designed to extract leaderboards defined as Task-Dataset-Metric tuples automatically from large collections of empirical research papers in Artificial…

Instruction Finetuning for Leaderboard Generation from Empirical AI Research

2024-08-19 · Salomon Kabongo, Jennifer D'Souza

This study demonstrates the application of instruction finetuning of pretrained Large Language Models (LLMs) to automate the generation of AI research leaderboards, extracting (Task, Dataset, Metric, Score) quadruples fr…

ArticlesNatural Language Inference

The Trust Paradox: How CS Researchers Engage LLM Leaderboards

2026-05-27 · Pouya Sadeghi, Anamaria Crisan, Jimmy Lin arxiv

Large language model (LLM) leaderboards rank AI models using standardized benchmarks and have become highly visible across computer science, despite known limitations in their reliability and robustness. Yet how they sha…

A Position Paper on the Automatic Generation of Machine Learning Leaderboards

2025-05-23 · Roelien C Timmer, Yufang Hou, Stephen Wan

An important task in machine learning (ML) research is comparing prior work, which is often performed via ML leaderboards: a tabular overview of experiments with comparable conditions (e.g., same task, dataset, and metri…

BenchmarkingPosition