paper-with-me

Papers

Quantifying Reproducibility in NLP and ML

2021-09-02 · Anya Belz

Reproducibility has become an intensely debated topic in NLP and ML over recent years, but no commonly accepted way of assessing reproducibility, let alone quantifying it, has so far emerged. The assumption has been that wider scientific reproducibility terminology and definitions are not applicable to NLP/ML, with the result that many different terms and definitions have been proposed, some diametrically opposed. In this paper, we test this assumption, by taking the standard terminology and definitions from metrology and applying them directly to NLP/ML. We find that we are able to straightforwardly derive a practical framework for assessing reproducibility which has the desirable property of yielding a quantified degree of reproducibility that is comparable across different reproduction studies.

📄 PDF Abstract BibTeX arXiv:2109.01211

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Quantifying the Reproducibility of Graph Neural Networks using Multigraph Brain Data

2021-09-06 · Mohammed Amine Gharsallaoui, Islem Rekik

Graph neural networks (GNNs) have witnessed an unprecedented proliferation in tackling several problems in computer vision, computer-aided diagnosis, and related fields. While prior studies have focused on boosting the m…

Prognosis

A Reproducibility Study on Quantifying Language Similarity: The Impact of Missing Values in the URIEL Knowledge Base

2024-05-17 · Hasti Toossi, Guo Qing Huai, Jinyu Liu, Eric Khiu 외

In the pursuit of supporting more languages around the world, tools that characterize properties of languages play a key role in expanding the existing multilingual NLP research. In this study, we focus on a widely used …

Missing ValuesMultilingual NLP

A Step Toward Quantifying Independently Reproducible Machine Learning Research

2019-09-14 · NeurIPS 2019 12 · Edward Raff

What makes a paper independently reproducible? Debates on reproducibility center around intuition or assumptions but lack empirical results. Our field focuses on releasing code, which is important, but is not sufficient …

BIG-bench Machine Learning

Reproducibility Study of “Quantifying Societal Bias Amplification in Image Captioning”

2023-09-26 · NeurIPS 2023 11

Scope of reproducibility - We study the reproducibility of the paper "Quantifying Societal Bias Amplification in Image Captioning" by Hirota et al. In this paper, the authors propose a new metric to measure bias amplific…

The C-index Multiverse

2025-08-20 · Begoña B. Sierra, Colin McLean, Peter S. Hall, Catalina A. Vallejos arxiv

Quantifying out-of-sample discrimination performance for time-to-event outcomes is a fundamental step for model evaluation and selection in the context of predictive modelling. The concordance index, or C-index, is a wid…