paper-with-me

Papers

Design Principles for Falsifiable, Replicable and Reproducible Empirical ML Research

2024-05-28 · Daniel Vranješ, Oliver Niggemann

Empirical research plays a fundamental role in the machine learning domain. At the heart of impactful empirical research lies the development of clear research hypotheses, which then shape the design of experiments. The execution of experiments must be carried out with precision to ensure reliable results, followed by statistical analysis to interpret these outcomes. This process is key to either supporting or refuting initial hypotheses. Despite its importance, there is a high variability in research practices across the machine learning community and no uniform understanding of quality criteria for empirical research. To address this gap, we propose a model for the empirical research process, accompanied by guidelines to uphold the validity of empirical research. By embracing these recommendations, greater consistency, enhanced reliability and increased impact can be achieved.

📄 PDF Abstract BibTeX arXiv:2405.18077

Code (1)

danvran/empiricalmachinelearningresearchguide

Methods 이 논문이 사용한 방법론

Uphold 설명 없음

Similar Papers 제목 키워드 기반

Replicable Constrained Bandits

2026-02-16 · Matteo Bollini, Gianmarco Genalti, Francesco Emanuele Stradi, Matteo Castiglioni 외 arxiv

Algorithmic \emph{replicability} has recently been introduced to address the need for reproducible experiments in machine learning. A \emph{replicable online learning} algorithm is one that takes the same sequence of dec…

SCENEREPLICA: Benchmarking Real-World Robot Manipulation by Creating Replicable Scenes

2023-06-27 · Ninad Khargonkar, Sai Haneesh Allu, Yangxiao Lu, Jishnu Jaykumar P 외

We present a new reproducible benchmark for evaluating robot manipulation in the real world, specifically focusing on pick-and-place. Our benchmark uses the YCB objects, a commonly used dataset in the robotics community,…

BenchmarkingMotion PlanningRobotic GraspingRobot Manipulation

From Generative to Episodic: Sample-Efficient Replicable Reinforcement Learning

2025-07-16 · Max Hopkins, Sihan Liu, Christopher Ye, Yuichi Yoshida arxiv

The epidemic failure of replicability across empirical science and machine learning has recently motivated the formal study of replicable learning algorithms [Impagliazzo et al. (2022)]. In batch settings where data come…

Reinforcement Learning

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences

2026-02-11 · Bang Nguyen, Dominik Soós, Qian Ma, Rochana R. Obadage 외 arxiv

The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on the computational aspect of this task, testing agents' ability to repro…

LUMINA: Foundation Models for Topology Transferable ACOPF

2026-03-04 · Yijiang Li, Zeeshan Memon, Hongwei Jin, Stefano Fenu 외 arxiv

Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constrained scientific systems, where predictions must satisfy physical laws an…