paper-with-me

홈 › Papers

Weak baselines and reporting biases lead to overoptimism in machine learning for fluid-related partial differential equations

2024-07-09 · Nick McGreivy, Ammar Hakim

One of the most promising applications of machine learning (ML) in computational physics is to accelerate the solution of partial differential equations (PDEs). The key objective of ML-based PDE solvers is to output a sufficiently accurate solution faster than standard numerical methods, which are used as a baseline comparison. We first perform a systematic review of the ML-for-PDE solving literature. Of articles that use ML to solve a fluid-related PDE and claim to outperform a standard numerical method, we determine that 79% (60/76) compare to a weak baseline. Second, we find evidence that reporting biases, especially outcome reporting bias and publication bias, are widespread. We conclude that ML-for-PDE solving research is overoptimistic: weak baselines lead to overly positive results, while reporting biases lead to underreporting of negative results. To a large extent, these issues appear to be caused by factors similar to those of past reproducibility crises: researcher degrees of freedom and a bias towards positive results. We call for bottom-up cultural changes to minimize biased reporting as well as top-down structural reforms intended to reduce perverse incentives for doing so.

📄 PDF Abstract BibTeX arXiv:2407.07218

Code (0)

등록된 구현이 없습니다.

Tasks

Articles

Similar Papers 제목 키워드 기반

Unraveling overoptimism and publication bias in ML-driven science

2024-05-23 · Pouria Saidi, Gautam Dasarathy, Visar Berisha

Machine Learning (ML) is increasingly used across many disciplines with impressive reported results. However, recent studies suggest published performance of ML models are often overoptimistic. Validity concerns are unde…

When the Ground Truth is not True: Modelling Human Biases in Temporal Annotations

2023-02-06 · Taku Yamagata, Emma L. Tonkin, Benjamin Arana Sanchez, Ian Craddock 외

In supervised learning, low quality annotations lead to poorly performing classification and detection models, while also rendering evaluation unreliable. This is particularly apparent on temporal data, where annotation …

Stronger Baselines for Trustable Results in Neural Machine Translation

2017-06-29 · WS 2017 8 · Michael Denkowski, Graham Neubig

Interest in neural machine translation has grown rapidly as its effectiveness has been demonstrated across language and data scenarios. New research regularly introduces architectural and algorithmic improvements that le…

Machine TranslationNMTTranslation

Perspective on Bias in Biomedical AI: Preventing Downstream Healthcare Disparities

2026-04-16 · Michal Rosen-Zvi, Yoav Kan-Tor, Michael Danziger, Agata Ferretti 외 arxiv

Healthcare disparities persist across socioeconomic boundaries, often attributed to unequal access to screening, diagnostics, and therapeutics. However, this perspective highlights that critical biases can emerge much ea…

Canonical Factors for Hybrid Neural Fields

2023-08-29 · ICCV 2023 1 · Brent Yi, Weijia Zeng, Sam Buchanan, Yi Ma

Factored feature volumes offer a simple way to build more compact, efficient, and intepretable neural fields, but also introduce biases that are not necessarily beneficial for real-world data. In this work, we (1) charac…