paper-with-me

홈 › Papers

Bridging the Reproducibility Divide: Open Source Software's Role in Standardizing Healthcare AI

2026-03-02 · John Wu, Zhenbang Wu, Jimeng Sun arxiv

Our analysis of recent AI4H publications reveals that, despite a trend toward utilizing open datasets and sharing modeling code, 74% of AI4H papers still rely on private datasets or do not share their code. This is especially concerning in healthcare applications, where trust is essential. Furthermore, inconsistent and poorly documented data preprocessing pipelines result in variable model performance reports, even for identical tasks and datasets, making it challenging to evaluate the true effectiveness of AI models. Despite the challenges posed by the reproducibility crisis, addressing these issues through open practices offers substantial benefits. For instance, while the reproducibility mandate adds extra effort to research and publication, it significantly enhances the impact of the work. Our analysis shows that papers that used both public datasets and shared code received, on average, 110% more citations than those that do neither--more than doubling the citation count. Given the clear benefits of enhancing reproducibility, it is imperative for the AI4H community to take concrete steps to overcome existing barriers. The community should promote open science practices, establish standardized guidelines for data preprocessing, and develop robust benchmarks. Tackling these challenges through open-source development can improve reproducibility, which is essential for ensuring that AI models are safe, effective, and beneficial for patient care. This approach will help build more trustworthy AI systems that can be integrated into healthcare settings, ultimately contributing to better patient outcomes and advancing the field of medicine.

📄 PDF Abstract BibTeX arXiv:2603.03367

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Reproducibility and FAIR Principles: The Case of a Segment Polarity Network Model

2023-04-18 · Pedro Mendes

The issue of reproducibility of computational models and the related FAIR principles (findable, accessible, interoperable, and reusable) are examined in a specific test case. I analyze a computational model of the segmen…

Bridging the Gap in the Responsible AI Divides

2026-03-15 · Bálint Gyevnár, Atoosa Kasirzadeh arxiv

Tensions between AI Safety (AIS) and AI Ethics (AIE) have increasingly surfaced in AI governance and public debates about AI, leading to what we term the "responsible AI divides". We introduce a model that categorizes fo…

ARVO: Atlas of Reproducible Vulnerabilities for Open-Source Software

2026-06-15 · Xiang Mei, Jordi Del Castillo, Pulkit Singh Singaria, Haoran Xi 외 arxiv

Achieving reproducibility, quantity, and diversity in vulnerability datasets has long been viewed as an inherent three-way trade-off, where improving one dimension often comes at the cost of the others. In practice, repr…

The risk of sub-optimal use of Open Source NLP Software: UKB is inadvertently state-of-the-art in knowledge-based WSD

2018-05-11 · WS 2018 7 · Eneko Agirre, Oier López de Lacalle, Aitor Soroa

UKB is an open source collection of programs for performing, among other tasks, knowledge-based Word Sense Disambiguation (WSD). Since it was released in 2009 it has been often used out-of-the-box in sub-optimal settings…

Word Sense Disambiguation

Recommendations to enhance rigor and reproducibility in biomedical research

2020-01-15 · Jaqueline J. Brito, Jun Li, Jason H. Moore, Casey S. Greene 외

Computational methods have reshaped the landscape of modern biology. While the biomedical community is increasingly dependent on computational tools, the mechanisms ensuring open data, open software, and reproducibility …