Position: More Rigorous Software Engineering Would Improve Reproducibility in Machine Learning Research
Experimental verification and falsification of scholarly work are part of the scientific method's core. To improve the Machine Learning (ML)-communities' ability to verify results from prior work, we argue for more robust software engineering. We estimate the adoption of common engineering best practices by examining repository links from all recently accepted International Conference on Machine Learning (ICML), International Conference on Learning Representations (ICLR) and Neural Information Processing Systems (NeurIPS) papers as well as ICML papers over time. Based on the results, we recommend how we, as a community, can improve reproducibility in ML-research.
Code (1)
Tasks
PositionSimilar Papers 제목 키워드 기반
Imandra CodeLogician: Neuro-Symbolic Reasoning for Precise Analysis of Software Logic
Large Language Models (LLMs) have shown strong performance on code understanding tasks, yet they fundamentally lack the ability to perform precise, exhaustive mathematical reasoning about program behavior. Existing bench…
Mathematical ReasoningMore Is Different: Toward a Theory of Emergence in AI-Native Software Ecosystems
Software engineering faces a fundamental challenge: multi-agent AI systems fail in ways that defy explanation by traditional theories. While individual agents perform correctly, their interactions degrade entire ecosyste…
Machine Learning for Software Engineering: A Systematic Mapping
Context: The software development industry is rapidly adopting machine learning for transitioning modern day software systems towards highly intelligent and self-learning systems. However, the full potential of machine l…
ArticlesBIG-bench Machine LearningSelf-LearningSeamless Digital Engineering: A Grand Challenge Driven by Needs
Digital Engineering currently relies on costly and often bespoke integration of disparate software products to assemble the authoritative source of truth of the system-of-interest. Tools not originally designed to work t…
Agentic AI for Software: thoughts from Software Engineering community
AI agents have recently shown significant promise in software engineering. Much public attention has been transfixed on the topic of code generation from Large Language Models (LLMs) via a prompt. However, software engin…
Code GenerationProgram Repair