MLDev: Data Science Experiment Automation and Reproducibility Software
In this paper we explore the challenges of automating experiments in data science. We propose an extensible experiment model as a foundation for integration of different open source tools for running research experiments. We implement our approach in a prototype open source MLDev software package and evaluate it in a series of experiments yielding promising results. Comparison with other state-of-the-art tools signifies novelty of our approach.
Code (1)
Similar Papers 제목 키워드 기반
34 Examples of LLM Applications in Materials Science and Chemistry: Towards Automation, Assistants, Agents, and Accelerated Scientific Discovery
Large Language Models (LLMs) are reshaping many aspects of materials science and chemistry research, enabling advances in molecular property prediction, materials design, scientific automation, knowledge extraction, and …
Large Language ModelMolecular Property PredictionProperty Predictionscientific discoveryLarge language models in materials science and the need for open-source approaches
Large language models (LLMs) are rapidly transforming materials science. This review examines recent LLM applications across the materials discovery pipeline, focusing on three key areas: mining scientific literature , p…
R-LAM: Reproducibility-Constrained Large Action Models for Scientific Workflow Automation
Large Action Models (LAMs) extend large language models by enabling autonomous decision-making and tool execution, making them promising for automating scientific workflows. However, scientific workflows impose strict re…
SciOps: Achieving Productivity and Reliability in Data-Intensive Research
Scientists are increasingly leveraging advances in instruments, automation, and collaborative tools to scale up their experiments and research goals, leading to new bursts of discovery. Various scientific disciplines, in…
Experimental DesignAn Analyst-Inspector Framework for Evaluating Reproducibility of LLMs in Data Science
Large Language Models (LLMs) have demonstrated potential for data science tasks via code generation. However, the exploratory nature of data science, alongside the stochastic and opaque outputs of LLMs, raise concerns ab…
BenchmarkingCode Generation