paper-with-me

Papers

Towards Trustworthy Machine Learning in Production: An Overview of the Robustness in MLOps Approach

2024-10-28 · Firas Bayram, Bestoun S. Ahmed

Artificial intelligence (AI), and especially its sub-field of Machine Learning (ML), are impacting the daily lives of everyone with their ubiquitous applications. In recent years, AI researchers and practitioners have introduced principles and guidelines to build systems that make reliable and trustworthy decisions. From a practical perspective, conventional ML systems process historical data to extract the features that are consequently used to train ML models that perform the desired task. However, in practice, a fundamental challenge arises when the system needs to be operationalized and deployed to evolve and operate in real-life environments continuously. To address this challenge, Machine Learning Operations (MLOps) have emerged as a potential recipe for standardizing ML solutions in deployment. Although MLOps demonstrated great success in streamlining ML processes, thoroughly defining the specifications of robust MLOps approaches remains of great interest to researchers and practitioners. In this paper, we provide a comprehensive overview of the trustworthiness property of MLOps systems. Specifically, we highlight technical practices to achieve robust MLOps systems. In addition, we survey the existing research approaches that address the robustness aspects of ML systems in production. We also review the tools and software available to build MLOps systems and summarize their support to handle the robustness aspects. Finally, we present the open challenges and propose possible future directions and opportunities within this emerging field. The aim of this paper is to provide researchers and practitioners working on practical AI applications with a comprehensive view to adopt robust ML solutions in production environments.

📄 PDF Abstract BibTeX arXiv:2410.21346

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Machine Learning Operations (MLOps): Overview, Definition, and Architecture

2022-05-04 · Dominik Kreuzberger, Niklas Kühl, Sebastian Hirschl

The final goal of all industrial machine learning (ML) projects is to develop ML products and rapidly bring them into production. However, it is highly challenging to automate and operationalize ML products and thus many…

BIG-bench Machine LearningCultural Vocal Bursts Intensity Prediction

Is Your Training Pipeline Production-Ready? A Case Study in the Healthcare Domain

2025-06-07 · Daniel Lawand, Lucas Quaresma, Roberto Bolgheroni, Alfredo Goldman 외

Deploying a Machine Learning (ML) training pipeline into production requires robust software engineering practices. This differs significantly from experimental workflows. This experience report investigates this challen…

Toward Production-Ready Federated Learning in Healthcare: Privacy, Orchestration, and Governance in MLOps

2026-07-11 · Sakshi Gorkhali, Jonesh Shrestha arxiv

Healthcare organizations often cannot freely centralize patient data because medical records are sensitive, regulated, and institutionally controlled. Federated learning offers a practical alternative by allowing hospita…

Federated Learning

MLOps -- Definitions, Tools and Challenges

2022-01-01 · G. Symeonidis, E. Nerantzis, A. Kazakis, G. A. Papakostas

This paper is an overview of the Machine Learning Operations (MLOps) area. Our aim is to define the operation and the components of such systems by highlighting the current problems and trends. In this context, we presen…

AutoMLBIG-bench Machine Learning

Architecturally Significant MLOps Guidelines for ML Model Integration and Deployment: a Gray Literature Review

2026-06-03 · Faezeh Amou Najafabad, Markus Haug, Keerthiga Rajenthiram, Justus Bogner 외 arxiv

Context. Despite the growing adoption of Machine Learning Operations (MLOps), teams often approach MLOps projects in an ad hoc manner due to the lack of consolidated architectural guidance. The community would benefit fr…