paper-with-me

Papers

Model-in-the-Loop (MILO): Accelerating Multimodal AI Data Annotation with LLMs

2024-09-16 · Yifan Wang, David Stevens, Pranay Shah, WenWen Jiang, Miao Liu, Xu Chen, Robert Kuo, Na Li, Boying Gong, Daniel Lee, Jiabo Hu, Ning Zhang, Bob Kamma

The growing demand for AI training data has transformed data annotation into a global industry, but traditional approaches relying on human annotators are often time-consuming, labor-intensive, and prone to inconsistent quality. We propose the Model-in-the-Loop (MILO) framework, which integrates AI/ML models into the annotation process. Our research introduces a collaborative paradigm that leverages the strengths of both professional human annotators and large language models (LLMs). By employing LLMs as pre-annotation and real-time assistants, and judges on annotator responses, MILO enables effective interaction patterns between human annotators and LLMs. Three empirical studies on multimodal data annotation demonstrate MILO's efficacy in reducing handling time, improving data quality, and enhancing annotator experiences. We also introduce quality rubrics for flexible evaluation and fine-grained feedback on open-ended annotations. The MILO framework has implications for accelerating AI/ML development, reducing reliance on human annotation alone, and promoting better alignment between human and machine values.

📄 PDF Abstract BibTeX arXiv:2409.10702

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Xiaomi MiMo-VL-Miloco Technical Report

2025-12-19 · Jiaze Li, Jingyang Chen, Yuxun Qu, Shijie Xu 외 arxiv

We open-source MiMo-VL-Miloco-7B and its quantized variant MiMo-VL-Miloco-7B-GGUF, a pair of home-centric vision-language models that achieve strong performance on both home-scenario understanding and general multimodal …

Reinforcement LearningMultimodal ReasoningGesture Recognition

GeneAnnotator: A Semi-automatic Annotation Tool for Visual Scene Graph

2021-09-06 · Zhixuan Zhang, Chi Zhang, Zhenning Niu, Le Wang 외

In this manuscript, we introduce a semi-automatic scene graph annotation tool for images, the GeneAnnotator. This software allows human annotators to describe the existing relationships between participators in the visua…

Graph GenerationGraph LearningImage CaptioningScene Graph Generation+1

Semi-Automated Data Annotation in Multisensor Datasets for Autonomous Vehicle Testing

2025-12-31 · Andrii Gamalii, Daniel Górniak, Robert Nowak, Bartłomiej Olber 외 arxiv

This report presents the design and implementation of a semi-automated data annotation pipeline developed within the DARTS project, whose goal is to create a large-scale, multimodal dataset of driving scenarios recorded …

3D Object DetectionDomain Adaptation

MiLo: Efficient Quantized MoE Inference with Mixture of Low-Rank Compensators

2025-04-03 · Beichen Huang, Yueming Yuan, Zelei Shao, Minjia Zhang

A critical approach for efficiently deploying Mixture-of-Experts (MoE) models with massive parameters is quantization. However, state-of-the-art MoE models suffer from non-negligible accuracy loss with extreme quantizati…

Mixture-of-ExpertsQuantization

MILO: A Lightweight Perceptual Quality Metric for Image and Latent-Space Optimization

2025-09-01 · Uğur Çoğalan, Mojtaba Bemana, Karol Myszkowski, Hans-Peter Seidel 외 arxiv

We present MILO (Metric for Image- and Latent-space Optimization), a lightweight, multiscale, perceptual metric for full-reference image quality assessment (FR-IQA). MILO is trained using pseudo-MOS (Mean Opinion Score) …

Image Quality Assessment