paper-with-me

홈 › Papers

LLaPipe: LLM-Guided Reinforcement Learning for Automated Data Preparation Pipeline Construction

2025-07-18 · Jing Chang, Chang Liu, Jinbin Huang, Rui Mao, Jianbin Qin arxiv

Automated data preparation is crucial for democratizing machine learning, yet existing reinforcement learning (RL) based approaches suffer from inefficient exploration in the vast space of possible preprocessing pipelines. We present LLaPipe, a novel framework that addresses this exploration bottleneck by integrating Large Language Models (LLMs) as intelligent policy advisors. Unlike traditional methods that rely solely on statistical features and blind trial-and-error, LLaPipe leverages the semantic understanding capabilities of LLMs to provide contextually relevant exploration guidance. Our framework introduces three key innovations: (1) an LLM Policy Advisor that analyzes dataset semantics and pipeline history to suggest promising preprocessing operations, (2) an Experience Distillation mechanism that mines successful patterns from past pipelines and transfers this knowledge to guide future exploration, and (3) an Adaptive Advisor Triggering strategy (Advisor\textsuperscript{+}) that dynamically determines when LLM intervention is most beneficial, balancing exploration effectiveness with computational cost. Through extensive experiments on 18 diverse datasets spanning multiple domains, we demonstrate that LLaPipe achieves up to 22.4\% improvement in pipeline quality and 2.3$\times$ faster convergence compared to state-of-the-art RL-based methods, while maintaining computational efficiency through selective LLM usage (averaging only 19.0\% of total exploration steps).

📄 PDF Abstract BibTeX arXiv:2507.13712

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyReinforcement Learning

Similar Papers 제목 키워드 기반

CollaPipe: Adaptive Segment-Optimized Pipeline Parallelism for Collaborative LLM Training in Heterogeneous Edge Networks

2025-09-24 · Jiewei Chen, Xiumei Deng, Zehui Xiong, Shaoyong Guo 외 arxiv

The increasing demand for intelligent mobile applications has made multi-agent collaboration with Transformer-based large language models (LLMs) essential in mobile edge computing (MEC) networks. However, training LLMs i…

SoftPipe: A Soft-Guided Reinforcement Learning Framework for Automated Data Preparation

2025-07-18 · Jing Chang, Chang Liu, Jinbin Huang, Shuyuan Zheng 외 arxiv

Data preparation is a foundational yet notoriously challenging component of the machine learning lifecycle, characterized by a vast combinatorial search space. While reinforcement learning (RL) offers a promising directi…

Reinforcement LearningBayesian Inference

Crown Jewels Analysis using Reinforcement Learning with Attack Graphs

2021-08-20 · Rohit Gangupantulu, Tyler Cody, Abdul Rahman, Christopher Redino 외

Cyber attacks pose existential threats to nations and enterprises. Current practice favors piece-wise analysis using threat-models in the stead of rigorous cyber terrain analysis and intelligence preparation of the battl…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

A Learned Simulation Environment to Model Student Engagement and Retention in Automated Online Courses

2022-12-22 · N. Imstepf, S. Senn, A. Fortin, B. Russell 외

We developed a simulator to quantify the effect of exercise ordering on both student engagement and retention. Our approach combines the construction of neural network representations for users and exercises using a dyna…

reinforcement-learningReinforcement Learning (RL)

DataAssist: A Machine Learning Approach to Data Cleaning and Preparation

2023-07-14 · Kartikay Goyle, Quin Xie, Vakul Goyle

Current automated machine learning (ML) tools are model-centric, focusing on model selection and parameter optimization. However, the majority of the time in data analysis is devoted to data cleaning and wrangling, for w…

AutoMLModel Selection