paper-with-me

홈 › Papers

Improve Large Language Model Systems with User Logs

2026-02-06 · Changyue Wang, Weihang Su, Qingyao Ai, Xingzhao Yue, Rui Zhang, Xiaojia Chang, Yiqun Liu arxiv

Scaling training data and model parameters has long driven progress in large language models (LLMs), but this paradigm is increasingly constrained by the scarcity of high-quality data and diminishing returns from rising computational costs. As a result, recent work is increasing the focus on continual learning from real-world deployment, where user interaction logs provide a rich source of authentic human feedback and procedural knowledge. However, learning from user logs is challenging due to their unstructured and noisy nature. Vanilla LLM systems often struggle to distinguish useful feedback signals from noisy user behavior, and the disparity between user log collection and model optimization (e.g., the off-policy optimization problem) further strengthens the problem. To this end, we propose UNO (User log-driveN Optimization), a unified framework for improving LLM systems (LLMsys) with user logs. UNO first distills logs into semi-structured rules and preference pairs, then employs query-and-feedback-driven clustering to manage data heterogeneity, and finally quantifies the cognitive gap between the model's prior knowledge and the log data. This assessment guides the LLMsys to adaptively filter out noisy feedback and construct different modules for primary and reflective experiences extracted from user logs, thereby improving future responses. Extensive experiments show that UNO achieves state-of-the-art effectiveness and efficiency, significantly outperforming Retrieval Augmented Generation (RAG) and memory-based baselines. We have open-sourced our code at https://github.com/bebr2/UNO .

📄 PDF Abstract BibTeX arXiv:2602.06470

Code (0)

등록된 구현이 없습니다.

Tasks

Continual Learning

Similar Papers 제목 키워드 기반

Contextual Data Augmentation for Task-Oriented Dialog Systems

2023-10-16 · Dustin Axman, Avik Ray, Shubham Garg, Jing Huang

Collection of annotated dialogs for training task-oriented dialog systems have been one of the key bottlenecks in improving current models. While dialog response generation has been widely studied on the agent side, it i…

Data AugmentationLanguage ModelingLanguage ModellingParaphrase Generation+2

A Prompt Log Analysis of Text-to-Image Generation Systems

2023-03-08 · Yutong Xie, Zhaoying Pan, Jinge Ma, Luo Jie 외

Recent developments in large language models (LLM) and generative AI have unleashed the astonishing capabilities of text-to-image generation systems to synthesize high-quality images that are faithful to a given referenc…

Image GenerationText to Image GenerationText-to-Image Generation

From Logs to Language: Learning Optimal Verbalization for LLM-Based Recommendation at Industry Scale

2026-02-24 · Yucheng Shi, Ying Li, Yu Wang, Yesu Feng 외 arxiv

Large language models (LLMs) are promising backbones for generative recommender systems, yet a key challenge remains underexplored: verbalization, i.e., converting structured user interaction logs into effective natural …

Reinforcement Learning

Face It Yourselves: An LLM-Based Two-Stage Strategy to Localize Configuration Errors via Logs

2024-03-31 · Shiwen Shan, Yintong Huo, Yuxin Su, Yichen Li 외

Configurable software systems are prone to configuration errors, resulting in significant losses to companies. However, diagnosing these errors is challenging due to the vast and complex configuration space. These errors…

Dialog Simulation with Realistic Variations for Training Goal-Oriented Conversational Systems

2020-11-16 · Chien-Wei Lin, Vincent Auvray, Daniel Elkind, Arijit Biswas 외

Goal-oriented dialog systems enable users to complete specific goals like requesting information about a movie or booking a ticket. Typically the dialog system pipeline contains multiple ML models, including natural lang…

Goal-Oriented DialogNatural Language Understanding