paper-with-me

홈 › Papers

LUMOS: Large User MOdels for User Behavior Prediction

2025-11-28 · Dhruv Nigam, Naman Agarwal, Krishna Murthy, Susmit Saha arxiv

User behavior prediction at scale remains a critical challenge for online B2C platforms. Traditional approaches rely heavily on task-specific models and domain-specific feature engineering. This is time-consuming, computationally expensive, and requires domain expertise and therefore, not scalable. We present LUMOS (Large User MOdel Series), a transformer-based architecture that eliminates task-specific models and manual feature engineering by learning multiple tasks jointly using only raw user activity data. LUMOS introduces a novel cross-attention mechanism that conditions predictions on future known events (e.g., holidays, sales, etc.), enabling the model to predict complex behavior patterns like "how will upcoming holidays affect user engagement?" The architecture also employs multi-modal tokenization, combining user activities, event context, and static user demographic attributes into rich representations processed through specialized embedding pathways. Through extensive experiments on a production dataset spanning 1.7 trillion user activity tokens from 250 million users, we demonstrate that LUMOS achieves superior performance compared to traditional task-specific models. Across 5 tasks with established baselines, we achieve an average improvement of 0.025 in ROC-AUC for binary classification tasks and 4.6\% reduction in MAPE for regression tasks. Online A/B testing validates these improvements translate to measurable business impact with a 3.15\% increase in Daily Active Users.

📄 PDF Abstract BibTeX arXiv:2512.08957

Code (0)

등록된 구현이 없습니다.

Tasks

Binary ClassificationFeature Engineering

Similar Papers 제목 키워드 기반

Lumos: A Library for Diagnosing Metric Regressions in Web-Scale Applications

2020-06-23 · Jamie Pool, Ebrahim Beyrami, Vishak Gopal, Ashkan Aazami 외

Web-scale applications can ship code on a daily to weekly cadence. These applications rely on online metrics to monitor the health of new releases. Regressions in metric values need to be detected and diagnosed as early …

Empowering Contrastive Federated Sequential Recommendation with LLMs

2026-02-10 · Thi Minh Chau Nguyen, Minh Hieu Nguyen, Duc Anh Nguyen, Xuan Huong Tran 외 arxiv

Federated sequential recommendation (FedSeqRec) aims to perform next-item prediction while keeping user data decentralised, yet model quality is frequently constrained by fragmented, noisy, and homogeneous interaction lo…

Sequential RecommendationRepresentation LearningData Augmentation

Lumos: Let there be Language Model System Certification

2025-12-02 · Isha Chaudhary, Vedaant Jain, Prineet Parhar, Kavya Sachdeva 외 arxiv

We introduce the first principled framework, Lumos, for specifying and formally certifying Language Model System (LMS) behaviors. Lumos is an imperative probabilistic programming DSL over graphs, with constructs to gener…

Autonomous Driving

LUMOS: A Semantic Operating-System Layer for Accessibility-Grounded AI Agents

2026-06-29 · Yogeswar Reddy Thota hf

Current operating systems expose interfaces optimized for human users but not for AI agents. Humans benefit from pixels, icons, windows, visual grouping, mouse movement, and keyboard shortcuts; AI agents instead need com…

Lumos: Efficient Performance Modeling and Estimation for Large-scale LLM Training

2025-04-12 · Mingyu Liang, Hiwot Tadese Kassa, Wenyin Fu, Brian Coutinho 외

Training LLMs in distributed environments presents significant challenges due to the complexity of model execution, deployment systems, and the vast space of configurable strategies. Although various optimization techniq…

Efficient Exploration