paper-with-me

홈 › Papers

HALO: Hindsight-Augmented Learning for Online Auto-Bidding

2025-08-05 · Pusen Dong, Chenglong Cao, Xinyu Zhou, Jirong You, Linhe Xu, Feifan Xu, Shuo Yuan arxiv

Digital advertising platforms operate millisecond-level auctions through Real-Time Bidding (RTB) systems, where advertisers compete for ad impressions through algorithmic bids. This dynamic mechanism enables precise audience targeting but introduces profound operational complexity due to advertiser heterogeneity: budgets and ROI targets span orders of magnitude across advertisers, from individual merchants to multinational brands. This diversity creates a demanding adaptation landscape for Multi-Constraint Bidding (MCB). Traditional auto-bidding solutions fail in this environment due to two critical flaws: 1) severe sample inefficiency, where failed explorations under specific constraints yield no transferable knowledge for new budget-ROI combinations, and 2) limited generalization under constraint shifts, as they ignore physical relationships between constraints and bidding coefficients. To address this, we propose HALO: Hindsight-Augmented Learning for Online Auto-Bidding. HALO introduces a theoretically grounded hindsight mechanism that repurposes all explorations into training data for arbitrary constraint configuration via trajectory reorientation. Further, it employs B-spline functional representation, enabling continuous, derivative-aware bid mapping across constraint spaces. HALO ensures robust adaptation even when budget/ROI requirements differ drastically from training scenarios. Industrial dataset evaluations demonstrate the superiority of HALO in handling multi-scale constraints, reducing constraint violations while improving GMV.

📄 PDF Abstract BibTeX arXiv:2508.03267

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DRIVE: Distributional and Retrieval-Augmented Bidding with Value Evaluation

2026-06-12 · Miduo Cui, Haochen Wang, Shangqin Mao, Xun Yang 외 arxiv

Auto-bidding is a core component of real-time advertising systems, where decisions must optimize long-term performance under budget and cost constraints, while online exploration is prohibitively risky. Offline reinforce…

Reinforcement LearningDecision Making

Learning-Augmented Online Bidding in Stochastic Settings

2025-10-29 · Spyros Angelopoulos, Bertrand Simon arxiv

Online bidding is a classic optimization problem, with several applications in online decision-making, the design of interruptible systems, and the analysis of approximation algorithms. In this work, we study online bidd…

HOB: A Holistically Optimized Bidding Strategy under Heterogeneous Bidding Environments

2025-10-17 · Qi Li, Wendong Huang, Qichen Ye, Wutong Xu 외 arxiv

Optimizing a single advertising campaign across heterogeneous channels is a central challenge in industrial autobidding. Auction mechanisms vary across channels in ranking rules (pure eCPM vs. UE-augmented scoring), pric…

AIGB: Generative Auto-bidding via Conditional Diffusion Modeling

2024-05-25 · Jiayan Guo, Yusen Huo, Zhilin Zhang, Tianyu Wang 외

Auto-bidding plays a crucial role in facilitating online advertising by automatically providing bids for advertisers. Reinforcement learning (RL) has gained popularity for auto-bidding. However, most current RL auto-bidd…

Reinforcement Learning (RL)

Sustainable Online Reinforcement Learning for Auto-bidding

2022-10-13 · Zhiyu Mou, Yusen Huo, Rongquan Bai, Mingzhou Xie 외

Recently, auto-bidding technique has become an essential tool to increase the revenue of advertisers. Facing the complex and ever-changing bidding environments in the real-world advertising system (RAS), state-of-the-art…

Q-Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)