paper-with-me

홈 › Papers

Learning to Infer User Hidden States for Online Sequential Advertising

2020-09-03 · Zhaoqing Peng, Junqi Jin, Lan Luo, Yaodong Yang, Rui Luo, Jun Wang, Wei-Nan Zhang, Haiyang Xu, Miao Xu, Chuan Yu, Tiejian Luo, Han Li, Jian Xu, Kun Gai

To drive purchase in online advertising, it is of the advertiser's great interest to optimize the sequential advertising strategy whose performance and interpretability are both important. The lack of interpretability in existing deep reinforcement learning methods makes it not easy to understand, diagnose and further optimize the strategy. In this paper, we propose our Deep Intents Sequential Advertising (DISA) method to address these issues. The key part of interpretability is to understand a consumer's purchase intent which is, however, unobservable (called hidden states). In this paper, we model this intention as a latent variable and formulate the problem as a Partially Observable Markov Decision Process (POMDP) where the underlying intents are inferred based on the observable behaviors. Large-scale industrial offline and online experiments demonstrate our method's superior performance over several baselines. The inferred hidden states are analyzed, and the results prove the rationality of our inference.

📄 PDF Abstract BibTeX arXiv:2009.01453

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learning

Methods 이 논문이 사용한 방법론

Interpretability 설명 없음

Similar Papers 제목 키워드 기반

Bayesian Non-parametric Hidden Markov Model for Agile Radar Pulse Sequences Streaming Analysis

2023-02-09 · Jiadi Bao, Yunjie Li, Mengtao Zhu, Shafei Wang

Multi-function radars (MFRs) are sophisticated types of sensors with the capabilities of complex agile inter-pulse modulation implementation and dynamic work mode scheduling. The developments in MFRs pose great challenge…

Bayesian InferenceChange DetectionChange Point Detectionparameter estimation+1

A Novel Hybrid Sequential Model for Review-based Rating Prediction

2019-04-14 · Advances in Knowledge Discovery and Data Mining (PAKDD 2019) 2019 4 · Yuanquan Lu, Wei zhang, Pan Lu, Jianyong Wang

Nowadays, the online interactions between users and items become diverse, and may include textual reviews as well as numerical ratings. Reviews often express various opinions and sentiments, which can alleviate the spars…

Multi-Domain Recommender SystemsRecommendation Systems

Detecting User Exits from Online Behavior: A Duration-Dependent Latent State Model

2022-08-08 · Tobias Hatt, Stefan Feuerriegel

In order to steer e-commerce users towards making a purchase, marketers rely upon predictions of when users exit without purchasing. Previously, such predictions were based upon hidden Markov models (HMMs) due to their a…

Neurons as Monte Carlo Samplers: Bayesian Inference and Learning in Spiking Networks

2014-12-01 · NeurIPS 2014 12 · Yanping Huang, Rajesh P. Rao

We propose a two-layer spiking network capable of performing approximate inference and learning for a hidden Markov model. The lower layer sensory neurons detect noisy measurements of hidden world states. The higher laye…

Bayesian Inference

SAGE:State-Aware Guided End-to-End Policy for Multi-Stage Sequential Tasks via Hidden Markov Decision Process

2025-09-24 · BinXu Wu, TengFei Zhang, Chen Yang, JiaHao Wen 외 arxiv

Multi-stage sequential (MSS) robotic manipulation tasks are prevalent and crucial in robotics. They often involve state ambiguity, where visually similar observations correspond to different actions. We present SAGE, a s…

Active Learning