paper-with-me

Papers

Learning from Delayed Outcomes via Proxies with Applications to Recommender Systems

2018-07-24 · Timothy A. Mann, Sven Gowal, András György, Ray Jiang, Huiyi Hu, Balaji Lakshminarayanan, Prav Srinivasan

Predicting delayed outcomes is an important problem in recommender systems (e.g., if customers will finish reading an ebook). We formalize the problem as an adversarial, delayed online learning problem and consider how a proxy for the delayed outcome (e.g., if customers read a third of the book in 24 hours) can help minimize regret, even though the proxy is not available when making a prediction. Motivated by our regret analysis, we propose two neural network architectures: Factored Forecaster (FF) which is ideal if the proxy is informative of the outcome in hindsight, and Residual Factored Forecaster (RFF) that is robust to a non-informative proxy. Experiments on two real-world datasets for predicting human behavior show that RFF outperforms both FF and a direct forecaster that does not make use of the proxy. Our results suggest that exploiting proxies by factorization is a promising way to mitigate the impact of long delays in human-behavior prediction tasks.

📄 PDF Abstract BibTeX arXiv:1807.09387

Code (0)

등록된 구현이 없습니다.

Tasks

Recommendation Systems

Similar Papers 제목 키워드 기반

Impatient Bandits: Optimizing for the Long-Term Without Delay

2025-01-14 · Kelly W. Zhang, Thomas Baldwin-McDonald, Kamil Ciosek, Lucas Maystre 외

Increasingly, recommender systems are tasked with improving users' long-term satisfaction. In this context, we study a content exploration task, which we formalize as a bandit problem with delayed rewards. There is an ap…

Recommendation Systems

Impatient Bandits: Optimizing Recommendations for the Long-Term Without Delay

2023-07-19 · Thomas M. McDonald, Lucas Maystre, Mounia Lalmas, Daniel Russo 외

Recommender systems are a ubiquitous feature of online platforms. Increasingly, they are explicitly tasked with increasing users' long-term satisfaction. In this context, we study a content exploration task, which we for…

Recommendation Systems

Generalized Delayed Feedback Model with Post-Click Information in Recommender Systems

2022-06-01 · Jia-Qi Yang, De-Chuan Zhan

Predicting conversion rate (e.g., the probability that a user will purchase an item) is a fundamental problem in machine learning based recommender systems. However, accurate conversion labels are revealed after a long d…

Recommendation Systems

Biased Error Attribution in Multi-Agent Human-AI Systems Under Delayed Feedback

2026-03-24 · Teerthaa Parakh, Karen M. Feigh arxiv

Human decision-making is strongly influenced by cognitive biases, particularly under conditions of uncertainty and risk. While prior work has examined bias in single-step decisions with immediate outcomes and in human in…

Games That Teach, Chats That Convince: Comparing Interactive and Static Formats for Persuasive Learning

2026-02-20 · Seyed Hossein Alavi, Zining Wang, Shruthi Chockkalingam, Raymond T. Ng 외 arxiv

Interactive systems such as chatbots and games are increasingly used to persuade and educate on sustainability-related topics, yet it remains unclear how different delivery formats shape learning and persuasive outcomes …