Consequential Ranking Algorithms and Long-term Welfare
Ranking models are typically designed to provide rankings that optimize some measure of immediate utility to the users. As a result, they have been unable to anticipate an increasing number of undesirable long-term consequences of their proposed rankings, from fueling the spread of misinformation and increasing polarization to degrading social discourse. Can we design ranking models that understand the consequences of their proposed rankings and, more importantly, are able to avoid the undesirable ones? In this paper, we first introduce a joint representation of rankings and user dynamics using Markov decision processes. Then, we show that this representation greatly simplifies the construction of consequential ranking models that trade off the immediate utility and the long-term welfare. In particular, we can obtain optimal consequential rankings just by applying weighted sampling on the rankings provided by models that maximize measures of immediate utility. However, in practice, such a strategy may be inefficient and impractical, specially in high dimensional scenarios. To overcome this, we introduce an efficient gradient-based algorithm to learn parameterized consequential ranking models that effectively approximate optimal ones. We showcase our methodology using synthetic and real data gathered from Reddit and show that ranking models derived using our methodology provide ranks that may mitigate the spread of misinformation and improve the civility of online discussions.
Code (0)
등록된 구현이 없습니다.
Tasks
MisinformationSimilar Papers 제목 키워드 기반
Trade-Offs Between Ranking Objectives: Reduced-Form Evidence and Structural Estimation
Online retailers and platforms typically present alternatives using ranked product lists. By adjusting the ranking, these platforms influence consumers' choices and, in turn, conversions, platform revenues, and consumer …
counterfactualFormPositionWhat is the general Welfare? Welfare Economic Perspectives
Researchers do not know what the framers of the United States Constitution intended when they wrote of the general Welfare. Nevertheless, economists can conjecture by specifying social welfare functions that aim to expre…
A Job I Like or a Job I Can Get: Designing Job Recommender Systems Using Field Experiments
Recommendation systems (RSs) are increasingly used to guide job seekers on online platforms, yet the algorithms currently deployed are typically optimized for predictive objectives such as clicks, applications, or hires,…
Recommendation SystemsThe Search for Stability: Learning Dynamics of Strategic Publishers with Initial Documents
We study a game-theoretic information retrieval model in which strategic publishers aim to maximize their chances of being ranked first by the search engine while maintaining the integrity of their original documents. We…
Information RetrievalRetrievalLearning to Bid Long-Term: Multi-Agent Reinforcement Learning with Long-Term and Sparse Reward in Repeated Auction Games
We propose a multi-agent distributed reinforcement learning algorithm that balances between potentially conflicting short-term reward and sparse, delayed long-term reward, and learns with partial information in a dynamic…
Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning (RL)