paper-with-me

Papers

Threshold Learning for Optimal Decision Making

2016-12-01 · NeurIPS 2016 12 · Nathan F. Lepora

Decision making under uncertainty is commonly modelled as a process of competitive stochastic evidence accumulation to threshold (the drift-diffusion model). However, it is unknown how animals learn these decision thresholds. We examine threshold learning by constructing a reward function that averages over many trials to Wald's cost function that defines decision optimality. These rewards are highly stochastic and hence challenging to optimize, which we address in two ways: first, a simple two-factor reward-modulated learning rule derived from Williams' REINFORCE method for neural networks; and second, Bayesian optimization of the reward function with a Gaussian process. Bayesian optimization converges in fewer trials than REINFORCE but is slower computationally with greater variance. The REINFORCE method is also a better model of acquisition behaviour in animals and a similar learning rule has been proposed for modelling basal ganglia function.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian OptimizationDecision MakingDecision Making Under Uncertainty

Methods 이 논문이 사용한 방법론

REINFORCE REINFORCE is a Monte Carlo variant of a policy gradient algorithm in reinforcement learning. The agent collects samples of an episode using its current policy, and uses it to…

Similar Papers 제목 키워드 기반

Decision-Making under Miscalibration

2022-03-18 · Guy N. Rothblum, Gal Yona

ML-based predictions are used to inform consequential decisions about individuals. How should we use predictions (e.g., risk of heart attack) to inform downstream binary classification decisions (e.g., undergoing a medic…

Binary ClassificationDecision MakingMedical Procedure

Enforcing Group Fairness in Algorithmic Decision Making: Utility Maximization Under Sufficiency

2022-06-05 · Joachim Baumann, Anikó Hannák, Christoph Heitz

Binary decision making classifiers are not fair by default. Fairness requirements are an additional element to the decision making rationale, which is typically driven by maximizing some utility function. In that sense, …

Decision MakingFairness

Threshold-Based Optimal Arm Selection in Monotonic Bandits: Regret Lower Bounds and Algorithms

2025-09-02 · Chanakya Varude, Jay Chaudhary, Siddharth Kaushik, Prasanna Chaporkar arxiv

In multi-armed bandit problems, the typical goal is to identify the arm with the highest reward. This paper explores a threshold-based bandit problem, aiming to select an arm based on its relation to a prescribed thresho…

Recommendation Systems

Bounded-Abstention Pairwise Learning to Rank

2025-05-29 · Antonio Ferrara, Andrea Pugnana, Francesco Bonchi, Salvatore Ruggieri

Ranking systems influence decision-making in high-stakes domains like health, education, and employment, where they can have substantial economic and social impacts. This makes the integration of safety mechanisms essent…

Decision MakingLearning-To-Rank

Strategic Impatience in Go/NoGo versus Forced-Choice Decision-Making

2012-12-01 · NeurIPS 2012 12 · Pradeep Shenoy, Angela J. Yu

Two-alternative forced choice (2AFC) and Go/NoGo (GNG) tasks are behavioral choice paradigms commonly used to study sensory and cognitive processing in choice behavior. While GNG is thought to isolate the sensory/decisio…

Decision Making