RLStop: A Reinforcement Learning Stopping Method for TAR
We present RLStop, a novel Technology Assisted Review (TAR) stopping rule based on reinforcement learning that helps minimise the number of documents that need to be manually reviewed within TAR applications. RLStop is trained on example rankings using a reward function to identify the optimal point to stop examining documents. Experiments at a range of target recall levels on multiple benchmark datasets (CLEF e-Health, TREC Total Recall, and Reuters RCV1) demonstrated that RLStop substantially reduces the workload required to screen a document collection for relevance. RLStop outperforms a wide range of alternative approaches, achieving performance close to the maximum possible for the task under some circumstances.
Code (1)
Tasks
reinforcement-learningReinforcement LearningTARSimilar Papers 제목 키워드 기반
A Generalised and Adaptable Reinforcement Learning Stopping Method
This paper presents a Technology Assisted Review (TAR) stopping approach based on Reinforcement Learning (RL). Previous such approaches offered limited control over stopping behaviour, such as fixing the target recall an…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)TARRobust Exploratory Stopping under Ambiguity in Reinforcement Learning
We propose and analyze a continuous-time robust reinforcement learning framework for optimal stopping under ambiguity. In this framework, an agent chooses a robust exploratory stopping time motivated by two objectives: r…
Reinforcement LearningContinuous-time Optimal Stopping through Deep Reinforcement Learning
Simulation based solvers for optimal stopping problems must discretize the stopping decision. Under classical dynamic programming, a coarse exercise grid with only a few stopping opportunities can materially undervalue t…
Computational EfficiencyReinforcement LearningExploratory Optimal Stopping: A Singular Control Formulation
This paper explores continuous-time and state-space optimal stopping problems from a reinforcement learning perspective. We begin by formulating the stopping problem using randomized stopping times, where the decision ma…
reinforcement-learningReinforcement LearningDeep Reinforcement Learning for Optimal Stopping with Application in Financial Engineering
Optimal stopping is the problem of deciding the right time at which to take a particular action in a stochastic system, in order to maximize an expected reward. It has many applications in areas such as finance, healthca…
Deep Reinforcement LearningQ-Learningreinforcement-learningReinforcement Learning (RL)