paper-with-me

Papers

Learning Not to Spoof

2023-06-09 · David Byrd

As intelligent trading agents based on reinforcement learning (RL) gain prevalence, it becomes more important to ensure that RL agents obey laws, regulations, and human behavioral expectations. There is substantial literature concerning the aversion of obvious catastrophes like crashing a helicopter or bankrupting a trading account, but little around the avoidance of subtle non-normative behavior for which there are examples, but no programmable definition. Such behavior may violate legal or regulatory, rather than physical or monetary, constraints. In this article, I consider a series of experiments in which an intelligent stock trading agent maximizes profit but may also inadvertently learn to spoof the market in which it participates. I first inject a hand-coded spoofing agent to a multi-agent market simulation and learn to recognize spoofing activity sequences. Then I replace the hand-coded spoofing trader with a simple profit-maximizing RL agent and observe that it independently discovers spoofing as the optimal strategy. Finally, I introduce a method to incorporate the recognizer as normative guide, shaping the agent's perceived rewards and altering its selected actions. The agent remains profitable while avoiding spoofing behaviors that would result in even higher profit. After presenting the empirical results, I conclude with some recommendations. The method should generalize to the reduction of any unwanted behavior for which a recognizer can be learned.

📄 PDF Abstract BibTeX arXiv:2306.06087

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning (RL)

Similar Papers 제목 키워드 기반

ASVspoof2019 vs. ASVspoof5: Assessment and Comparison

2025-05-21 · Avishai Weizman, Yehuda Ben-Shimol, Itshak Lapidot

ASVspoof challenges are designed to advance the understanding of spoofing speech attacks and encourage the development of robust countermeasure systems. These challenges provide a standardized database for assessing and …

Speaker VerificationVoice Anti-spoofing

On Disentangling Spoof Trace for Generic Face Anti-Spoofing

2020-07-17 · ECCV 2020 8 · Yaojie Liu, Joel Stehouwer, Xiaoming Liu

Prior studies show that the key to face anti-spoofing lies in the subtle image pattern, termed "spoof trace", e.g., color distortion, 3D mask edge, Moire pattern, and many others. Designing a generic anti-spoofing model …

DiversityFace Anti-Spoofing

Face De-Spoofing: Anti-Spoofing via Noise Modeling

2018-07-26 · ECCV 2018 9 · Amin Jourabloo, Yaojie Liu, Xiaoming Liu

Many prior face anti-spoofing works develop discriminative models for recognizing the subtle differences between live and spoof faces. Those approaches often regard the image as an indivisible unit, and process it holist…

DenoisingFace Anti-Spoofing

Physics-Guided Spoof Trace Disentanglement for Generic Face Anti-Spoofing

2020-12-09 · Yaojie Liu, Xiaoming Liu

Prior studies show that the key to face anti-spoofing lies in the subtle image pattern, termed "spoof trace", e.g., color distortion, 3D mask edge, Moire pattern, and many others. Designing a generic face anti-spoofing m…

DisentanglementFace Anti-Spoofing

Spoof Trace Disentanglement for generic face antispoofing

2023-03-03 · journal 2023 3 · Yaojie Liu and Xiaoming Liu, Member, IEEE

Prior studies show that the key to face anti-spoofing lies in the subtle image patterns, termed “spoof trace,” e.g., color distortion, 3D mask edge, and Moire pattern. Spoof detection rooted on those spoof traces can im…

Data AugmentationDisentanglementFace Anti-Spoofing