paper-with-me

Papers

Learning by Cheating

2019-12-27 · Dian Chen, Brady Zhou, Vladlen Koltun, Philipp Krähenbühl

Vision-based urban driving is hard. The autonomous system needs to learn to perceive the world and act in it. We show that this challenging learning problem can be simplified by decomposing it into two stages. We first train an agent that has access to privileged information. This privileged agent cheats by observing the ground-truth layout of the environment and the positions of all traffic participants. In the second stage, the privileged agent acts as a teacher that trains a purely vision-based sensorimotor agent. The resulting sensorimotor agent does not have access to any privileged information and does not cheat. This two-stage training procedure is counter-intuitive at first, but has a number of important advantages that we analyze and empirically demonstrate. We use the presented approach to train a vision-based autonomous driving system that substantially outperforms the state of the art on the CARLA benchmark and the recent NoCrash benchmark. Our approach achieves, for the first time, 100% success rate on all tasks in the original CARLA benchmark, sets a new record on the NoCrash benchmark, and reduces the frequency of infractions by an order of magnitude compared to the prior state of the art. For the video that summarizes this work, see https://youtu.be/u9ZCxxD-UUw

📄 PDF Abstract BibTeX arXiv:1912.12294

Code (9)

dotchen/LearningByCheating 공식 구현 pytorch
SimarKareer/legged_gym pytorch
bradyz/2020_CARLA_challenge
deepsense-ai/carla-birdeye-view
jostl/masters-thesis pytorch
piazzesiNiccolo/myLbc pytorch
scope-lab-vu/anti-carla
scope-lab-vu/risk-aware-scene-generation-cps
zwc662/SequentialAttack pytorch

Tasks

Autonomous Driving

Methods 이 논문이 사용한 방법론

Entropy Regularization 설명 없음
PPO Proximal Policy Optimization, or PPO, is a policy gradient method for reinforcement learning. The motivation was to have an algorithm with the data efficiency and reliable…
CARLA CARLA is an open-source simulator for autonomous driving research. CARLA has been developed from the ground up to support development, training, and validation of autonomous urban…

Similar Papers 제목 키워드 기반

On Perception of Prevalence of Cheating and Usage of Generative AI

2024-05-29 · Roman Denkin

This report investigates the perceptions of teaching staff on the prevalence of student cheating and the impact of Generative AI on academic integrity. Data was collected via an anonymous survey of teachers at the Depart…

Do Coding Agents Deceive Us? Detecting and Preventing Cheating via Capped Evaluation with Randomized Tests

2026-06-05 · Thanawat Lodkaew, Johannes Ackermann, Soichiro Nishimori, Nontawat Charoenphakdee 외 arxiv

A growing failure mode in agent evaluation and training is that models can achieve high evaluation scores by exploiting shortcuts instead of solving the intended task, producing deceptive performance. This makes evaluati…

Human-in-the-Loop AI for Cheating Ring Detection

2024-03-18 · Yong-Siang Shih, Manqian Liao, Ruidong Liu, Mirza Basim Baig

Online exams have become popular in recent years due to their accessibility. However, some concerns have been raised about the security of the online exams, particularly in the context of professional cheating services a…

Fairness

How Much Can a Few Engine Moves Help? Quantifying Limited Cheating in Chess

2026-01-08 · Daniel Keren arxiv

Cheating in chess, by using advice from powerful software, has become a major problem, reaching the highest levels. As opposed to the large majority of previous work, which concerned {\em detection} of cheating, here we …

Hyperparameter Optimization

Applying IRT to Distinguish Between Human and Generative AI Responses to Multiple-Choice Assessments

2024-11-28 · Alona Strugatski, Giora Alexandron

Generative AI is transforming the educational landscape, raising significant concerns about cheating. Despite the widespread use of multiple-choice questions in assessments, the detection of AI cheating in MCQ-based test…

Multiple-choice