paper-with-me

Papers

Learning Policies from Human Data for Skat

2019-05-27 · Douglas Rebstock, Christopher Solinas, Michael Buro

Decision-making in large imperfect information games is difficult. Thanks to recent success in Poker, Counterfactual Regret Minimization (CFR) methods have been at the forefront of research in these games. However, most of the success in large games comes with the use of a forward model and powerful state abstractions. In trick-taking card games like Bridge or Skat, large information sets and an inability to advance the simulation without fully determinizing the state make forward search problematic. Furthermore, state abstractions can be especially difficult to construct because the precise holdings of each player directly impact move values. In this paper we explore learning model-free policies for Skat from human game data using deep neural networks (DNN). We produce a new state-of-the-art system for bidding and game declaration by introducing methods to a) directly vary the aggressiveness of the bidder and b) declare games based on expected value while mitigating issues with rarely observed state-action pairs. Although cardplay policies learned through imitation are slightly weaker than the current best search-based method, they run orders of magnitude faster. We also explore how these policies could be learned directly from experience in a reinforcement learning setting and discuss the value of incorporating human data for this task.

📄 PDF Abstract BibTeX arXiv:1905.10907

Code (0)

등록된 구현이 없습니다.

Tasks

Card GamescounterfactualDecision MakingReinforcement Learning

Similar Papers 제목 키워드 기반

Learning Roller-Skating Motions of Humanoid Robots Based on Adversarial Motion Priors

2026-07-12 · Yunkang Cheng, Yutong Wu, Menghan Li, Shihe Zhou 외 arxiv

Humanoid roller-skating is difficult because the robot must coordinate whole-body balance, rolling contacts, and velocity-dependent posture regulation. This paper presents an adversarial motion prior based reinforcement …

Reinforcement Learning

The SkatingVerse Workshop & Challenge: Methods and Results

2024-05-27 · Jian Zhao, Lei Jin, Jianshu Li, Zheng Zhu 외

The SkatingVerse Workshop & Challenge aims to encourage research in developing novel and accurate methods for human action understanding. The SkatingVerse dataset used for the SkatingVerse Challenge has been publicly rel…

Action Understanding

Reinforcement Learning-Based Control for an Inline Skating Humanoid Robot

2026-06-30 · Ethan Marot, Thomas Bi, Clemens Schwarke, Victor Klemm 외 arxiv

As humanoid robots become increasingly dynamic, coupling them with reinforcement learning offers a promising approach to solving the complex, underactuated mechanics of passive inline skating. Equipping a humanoid robot …

Reinforcement Learning

3D Pose-Based Temporal Action Segmentation for Figure Skating: A Fine-Grained and Jump Procedure-Aware Annotation Approach

2024-08-29 · Ryota Tanaka, Tomohiro Suzuki, Keisuke Fujii

Understanding human actions from videos is essential in many domains, including sports. In figure skating, technical judgments are performed by watching skaters' 3D movements, and its part of the judging procedure can be…

Action SegmentationMarkerless Motion CaptureTemporal Action Segmentation

SKATER: Synthesized Kinematics for Advanced Traversing Efficiency on a Humanoid Robot via Roller Skate Swizzles

2026-01-08 · Junchi Gu, Feiyang Yuan, Weize Shi, Tianchen Huang 외 arxiv

Although recent years have seen significant progress of humanoid robots in walking and running, the frequent foot strikes with ground during these locomotion gaits inevitably generate high instantaneous impact forces, wh…

Reinforcement Learning