paper-with-me

홈 › Papers

Hearts Gym: Learning Reinforcement Learning as a Team Event

2022-09-07 · Jan Ebert, Danimir T. Doncevic, Ramona Kloß, Stefan Kesselheim

Amidst the COVID-19 pandemic, the authors of this paper organized a Reinforcement Learning (RL) course for a graduate school in the field of data science. We describe the strategy and materials for creating an exciting learning experience despite the ubiquitous Zoom fatigue and evaluate the course qualitatively. The key organizational features are a focus on a competitive hands-on setting in teams, supported by a minimum of lectures providing the essential background on RL. The practical part of the course revolved around Hearts Gym, an RL environment for the card game Hearts that we developed as an entry-level tutorial to RL. Participants were tasked with training agents to explore reward shaping and other RL hyperparameters. For a final evaluation, the agents of the participants competed against each other.

📄 PDF Abstract BibTeX arXiv:2209.05466

Code (1)

helmholtzai-fzj/hearts-gym 공식 구현 tf

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Personalized HeartSteps: A Reinforcement Learning Algorithm for Optimizing Physical Activity

2019-09-08 · Peng Liao, Kristjan Greenewald, Predrag Klasnja, Susan Murphy

With the recent evolution of mobile health technologies, health scientists are increasingly interested in developing just-in-time adaptive interventions (JITAIs), typically delivered via notification on mobile device and…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

HeartSpot: Privatized and Explainable Data Compression for Cardiomegaly Detection

2022-10-05 · Elvin Johnson, Shreshta Mohan, Alex Gaudio, Asim Smailagic 외

Advances in data-driven deep learning for chest X-ray image analysis underscore the need for explainability, privacy, large datasets and significant computational resources. We frame privacy and explainability as a lossy…

Data CompressionImage Compression

Bold Hearts Team Description for RoboCup 2019 (Humanoid Kid Size League)

2019-04-22 · Marcus M. Scheunemann, Sander G. van Dijk, Rebecca Miko, Daniel Barry 외

We participated in the RoboCup 2018 competition in Montreal with our newly developed BoldBot based on the Darwin-OP and mostly self-printed custom parts. This paper is about the lessons learnt from that competition and f…

Semantic Segmentation

Cardiac kinematic parameters computed from video of $\textit{in situ}$ beating heart

2020-02-05 · Lorenzo Fassina, Giacomo Rozzi, Stefano Rossi, Simone Scacchi 외

Mechanical function of the heart during open-chest cardiac surgery is exclusively monitored by echocardiographic techniques. However, little is known about local kinematics, particularly for the reperfused regions after …

HEARTS: Benchmarking LLM Reasoning on Health Time Series

2026-02-25 · Sirui Li, Shuhan Xiao, Mihir Joshi, Ahmed Metwally 외 arxiv

The rise of large language models (LLMs) has shifted time series analysis from narrow analytics to general-purpose reasoning. Yet, existing benchmarks cover only a small set of health time series modalities and tasks, fa…

Time Series Analysis