paper-with-me

홈 › Papers

Partner Approximating Learners (PAL): Simulation-Accelerated Learning with Explicit Partner Modeling in Multi-Agent Domains

2019-09-09 · Florian Köpf, Alexander Nitsch, Michael Flad, Sören Hohmann

Mixed cooperative-competitive control scenarios such as human-machine interaction with individual goals of the interacting partners are very challenging for reinforcement learning agents. In order to contribute towards intuitive human-machine collaboration, we focus on problems in the continuous state and control domain where no explicit communication is considered and the agents do not know the others' goals or control laws but only sense their control inputs retrospectively. Our proposed framework combines a learned partner model based on online data with a reinforcement learning agent that is trained in a simulated environment including the partner model. Thus, we overcome drawbacks of independent learners and, in addition, benefit from a reduced amount of real world data required for reinforcement learning which is vital in the human-machine context. We finally analyze an example that demonstrates the merits of our proposed framework which learns fast due to the simulated environment and adapts to the continuously changing partner due to the partner approximation.

📄 PDF Abstract BibTeX arXiv:1909.03868

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

InqEduAgent: Adaptive AI Learning Partners with Gaussian Process Augmentation

2025-08-05 · Wen-Xi Yang, Tian-Fang Zhao, Guan Liu, Liang Yang 외 arxiv

Collaborative partnership matters in inquiry-oriented education. However, most study partners are selected either rely on experience-based assignments with little scientific planning or build on rule-based machine assist…

Learnersourcing in the Age of AI: Student, Educator and Machine Partnerships for Content Creation

2023-06-10 · Hassan Khosravi, Paul Denny, Steven Moore, John Stamper

Engaging students in creating novel content, also referred to as learnersourcing, is increasingly recognised as an effective approach to promoting higher-order learning, deeply engaging students with course material and …

Structural plasticity on an accelerated analog neuromorphic hardware system

2019-12-27 · Sebastian Billaudelle, Benjamin Cramer, Mihai A. Petrovici, Korbinian Schreiber 외

In computational neuroscience, as well as in machine learning, neuromorphic devices promise an accelerated and scalable alternative to neural network simulations. Their neural connectivity and synaptic capacity depends o…

Computational Efficiency

Fast Posterior Estimation of Cardiac Electrophysiological Model Parameters via Bayesian Active Learning

2021-10-13 · Md Shakil Zaman, Jwala Dhamala, Pradeep Bajracharya, John L. Sapp 외

Probabilistic estimation of cardiac electrophysiological model parameters serves an important step towards model personalization and uncertain quantification. The expensive computation associated with these model simulat…

Active Learning

Embrace rejection: Kernel matrix approximation by accelerated randomly pivoted Cholesky

2024-10-04 · Ethan N. Epperly, Joel A. Tropp, Robert J. Webber

Randomly pivoted Cholesky (RPCholesky) is an algorithm for constructing a low-rank approximation of a positive-semidefinite matrix using a small number of columns. This paper develops an accelerated version of RPCholesky…

Computational chemistry