paper-with-me

홈 › Papers

MASAI: Multi-agent Summative Assessment Improvement for Unsupervised Environment Design

2021-06-13 · ICML Workshop URL 2021 7 · Yiping Wang, Michael Brandon Haworth

Reinforcement Learning agents require a distribution of environments for their policy to be trained on. The method or process of defining these environments directly impacts robustness and generalization of the learned agent policies. In single agent reinforcement learning, this problem is often solved by domain randomization, or randomizing the environment and tasks within the scope of the desired operating domain of the agent. The challenge here is to generate both structured and solvable environments that guide the agent's learning process. Most recently, works have sought to produce the environments under the Unsupervised Environment Design (UED) formulation. However, these methods lead to a proliferation of adversarial agents to train one agent for a single agent problem in a discretized task domain. In this work, we aim to automatically generate environments that are solvable and challenging for the continuous multi-agent setting. We base our solution on the Teacher-Student relationship with parameter sharing $\textit{Students}$ where we re-imagine the $\textit{Teacher}$ as an environment generator for UED. Our approach uses one environment generator agent ($\textit{Teacher}$) for any number of learning agents ($\textit{Students}$). We qualitatively and quantitatively demonstrate that, in terms of multi-agent ($\geq$ 8 agents) navigation and steering, $\textit{Students}$ trained by our approach outperform agents using heuristic search, as well as agents trained by domain randomization. Our code is available at https://github.com/GAIDG-Lab/MASAI.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Heuristic Searchreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

MASAI: Modular Architecture for Software-engineering AI Agents

2024-06-17 · Daman Arora, Atharv Sonwane, Nalin Wadhwa, Abhav Mehrotra 외

A common method to solve complex problems in software engineering, is to divide the problem into multiple sub-problems. Inspired by this, we propose a Modular Architecture for Software-engineering AI (MASAI) agents, wher…

Video-based Formative and Summative Assessment of Surgical Tasks using Deep Learning

2022-03-17 · Erim Yanik, Uwe Kruger, Xavier Intes, Rahul Rahul 외

To ensure satisfactory clinical outcomes, surgical skill assessment must be objective, time-efficient, and preferentially automated - none of which is currently achievable. Video-based assessment (VBA) is being deployed …

Hybrid E-Assessment in Higher Education: Semi-Automated Grading of Paper-Based Written Examinations

2026-06-07 · Hartwig Grabowski, Michael Canz arxiv

This paper examines the limitations of fully digital and partially digital e-assessment approaches in summative examinations in higher education. The analysis focuses on the didactic narrowing caused by closed question f…

Comparing Human and Automated Evaluation of Open-Ended Student Responses to Questions of Evolution

2016-03-22 · Michael J Wiser, Louise S Mead, James J Smith, Robert T. Pennock

Written responses can provide a wealth of data in understanding student reasoning on a topic. Yet they are time- and labor-intensive to score, requiring many instructors to forego them except as limited parts of summativ…

Tell Me Who Your Students Are: GPT Can Generate Valid Multiple-Choice Questions When Students' (Mis)Understanding Is Hinted

2025-05-09 · Machi Shimmei, Masaki Uto, Yuichiroh Matsubayashi, Kentaro Inui 외

The primary goal of this study is to develop and evaluate an innovative prompting technique, AnaQuest, for generating multiple-choice questions (MCQs) using a pre-trained large language model. In AnaQuest, the choice ite…

Language ModelingLanguage ModellingLarge Language ModelMultiple-choice+2