paper-with-me

Papers

Synthetically Generating Human-like Data for Sequential Decision Making Tasks via Reward-Shaped Imitation Learning

2023-04-14 · Bryan Brandt, Prithviraj Dasgupta

We consider the problem of synthetically generating data that can closely resemble human decisions made in the context of an interactive human-AI system like a computer game. We propose a novel algorithm that can generate synthetic, human-like, decision making data while starting from a very small set of decision making data collected from humans. Our proposed algorithm integrates the concept of reward shaping with an imitation learning algorithm to generate the synthetic data. We have validated our synthetic data generation technique by using the synthetically generated data as a surrogate for human interaction data to solve three sequential decision making tasks of increasing complexity within a small computer game-like setup. Different empirical and statistical analyses of our results show that the synthetically generated data can substitute the human data and perform the game-playing tasks almost indistinguishably, with very low divergence, from a human performing the same tasks.

📄 PDF Abstract BibTeX arXiv:2304.07280

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingImitation LearningSequential Decision MakingSynthetic Data Generation

Similar Papers 제목 키워드 기반

Unified Pragmatic Models for Generating and Following Instructions

2017-11-14 · NAACL 2018 6 · Daniel Fried, Jacob Andreas, Dan Klein

We show that explicit pragmatic inference aids in correctly generating and following natural language instructions for complex, sequential tasks. Our pragmatics-enabled models reason about why speakers produce certain in…

Text Generation

From Seed to Harvest: Augmenting Human Creativity with AI for Red-teaming Text-to-Image Models

2025-07-23 · Jessica Quaye, Charvi Rastogi, Alicia Parrish, Oana Inel 외 arxiv

Text-to-image (T2I) models have become prevalent across numerous applications, making their robust evaluation against adversarial attacks a critical priority. Continuous access to new and challenging adversarial prompts …

The Good, the Bad and the Constructive: Automatically Measuring Peer Review's Utility for Authors

2025-08-31 · Abdelrahman Sadallah, Tim Baumgärtner, Iryna Gurevych, Ted Briscoe arxiv

Providing constructive feedback to paper authors is a core component of peer review. With reviewers increasingly having less time to perform reviews, automated support systems are required to ensure high reviewing qualit…

Making Task-Oriented Dialogue Datasets More Natural by Synthetically Generating Indirect User Requests

2024-06-12 · Amogh Mannekote, Jinseok Nam, Ziming Li, Jian Gao 외

Indirect User Requests (IURs), such as "It's cold in here" instead of "Could you please increase the temperature?" are common in human-human task-oriented dialogue and require world knowledge and pragmatic reasoning from…

Dialogue State TrackingNatural Language UnderstandingTask-Oriented Dialogue SystemsWorld Knowledge

AnthroNet: Conditional Generation of Humans via Anthropometrics

2023-09-07 · Francesco Picetti, Shrinath Deshpande, Jonathan Leban, Soroosh Shahtalebi 외

We present a novel human body model formulated by an extensive set of anthropocentric measurements, which is capable of generating a wide range of human body shapes and poses. The proposed model enables direct modeling o…

3D human pose and shape estimation3D Human Reconstruction3D Human Shape EstimationDiversity+2