paper-with-me

Papers

Alexa Arena: A User-Centric Interactive Platform for Embodied AI

2023-03-02 · NeurIPS 2023 11 · Qiaozi Gao, Govind Thattai, Suhaila Shakiah, Xiaofeng Gao, Shreyas Pansare, Vasu Sharma, Gaurav Sukhatme, Hangjie Shi, Bofei Yang, Desheng Zheng, Lucy Hu, Karthika Arumugam, Shui Hu, Matthew Wen, Dinakar Guthy, Cadence Chung, Rohan Khanna, Osman Ipek, Leslie Ball, Kate Bland, Heather Rocker, Yadunandana Rao, Michael Johnston, Reza Ghanadan, Arindam Mandal, Dilek Hakkani Tur, Prem Natarajan

We introduce Alexa Arena, a user-centric simulation platform for Embodied AI (EAI) research. Alexa Arena provides a variety of multi-room layouts and interactable objects, for the creation of human-robot interaction (HRI) missions. With user-friendly graphics and control mechanisms, Alexa Arena supports the development of gamified robotic tasks readily accessible to general human users, thus opening a new venue for high-efficiency HRI data collection and EAI system evaluation. Along with the platform, we introduce a dialog-enabled instruction-following benchmark and provide baseline results for it. We make Alexa Arena publicly available to facilitate research in building generalizable and assistive embodied agents.

📄 PDF Abstract BibTeX arXiv:2303.01586

Code (1)

amazon-science/alexa-arena 공식 구현 pytorch

Tasks

Instruction Following

Similar Papers 제목 키워드 기반

Arena 4.0: A Comprehensive ROS2 Development and Benchmarking Platform for Human-centric Navigation Using Generative-Model-based Environment Generation

2024-09-19 · Volodymyr Shcherbyna1, Linh Kästner, Diego Diaz, Huu Giang Nguyen 외

Building on the foundations of our previous work, this paper introduces Arena 4.0, a significant advancement over Arena 3.0, Arena-Bench, Arena 1.0, and Arena 2.0. Arena 4.0 offers three key novel contributions: (1) a ge…

BenchmarkingSocial Navigation

Alexa, Let's Work Together: Introducing the First Alexa Prize TaskBot Challenge on Conversational Task Assistance

2022-09-13 · Anna Gottardi, Osman Ipek, Giuseppe Castellucci, Shui Hu 외

Since its inception in 2016, the Alexa Prize program has enabled hundreds of university students to explore and compete to develop conversational agents through the SocialBot Grand Challenge. The goal of the challenge is…

WorldArena 2.0: Extending Embodied World Model Benchmarking on Modality, Functionality and Platform

2026-05-18 · Yu Shang, Yinzhou Tang, Yiding Ma, Zhuohang Li 외 arxiv

World models have emerged as a central paradigm for embodied intelligence, enabling agents to predict action-conditioned future and reason about environmental dynamics. However, existing embodied world model benchmarks a…

CRS Arena: Crowdsourced Benchmarking of Conversational Recommender Systems

2024-12-13 · Nolwenn Bernard, Hideaki Joko, Faegheh Hasibi, Krisztian Balog

We introduce CRS Arena, a research platform for scalable benchmarking of Conversational Recommender Systems (CRS) based on human feedback. The platform displays pairwise battles between anonymous conversational recommend…

BenchmarkingRecommendation Systems

Inclusion Arena: An Open Platform for Evaluating Large Foundation Models with Real-World Apps

2025-08-15 · Kangyu Wang, Hongliang He, Lin Liu, Ruiqi Liang 외 arxiv

Large Language Models (LLMs) and Multimodal Large Language Models (MLLMs) have ushered in a new era of AI capabilities, demonstrating near-human-level performance across diverse scenarios. While numerous benchmarks (e.g.…