paper-with-me

홈 › Papers

Assistant Placement Aria: A Benchmark for Egocentric Placement Assistance

2026-08-01 · Amir Belder, Gonçalo Dias Pais, Refael Vivanti, Omri Carmi, Daniel DeTone, Oren Shrout, Ido Gattegno, Ayellet Tal arxiv

Human assistance in robotics spans around several tasks such as navigation, object manipulation, and placement, where a key challenge is selecting target destinations that align with human intentions or preferences. We focus on this challenge in the context of Virtual Placement (VP), the task of identifying all plausible target locations given scene context and human-centric constraints. This differs from traditional placement tasks that typically focus on a single, predefined target location. The VP problem is complex, as it requires both global and local reasoning about the scene's geometry, semantics, and plausibility. To address this gap, we introduce {\bf Assistant Placement Aria}, the first benchmark to explore diverse aspects of VP, including global, local, and human-centric constraints. It contains both synthetic and real indoor scenes annotated for three tasks: (i)~2D Panel Placement, (ii)~Sitting Suggestion, and (iii)~TV Placement. Each scene includes 2D images, a 3D point cloud, and a textual description of the objects within the scene. By contributing this benchmark, we aim to encourage further research in this underexplored and challenging field that is critically dependent on relevant data. We also evaluate several foundation models for object detection and segmentation on our benchmark.

📄 PDF Abstract BibTeX arXiv:2608.00652

Code (0)

등록된 구현이 없습니다.

Tasks

Object Detection

Similar Papers 제목 키워드 기반

WEAR: An Outdoor Sports Dataset for Wearable and Egocentric Activity Recognition

2023-04-11 · Marius Bock, Hilde Kuehne, Kristof Van Laerhoven, Michael Moeller

Research has shown the complementarity of camera- and inertial-based data for modeling human activities, yet datasets with both egocentric video and inertial-based sensor data remain scarce. In this paper, we introduce W…

Action DetectionAction LocalizationEgocentric Activity RecognitionHuman Activity Recognition+2

GOPLA: Generalizable Object Placement Learning via Synthetic Augmentation of Human Arrangement

2025-10-16 · Yao Zhong, Hanzhi Chen, Simon Schaefer, Anran Zhang 외 arxiv

Robots are expected to serve as intelligent assistants, helping humans with everyday household organization. A central challenge in this setting is the task of object placement, which requires reasoning about both semant…

Collision Avoidance

Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions

2025-08-06 · Liang Xu, Chengqun Yang, Zili Lin, Fei Xu 외 arxiv

Learning action models from real-world human-centric interaction datasets is important towards building general-purpose intelligent assistants with efficiency. However, most existing datasets only offer specialist intera…

EgoArgus: Benchmarking VLMs as Situational Assistants for Modality-Grounded User Supports

2026-08-26 · Yu-Chien Tang, Yu-Hsiang Liu, An-Zi Yen arxiv

VLMs are increasingly positioned as daily assistants that perceive first-person environments, follow user dialogue, and decide how to help. Existing egocentric benchmarks mainly evaluate visual understanding in isolation…

Building Egocentric Procedural AI Assistant: Methods, Benchmarks, and Challenges

2025-11-17 · Junlong Li, Huaiyuan Xu, Sijie Cheng, Kejun Wu 외 arxiv

Driven by recent advances in vision-language models (VLMs) and egocentric perception research, the emerging topic of an egocentric procedural AI assistant (EgoProceAssist) is introduced to step-by-step support daily proc…

Question Answering