paper-with-me

홈 › Papers

The Toybox Dataset of Egocentric Visual Object Transformations

2018-06-15 · Xiaohan Wang, Tengyu Ma, James Ainooson, Seunghwan Cha, Xiaotian Wang, Azhar Molla, Maithilee Kunda

In object recognition research, many commonly used datasets (e.g., ImageNet and similar) contain relatively sparse distributions of object instances and views, e.g., one might see a thousand different pictures of a thousand different giraffes, mostly taken from a few conventionally photographed angles. These distributional properties constrain the types of computational experiments that are able to be conducted with such datasets, and also do not reflect naturalistic patterns of embodied visual experience. As a contribution to the small (but growing) number of multi-view object datasets that have been created to bridge this gap, we introduce a new video dataset called Toybox that contains egocentric (i.e., first-person perspective) videos of common household objects and toys being manually manipulated to undergo structured transformations, such as rotation, translation, and zooming. To illustrate potential uses of Toybox, we also present initial neural network experiments that examine 1) how training on different distributions of object instances and views affects recognition performance, and 2) how viewpoint-dependent object concepts are represented within the hidden layers of a trained network.

📄 PDF Abstract BibTeX arXiv:1806.06034

Code (0)

등록된 구현이 없습니다.

Tasks

ObjectObject RecognitionTranslation

Similar Papers 제목 키워드 기반

A Computational Account Of Self-Supervised Visual Learning From Egocentric Object Play

2023-05-30 · Deepayan Sanyal, Joel Michelson, Yuan Yang, James Ainooson 외

Research in child development has shown that embodied experience handling physical objects contributes to many cognitive abilities, including visual learning. One characteristic of such experience is that the learner see…

Contrastive Learningimage-classificationImage ClassificationObject

Toybox: A Suite of Environments for Experimental Evaluation of Deep Reinforcement Learning

2019-05-07 · Emma Tosch, Kaleigh Clary, John Foley, David Jensen

Evaluation of deep reinforcement learning (RL) is inherently challenging. In particular, learned policies are largely opaque, and hypotheses about the behavior of deep RL agents are difficult to test in black-box environ…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Ego-InBetween: Generating Object State Transitions in Ego-Centric Videos

2026-04-20 · Mengmeng Ge, Takashi Isobe, Xu Jia, Yanan Sun 외 arxiv

Understanding physical transformation processes is crucial for both human cognition and artificial intelligence systems, particularly from an egocentric perspective, which serves as a key bridge between humans and machin…

Visual Intention Grounding for Egocentric Assistants

2025-04-18 · Pengzhan Sun, Junbin Xiao, Tze Ho Elden Tse, Yicong Li 외

Visual grounding associates textual descriptions with objects in an image. Conventional methods target third-person image inputs and named object queries. In applications such as AI assistants, the perspective shifts -- …

ObjectVisual Grounding

SymmGrid: Super-Scaling On-Robot Learning with Parallelized Symmetries and Egocentric-Exocentric Visual Perception

2026-07-29 · Gabe Everett, Brice Gunter, Ryan Vander Stelt, Cleiver Ruiz-Martinez 외 arxiv

Deep reinforcement policy learning directly in physical robots (on-robot learning) remains bottlenecked by slow wall-clock training times. We present SymmGrid, a trajectory level augmentation framework inspired by parall…

Robot Manipulation