paper-with-me

Papers

Human Gaze Boosts Object-Centered Representation Learning

2025-01-06 · Timothy Schaumlöffel, Arthur Aubret, Gemma Roig, Jochen Triesch

Recent self-supervised learning (SSL) models trained on human-like egocentric visual inputs substantially underperform on image recognition tasks compared to humans. These models train on raw, uniform visual inputs collected from head-mounted cameras. This is different from humans, as the anatomical structure of the retina and visual cortex relatively amplifies the central visual information, i.e. around humans' gaze location. This selective amplification in humans likely aids in forming object-centered visual representations. Here, we investigate whether focusing on central visual information boosts egocentric visual object learning. We simulate 5-months of egocentric visual experience using the large-scale Ego4D dataset and generate gaze locations with a human gaze prediction model. To account for the importance of central vision in humans, we crop the visual area around the gaze location. Finally, we train a time-based SSL model on these modified inputs. Our experiments demonstrate that focusing on central vision leads to better object-centered representations. Our analysis shows that the SSL model leverages the temporal dynamics of the gaze movements to build stronger visual representations. Overall, our work marks a significant step toward bio-inspired learning of visual representations.

📄 PDF Abstract BibTeX arXiv:2501.02966

Code (0)

등록된 구현이 없습니다.

Tasks

Gaze PredictionObjectRepresentation LearningSelf-Supervised Learning

Similar Papers 제목 키워드 기반

Active Gaze Behavior Boosts Self-Supervised Object Learning

2024-11-04 · Zhengyang Yu, Arthur Aubret, Marcel C. Raabe, Jane Yang 외

Due to significant variations in the projection of the same object from different viewpoints, machine learning algorithms struggle to recognize the same object across various perspectives. In contrast, toddlers quickly l…

ObjectObject RecognitionSelf-Supervised Learning

Active Gaze Control for Foveal Scene Exploration

2022-08-24 · Alexandre M. F. Dias, Luís Simões, Plinio Moreno, Alexandre Bernardino

Active perception and foveal vision are the foundations of the human visual system. While foveal vision reduces the amount of information to process during a gaze fixation, active perception will change the gaze directio…

Learning Unsupervised Gaze Representation via Eye Mask Driven Information Bottleneck

2024-06-29 · Yangzhou Jiang, Yinxin Lin, Yaoming Wang, Teng Li 외

Appearance-based supervised methods with full-face image input have made tremendous advances in recent gaze estimation tasks. However, intensive human annotation requirement inhibits current methods from achieving indust…

Gaze EstimationUnsupervised Pre-training

Object-aware Gaze Target Detection

2023-07-18 · ICCV 2023 1 · Francesco Tonini, Nicola Dall'Asen, Cigdem Beyan, Elisa Ricci

Gaze target detection aims to predict the image location where the person is looking and the probability that a gaze is out of the scene. Several works have tackled this task by regressing a gaze heatmap centered on the …

Object

A Self Validation Network for Object-Level Human Attention Estimation

2019-10-31 · NeurIPS 2019 12 · Zehua Zhang, Chen Yu, David Crandall

Due to the foveated nature of the human vision system, people can focus their visual attention on a small region of their visual field at a time, which usually contains only a single object. Estimating this object of att…

Object