paper-with-me

홈 › Papers

GazePrior: Zero-Shot AR/VR Eye Tracking via Learned 3D Gaze Reconstruction

2026-05-21 · Corentin Dumery, David Colmenares, Alexander Fix, Pascal Fua, Ali Behrooz, Jogendra Kundu arxiv

Eye tracking (ET) is a foundational technology for advanced AR/VR applications. However, training ET models for every new ET device is challenging: real data collection is costly and time-consuming, while existing synthetic data generation methods lack realism. To remove the need for additional data collection while maintaining data quality, we introduce a data-driven 3D prior that models the distribution of human eyes across diverse identities, gaze directions, and light settings. This model, which we coin GazePrior, then enables sparse-input 3D reconstruction of annotated data collected with previous ET devices, which can in turn be rendered from the cameras of any target ET device. Our approach synthesizes data with the realism, diversity and ground-truth accuracy of real data collection without its prohibitive costs. Our experiments demonstrate that ET models trained with our synthesized data outperform previous zero-shot methods, achieving higher accuracy and robustness.

📄 PDF Abstract BibTeX arXiv:2605.22359

Code (0)

등록된 구현이 없습니다.

Tasks

Synthetic Data Generation3D Reconstruction

Similar Papers 제목 키워드 기반

Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs

2026-03-05 · Yiwei Li, Zihao Wu, Yanjun Lv, Hanqi Jiang 외 arxiv

Vision--language models (VLMs) process images as visual tokens, yet their intermediate reasoning is often carried out in text, which can be suboptimal for visually grounded radiology tasks. Radiologists instead diagnose …

Visual Reasoning

Zero-Shot Gaze-based Volumetric Medical Image Segmentation

2025-05-21 · Tatyana Shmykova, Leila Khaertdinova, Ilya Pershin

Accurate segmentation of anatomical structures in volumetric medical images is crucial for clinical applications, including disease monitoring and cancer treatment planning. Contemporary interactive segmentation models, …

Image SegmentationInteractive SegmentationMedical Image SegmentationSegmentation+2

Gaze Embeddings for Zero-Shot Image Classification

2016-11-28 · CVPR 2017 7 · Nour Karessli, Zeynep Akata, Bernt Schiele, Andreas Bulling

Zero-shot image classification using auxiliary information, such as attributes describing discriminative object properties, requires time-consuming annotation by domain experts. We instead propose a method that relies on…

ClassificationFine-Grained Image ClassificationGeneral Classificationimage-classification+2

Goal-Oriented Gaze Estimation for Zero-Shot Learning

2021-03-05 · CVPR 2021 1 · Yang Liu, Lei Zhou, Xiao Bai, Yifei HUANG 외

Zero-shot learning (ZSL) aims to recognize novel classes by transferring semantic knowledge from seen classes to unseen classes. Since semantic knowledge is built on attributes shared between different classes, which are…

AttributeGaze EstimationGeneralized Zero-Shot LearningZero-Shot Learning

A World Model of Radiologist Reading for Medical Image Representation Learning

2026-05-17 · Yiwei Li, Zihao Wu, Huaqin Zhao, Yifan Zhou 외 arxiv

Radiologist eye-tracking data provide a rich record of how experts search, compare, and accumulate evidence during image reading; yet, existing methods exploit this signal only partially, either as a static spatial prior…

Representation Learning