paper-with-me

홈 › Papers

Panonut360: A Head and Eye Tracking Dataset for Panoramic Video

2024-03-26 · Yutong Xu, Junhao Du, Jiahe Wang, Yuwei Ning, Sihan Zhou Yang Cao

With the rapid development and widespread application of VR/AR technology, maximizing the quality of immersive panoramic video services that match users' personal preferences and habits has become a long-standing challenge. Understanding the saliency region where users focus, based on data collected with HMDs, can promote multimedia encoding, transmission, and quality assessment. At the same time, large-scale datasets are essential for researchers and developers to explore short/long-term user behavior patterns and train AI models related to panoramic videos. However, existing panoramic video datasets often include low-frequency user head or eye movement data through short-term videos only, lacking sufficient data for analyzing users' Field of View (FoV) and generating video saliency regions. Driven by these practical factors, in this paper, we present a head and eye tracking dataset involving 50 users (25 males and 25 females) watching 15 panoramic videos. The dataset provides details on the viewport and gaze attention locations of users. Besides, we present some statistics samples extracted from the dataset. For example, the deviation between head and eye movements challenges the widely held assumption that gaze attention decreases from the center of the FoV following a Gaussian distribution. Our analysis reveals a consistent downward offset in gaze fixations relative to the FoV in experimental settings involving multiple users and videos. That's why we name the dataset Panonut, a saliency weighting shaped like a donut. Finally, we also provide a script that generates saliency distributions based on given head or eye coordinates and pre-generated saliency distribution map sets of each video from the collected eye tracking data. The dataset is available on website: https://dianvrlab.github.io/Panonut360/.

📄 PDF Abstract BibTeX arXiv:2403.17708

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Using Panoramic Videos for Multi-person Localization and Tracking in a 3D Panoramic Coordinate

2019-11-24 · Fan Yang, Feiran Li, Yang Wu, Sakriani Sakti 외

3D panoramic multi-person localization and tracking are prominent in many applications, however, conventional methods using LiDAR equipment could be economically expensive and also computationally inefficient due to the …

Multi-Object Tracking

PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation

2023-09-21 · Shilin Yan, Xiaohao Xu, Renrui Zhang, Lingyi Hong 외

Panoramic videos contain richer spatial information and have attracted tremendous amounts of attention due to their exceptional experience in some fields such as autonomous driving and virtual reality. However, existing …

Autonomous DrivingSegmentationSemantic SegmentationVideo Object Segmentation+2

A Log-Rectilinear Transformation for Foveated 360-degree Video Streaming

2021-03-22 · David Li, Ruofei Du, Adharsh Babu, Camelia Brumar 외

With the rapidly increasing resolutions of 360° cameras, head-mounted displays, and live-streaming services, streaming high-resolution panoramic videos over limited-bandwidth networks is becoming a critical challenge. Fo…

People Tracking in Panoramic Video for Guiding Robots

2022-06-06 · Alberto Bacchin, Filippo Berno, Emanuele Menegatti, Alberto Pretto

A guiding robot aims to effectively bring people to and from specific places within environments that are possibly unknown to them. During this operation the robot should be able to detect and track the accompanied perso…

Skeleton-based Group Activity Recognition via Spatial-Temporal Panoramic Graph

2024-07-28 · Zhengcen Li, Xinle Chang, Yueran Li, Jingyong Su

Group Activity Recognition aims to understand collective activities from videos. Existing solutions primarily rely on the RGB modality, which encounters challenges such as background variations, occlusions, motion blurs,…

Activity RecognitionGroup Activity RecognitionPose Estimation