A gaze driven fast-forward method for first-person videos
The growing data sharing and life-logging cultures are driving an unprecedented increase in the amount of unedited First-Person Videos. In this paper, we address the problem of accessing relevant information in First-Person Videos by creating an accelerated version of the input video and emphasizing the important moments to the recorder. Our method is based on an attention model driven by gaze and visual scene analysis that provides a semantic score of each frame of the input video. We performed several experimental evaluations on publicly available First-Person Videos datasets. The results show that our methodology can fast-forward videos emphasizing moments when the recorder visually interact with scene components while not including monotonous clips.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Sharingan: A Transformer-based Architecture for Gaze Following
Gaze is a powerful form of non-verbal communication and social interaction that humans develop from an early age. As such, modeling this behavior is an important task that can benefit a broad set of application domains r…
Gaze PredictionPredictionSociologyGazeOnce: Real-Time Multi-Person Gaze Estimation
Appearance-based gaze estimation aims to predict the 3D eye gaze direction from a single image. While recent deep learning-based approaches have demonstrated excellent performance, they usually assume one calibrated face…
Gaze EstimationGazeGPT: Augmenting Human Capabilities using Gaze-contingent Contextual AI for Smart Eyewear
Multimodal large language models (LMMs) excel in world knowledge and problem-solving abilities. Through the use of a world-facing camera and contextual AI, emerging smart accessories aim to provide a seamless interface b…
World KnowledgeGazeGen: Gaze-Driven User Interaction for Visual Content Generation
We present GazeGen, a user interaction system that generates visual content (images and videos) for locations indicated by the user's eye gaze. GazeGen allows intuitive manipulation of visual content by targeting regions…
Gaze EstimationKnowledge DistillationRaspberry Pi 4Personality-Driven Gaze Animation with Conditional Generative Adversarial Networks
We present a generative adversarial learning approach to synthesize gaze behavior of a given personality. We train the model using an existing data set that comprises eye-tracking data and personality traits of 42 partic…
Time SeriesTime Series Analysis