paper-with-me

Papers

Temporal Coherence for Active Learning in Videos

2019-08-30 · Javad Zolfaghari Bengar, Abel Gonzalez-Garcia, Gabriel Villalonga, Bogdan Raducanu, Hamed H. Aghdam, Mikhail Mozerov, Antonio M. Lopez, Joost Van de Weijer

Autonomous driving systems require huge amounts of data to train. Manual annotation of this data is time-consuming and prohibitively expensive since it involves human resources. Therefore, active learning emerged as an alternative to ease this effort and to make data annotation more manageable. In this paper, we introduce a novel active learning approach for object detection in videos by exploiting temporal coherence. Our active learning criterion is based on the estimated number of errors in terms of false positives and false negatives. The detections obtained by the object detector are used to define the nodes of a graph and tracked forward and backward to temporally link the nodes. Minimizing an energy function defined on this graphical model provides estimates of both false positives and false negatives. Additionally, we introduce a synthetic video dataset, called SYNTHIA-AL, specially designed to evaluate active learning for video object detection in road scenes. Finally, we show that our approach outperforms active learning baselines tested on two datasets.

📄 PDF Abstract BibTeX arXiv:1908.11757

Code (0)

등록된 구현이 없습니다.

Tasks

Active LearningAutonomous DrivingObjectobject-detectionObject DetectionVideo Object Detection

Similar Papers 제목 키워드 기반

EmoCo: Visual Analysis of Emotion Coherence in Presentation Videos

2019-07-29 · Haipeng Zeng, Xingbo Wang, Aoyu Wu, Yong Wang 외

Emotions play a key role in human communication and public presentations. Human emotions are usually expressed through multiple modalities. Therefore, exploring multimodal emotions and their coherence is of great value f…

ClusteringSentence

Video Face Editing Using Temporal-Spatial-Smooth Warping

2014-08-11 · Xiaoyan Li, DaCheng Tao

Editing faces in videos is a popular yet challenging aspect of computer vision and graphics, which encompasses several applications including facial attractiveness enhancement, makeup transfer, face replacement, and expr…

Exploring Temporal Coherence for More General Video Face Forgery Detection

2021-08-15 · ICCV 2021 10 · Yinglin Zheng, Jianmin Bao, Dong Chen, Ming Zeng 외

Although current face manipulation techniques achieve impressive performance regarding quality and controllability, they are struggling to generate temporal coherent face videos. In this work, we explore to take full adv…

DeepFake Detection

Edit Temporal-Consistent Videos with Image Diffusion Model

2023-08-17 · Yuanzhi Wang, Yong Li, Xiaoya Zhang, Xin Liu 외

Large-scale text-to-image (T2I) diffusion models have been extended for text-guided video editing, yielding impressive zero-shot video editing performance. Nonetheless, the generated videos usually show spatial irregular…

modelVideo EditingVideo Temporal Consistency

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence

2026-04-25 · Yahui Li, Yinfeng Yu, Liejun Wang, Shengjie Shen arxiv

Emotionally talking head video generation aims to generate expressive portrait videos with accurate lip synchronization and emotional facial expressions. Current methods rely on simple emotional labels, leading to insuff…

Graph structure learningComputational EfficiencyTalking Head GenerationVideo Generation