paper-with-me

홈 › Papers

Three Birds with One Stone: Multi-Task Temporal Action Detection via Recycling Temporal Annotations

2021-06-19 · CVPR 2021 1 · Zhihui Li, Lina Yao

Temporal action detection on unconstrained videos has seen significant research progress in recent years. Deep learning has achieved enormous success in this direction. However, collecting large-scale temporal detection datasets to ensuring promising performance in the real-world is a laborious, impractical and time consuming process. Accordingly, we present a novel improved temporal action localization model that is better able to take advantage of limited labeled data available. Specifically, we design two auxiliary tasks by reconstructing the available label information and then facilitate the learning of the temporal action detection model. Each task generates their supervision signal by recycling the original annotations, and are jointly trained with the temporal action detection model in a multi-task learning fashion. Note that the proposed approach can be pluggable to any region proposal based temporal action detection models. We conduct extensive experiments on three benchmark datasets, namely THUMOS'14, Charades and ActivityNet. Our experimental results confirm the effectiveness of the proposed model.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Action DetectionAction LocalizationMulti-Task LearningRegion ProposalTemporal Action Localization

Similar Papers 제목 키워드 기반

ECNU: One Stone Two Birds: Ensemble of Heterogenous Measures for Semantic Relatedness and Textual Entailment

2014-08-01 · SEMEVAL 2014 8 · Jiang Zhao, Tiantian Zhu, Man Lan
Natural Language InferenceSemantic Textual Similarity

Transduction with Matrix Completion: Three Birds with One Stone

2010-12-01 · NeurIPS 2010 12 · Andrew Goldberg, Ben Recht, Jun-Ming Xu, Robert Nowak 외

We pose transductive classification as a matrix completion problem. By assuming the underlying matrix has a low rank, our formulation is able to handle three problems simultaneously: i) multi-label learning, where each i…

General ClassificationMatrix CompletionMulti-Label Learning

One Stone, Three Birds: Self-adaptive Optimal Transport for Multi-VLM Selection, Adaptation, and Ensembling

2026-06-06 · Qiyu Xu, Zhanxuan Hu, Yu Duan, Yonghang Tai 외 arxiv

Vision-language models (VLMs) enable visual recognition from semantic class descriptions, which makes them attractive when target annotations are scarce or unavailable. Most deployment pipelines, however, first choose a …

Automatic recognition of element classes and boundaries in the birdsong with variable sequences

2016-01-23 · Takuya Koumura, Kazuo Okanoya

Researches on sequential vocalization often require analysis of vocalizations in long continuous sounds. In such studies as developmental ones or studies across generations in which days or months of vocalizations must b…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Boundary DetectionGeneral Classification+2

Three Birds One Stone: A General Architecture for Salient Object Segmentation, Edge Detection and Skeleton Extraction

2018-03-27 · Qibin Hou, Jiang-Jiang Liu, Ming-Ming Cheng, Ali Borji 외

In this paper, we aim at solving pixel-wise binary problems, including salient object segmentation, skeleton extraction, and edge detection, by introducing a unified architecture. Previous works have proposed tailored me…

Edge DetectionSemantic Segmentation