The Runner-up Solution for YouTube-VIS Long Video Challenge 2022
This technical report describes our 2nd-place solution for the ECCV 2022 YouTube-VIS Long Video Challenge. We adopt the previously proposed online video instance segmentation method IDOL for this challenge. In addition, we use pseudo labels to further help contrastive learning, so as to obtain more temporally consistent instance embedding to improve tracking performance between frames. The proposed method obtains 40.2 AP on the YouTube-VIS 2022 long video dataset and was ranked second place in this challenge. We hope our simple and effective method could benefit further research.
Code (0)
등록된 구현이 없습니다.
Tasks
Contrastive LearningInstance SegmentationSemantic SegmentationVideo Instance SegmentationSimilar Papers 제목 키워드 기반
DreamRunner: Fine-Grained Storytelling Video Generation with Retrieval-Augmented Motion Adaptation
Storytelling video generation (SVG) has recently emerged as a task to create long, multi-motion, multi-scene videos that consistently represent the story described in the input text script. SVG holds great potential for …
Large Language ModelMotion PlanningRetrievalTest-time Adaptation+2Heuristics2Annotate: Efficient Annotation of Large-Scale Marathon Dataset For Bounding Box Regression
Annotating a large-scale in-the-wild person re-identification dataset especially of marathon runners is a challenging task. The variations in the scenarios such as camera viewpoints, resolution, occlusion, and illuminati…
Person Re-IdentificationregressionConstrained-size Tensorflow Models for YouTube-8M Video Understanding Challenge
This paper presents our 7th place solution to the second YouTube-8M video understanding competition which challenges participates to build a constrained-size model to classify millions of YouTube videos into thousands of…
Video Understanding5th Place Solution for YouTube-VOS Challenge 2022: Video Object Segmentation
Video object segmentation (VOS) has made significant progress with the rise of deep learning. However, there still exist some thorny problems, for example, similar objects are easily confused and tiny objects are difficu…
ObjectSegmentationSemantic SegmentationVideo Object Segmentation+2YouTube-VOS: Sequence-to-Sequence Video Object Segmentation
Learning long-term spatial-temporal features are critical for many video analysis tasks. However, existing video segmentation methods predominantly rely on static image segmentation techniques, and methods capturing temp…
Image SegmentationObjectOne-shot visual object segmentationOptical Flow Estimation+7