paper-with-me

홈 › Papers

Memory Based Video Scene Parsing

2021-09-01 · Zhenchao Jin, Dongdong Yu, Kai Su, Zehuan Yuan, Changhu Wang

Video scene parsing is a long-standing challenging task in computer vision, aiming to assign pre-defined semantic labels to pixels of all frames in a given video. Compared with image semantic segmentation, this task pays more attention on studying how to adopt the temporal information to obtain higher predictive accuracy. In this report, we introduce our solution for the 1st Video Scene Parsing in the Wild Challenge, which achieves a mIoU of 57.44 and obtained the 2nd place (our team name is CharlesBLWX).

📄 PDF Abstract BibTeX arXiv:2109.00373

Code (0)

등록된 구현이 없습니다.

Tasks

Scene ParsingSemantic Segmentation

Similar Papers 제목 키워드 기반

Video Scene Parsing with Predictive Feature Learning

2016-12-01 · ICCV 2017 10 · Xiaojie Jin, Xin Li, Huaxin Xiao, Xiaohui Shen 외

In this work, we address the challenging video scene parsing problem by developing effective representation learning methods given limited parsing annotations. In particular, we contribute two novel methods that constitu…

Representation LearningScene Parsing

VSPW: A Large-scale Dataset for Video Scene Parsing in the Wild

2021-06-19 · CVPR 2021 1 · Jiaxu Miao, Yunchao Wei, Yu Wu, Chen Liang 외

In this paper, we present a new dataset with the target of advancing the scene parsing task from images to videos. Our dataset aims to perform Video Scene Parsing in the Wild (VSPW), which covers a wide range of real…

4kScene Parsing

Semantic Segmentation on VSPW Dataset through Aggregation of Transformer Models

2021-09-03 · Zixuan Chen, Junhong Zou, Xiaotao Wang

Semantic segmentation is an important task in computer vision, from which some important usage scenarios are derived, such as autonomous driving, scene parsing, etc. Due to the emphasis on the task of video semantic segm…

Autonomous DrivingScene ParsingSegmentationSemantic Segmentation+1

Discourse Parsing in Videos: A Multi-modal Appraoch

2019-03-06 · Arjun R. Akula, Song-Chun Zhu

Text-level discourse parsing aims to unmask how two sentences in the text are related to each other. We propose the task of Visual Discourse Parsing, which requires understanding discourse relations among scenes in a vid…

Discourse ParsingVisual DialogVisual Storytelling

Semi-supervised Video Semantic Segmentation Using Unreliable Pseudo Labels for PVUW2024

2024-06-02 · Biao Wu, Diankai Zhang, Si Gao, Chengjian Zheng 외

Pixel-level Scene Understanding is one of the fundamental problems in computer vision, which aims at recognizing object classes, masks and semantics of each pixel in the given image. Compared with image scene parsing, vi…

Scene ParsingScene UnderstandingSemantic SegmentationVideo Semantic Segmentation