paper-with-me

Papers

STD2P: RGBD Semantic Segmentation Using Spatio-Temporal Data-Driven Pooling

2016-04-08 · CVPR 2017 7 · Yang He, Wei-Chen Chiu, Margret Keuper, Mario Fritz

We propose a novel superpixel-based multi-view convolutional neural network for semantic image segmentation. The proposed network produces a high quality segmentation of a single image by leveraging information from additional views of the same scene. Particularly in indoor videos such as captured by robotic platforms or handheld and bodyworn RGBD cameras, nearby video frames provide diverse viewpoints and additional context of objects and scenes. To leverage such information, we first compute region correspondences by optical flow and image boundary-based superpixels. Given these region correspondences, we propose a novel spatio-temporal pooling layer to aggregate information over space and time. We evaluate our approach on the NYU--Depth--V2 and the SUN3D datasets and compare it to various state-of-the-art single-view and multi-view approaches. Besides a general improvement over the state-of-the-art, we also show the benefits of making use of unlabeled frames during training for multi-view as well as single-view prediction.

📄 PDF Abstract BibTeX arXiv:1604.02388

Code (1)

SSAW14/STD2P 공식 구현

Tasks

Image SegmentationOptical Flow EstimationRGBD Semantic SegmentationSegmentationSemantic SegmentationSuperpixels

Similar Papers 제목 키워드 기반

Attention-based Dual Supervised Decoder for RGBD Semantic Segmentation

2022-01-05 · Yang Zhang, Yang Yang, Chenyun Xiong, Guodong Sun 외

Encoder-decoder models have been widely used in RGBD semantic segmentation, and most of them are designed via a two-stream network. In general, jointly reasoning the color and geometric information from RGBD is beneficia…

DecoderRGBD Semantic SegmentationSegmentationSemantic Segmentation

STFCN: Spatio-Temporal FCN for Semantic Video Segmentation

2016-08-21 · Mohsen Fayyaz, Mohammad Hajizadeh Saffar, Mohammad Sabokrou, Mahmood Fathy 외

This paper presents a novel method to involve both spatial and temporal features for semantic video segmentation. Current work on convolutional neural networks(CNNs) has shown that CNNs provide advanced spatial features …

SegmentationSemantic SegmentationVideo SegmentationVideo Semantic Segmentation

Rescan: Inductive Instance Segmentation for Indoor RGBD Scans

2019-09-25 · ICCV 2019 10 · Maciej Halber, Yifei Shi, Kai Xu, Thomas Funkhouser

In depth-sensing applications ranging from home robotics to AR/VR, it will be common to acquire 3D scans of interior spaces repeatedly at sparse time intervals (e.g., as part of regular daily use). We propose an algorith…

Instance SegmentationSegmentationSemantic Segmentation

A Spatiotemporal Correspondence Approach to Unsupervised LiDAR Segmentation with Traffic Applications

2023-08-23 · Xiao Li, Pan He, Aotian Wu, Sanjay Ranka 외

We address the problem of unsupervised semantic segmentation of outdoor LiDAR point clouds in diverse traffic scenarios. The key idea is to leverage the spatiotemporal nature of a dynamic point cloud sequence and introdu…

ClusteringPseudo LabelRepresentation LearningSegmentation+2

A spatio-temporal network for video semantic segmentation in surgical videos

2023-06-19 · Maria Grammatikopoulou, Ricardo Sanchez-Matilla, Felix Bragman, David Owen 외

Semantic segmentation in surgical videos has applications in intra-operative guidance, post-operative analytics and surgical education. Segmentation models need to provide accurate and consistent predictions since tempor…

DecoderSegmentationSemantic SegmentationVideo Semantic Segmentation