paper-with-me

홈 › Papers

Learning to Refine Object Segments

2016-03-29 · Pedro O. Pinheiro, Tsung-Yi Lin, Ronan Collobert, Piotr Dollàr

Object segmentation requires both object-level information and low-level pixel data. This presents a challenge for feedforward networks: lower layers in convolutional nets capture rich spatial information, while upper layers encode object-level knowledge but are invariant to factors such as pose and appearance. In this work we propose to augment feedforward nets for object segmentation with a novel top-down refinement approach. The resulting bottom-up/top-down architecture is capable of efficiently generating high-fidelity object masks. Similarly to skip connections, our approach leverages features at all layers of the net. Unlike skip connections, our approach does not attempt to output independent predictions at each layer. Instead, we first output a coarse `mask encoding' in a feedforward pass, then refine this mask encoding in a top-down pass utilizing features at successively lower layers. The approach is simple, fast, and effective. Building on the recent DeepMask network for generating object proposals, we show accuracy improvements of 10-20% in average recall for various setups. Additionally, by optimizing the overall network architecture, our approach, which we call SharpMask, is 50% faster than the original DeepMask network (under .8s per image).

📄 PDF Abstract BibTeX arXiv:1603.08695

Code (2)

aby2s/sharpmask tf
facebookresearch/deepmask torch

Tasks

ObjectSemantic Segmentation

Methods 이 논문이 사용한 방법론

Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Ethereum Customer Service Number +1-833-534-1729 설명 없음
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
1x1 Convolution A 1 x 1 Convolution is a convolution with some special properties in that it can be used for dimensionality reduction,…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…

Similar Papers 제목 키워드 기반

Indoor Frame Recovery from Refined Line Segments

2017-04-30 · Luanzheng Guo, Jun Chu

An important yet challenging problem in understanding indoor scene is recovering indoor frame structure from a monocular image. It is more difficult when occlusions and illumination vary, and object boundaries are weak. …

Joint Multi-Person Pose Estimation and Semantic Part Segmentation

2017-08-10 · CVPR 2017 7 · Fangting Xia, Peng Wang, Xianjie Chen, Alan Yuille

Human pose estimation and semantic part segmentation are two complementary tasks in computer vision. In this paper, we propose to solve the two tasks jointly for natural multi-person images, in which the estimated pose p…

Human DetectionMulti-Person Pose EstimationPose Estimation

Multiscale Vision Transformer With Deep Clustering-Guided Refinement for Weakly Supervised Object Localization

2023-12-15 · David Kim, Sinhae Cha, Byeongkeun Kang

This work addresses the task of weakly-supervised object localization. The goal is to learn object localization using only image-level class labels, which are much easier to obtain compared to bounding box annotations. T…

ClusteringDeep ClusteringObjectObject Localization+1

Occlusion Resistant Object Rotation Regression from Point Cloud Segments

2018-08-16 · Ge Gao, Mikko Lauri, Jianwei Zhang, Simone Frintrop

Rotation estimation of known rigid objects is important for robotic applications such as dexterous manipulation. Most existing methods for rotation estimation use intermediate representations such as templates, global or…

Objectregression

TIBR4D: Tracing-Guided Iterative Boundary Refinement for Efficient 4D Gaussian Segmentation

2026-02-09 · He Wu, Xia Yan, Yanghui Xu, Liegang Xia 외 arxiv

Object-level segmentation in dynamic 4D Gaussian scenes remains challenging due to complex motion, occlusions, and ambiguous boundaries. In this paper, we present an efficient learning-free 4D Gaussian segmentation frame…

Video SegmentationPoint Clouds