paper-with-me

Papers

HS3: Learning with Proper Task Complexity in Hierarchically Supervised Semantic Segmentation

2021-11-03 · Shubhankar Borse, Hong Cai, Yizhe Zhang, Fatih Porikli

While deeply supervised networks are common in recent literature, they typically impose the same learning objective on all transitional layers despite their varying representation powers. In this paper, we propose Hierarchically Supervised Semantic Segmentation (HS3), a training scheme that supervises intermediate layers in a segmentation network to learn meaningful representations by varying task complexity. To enforce a consistent performance vs. complexity trade-off throughout the network, we derive various sets of class clusters to supervise each transitional layer of the network. Furthermore, we devise a fusion framework, HS3-Fuse, to aggregate the hierarchical features generated by these layers, which can provide rich semantic contexts and further enhance the final segmentation. Extensive experiments show that our proposed HS3 scheme considerably outperforms vanilla deep supervision with no added inference cost. Our proposed HS3-Fuse framework further improves segmentation predictions and achieves state-of-the-art results on two large segmentation benchmarks: NYUD-v2 and Cityscapes.

📄 PDF Abstract BibTeX arXiv:2111.02333

Code (0)

등록된 구현이 없습니다.

Tasks

SegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

Unsupervised semantic discovery through visual patterns detection

2021-02-24 · Francesco Pelosin, Andrea Gasparetto, Andrea Albarelli, Andrea Torsello

We propose a new fast fully unsupervised method to discover semantic patterns. Our algorithm is able to hierarchically find visual categories and produce a segmentation mask where previous methods fail. Through the model…

ClusteringSuperpixels

Hierarchically Decoupled Spatial-Temporal Contrast for Self-supervised Video Representation Learning

2020-11-23 · Zehua Zhang, David Crandall

We present a novel technique for self-supervised video representation learning by: (a) decoupling the learning objective into two contrastive subtasks respectively emphasizing spatial and temporal features, and (b) perfo…

Action RecognitionContrastive LearningRepresentation Learning

Generative Adversarial Image Synthesis with Decision Tree Latent Controller

2018-05-27 · CVPR 2018 6 · Takuhiro Kaneko, Kaoru Hiramatsu, Kunio Kashino

This paper proposes the decision tree latent controller generative adversarial network (DTLC-GAN), an extension of a GAN that can learn hierarchically interpretable representations without relying on detailed supervision…

Generative Adversarial NetworkImage GenerationImage RetrievalRepresentation Learning+1

Formal Concept Lattices are Good Semantic Scaffolds for Concept-Based Learning

2026-06-03 · Deepika SN Vemuri, Sayanta Adhikari, Ankit Saha, Krishn Vishwas Kher 외 arxiv

Learning semantics is essential for deep learning models to be interpretable and better aligned with human reasoning. Concept-based models approach this by representing classes through meaningful semantic abstractions, b…

Weakly-Supervised Action Localization by Hierarchically-structured Latent Attention Modeling

2023-08-19 · ICCV 2023 1 · Guiqin Wang, Peng Zhao, Cong Zhao, Shusen Yang 외

Weakly-supervised action localization aims to recognize and localize action instancese in untrimmed videos with only video-level labels. Most existing models rely on multiple instance learning(MIL), where the predictions…

Action LocalizationMultiple Instance LearningWeakly Supervised Action Localization