paper-with-me

Papers

UniNet: A Unified Scene Understanding Network and Exploring Multi-Task Relationships through the Lens of Adversarial Attacks

2021-08-10 · Naresh Kumar Gurulingan, Elahe Arani, Bahram Zonooz

Scene understanding is crucial for autonomous systems which intend to operate in the real world. Single task vision networks extract information only based on some aspects of the scene. In multi-task learning (MTL), on the other hand, these single tasks are jointly learned, thereby providing an opportunity for tasks to share information and obtain a more comprehensive understanding. To this end, we develop UniNet, a unified scene understanding network that accurately and efficiently infers vital vision tasks including object detection, semantic segmentation, instance segmentation, monocular depth estimation, and monocular instance depth prediction. As these tasks look at different semantic and geometric information, they can either complement or conflict with each other. Therefore, understanding inter-task relationships can provide useful cues to enable complementary information sharing. We evaluate the task relationships in UniNet through the lens of adversarial attacks based on the notion that they can exploit learned biases and task interactions in the neural network. Extensive experiments on the Cityscapes dataset, using untargeted and targeted attacks reveal that semantic tasks strongly interact amongst themselves, and the same holds for geometric tasks. Additionally, we show that the relationship between semantic and geometric tasks is asymmetric and their interaction becomes weaker as we move towards higher-level representations.

📄 PDF Abstract BibTeX arXiv:2108.04584

Code (1)

NeurAI-Lab/UniNet 공식 구현 pytorch

Tasks

Depth EstimationDepth PredictionInstance SegmentationMonocular Depth EstimationMulti-Task Learningobject-detectionObject DetectionScene UnderstandingSemantic Segmentation

Similar Papers 제목 키워드 기반

UniNet: A Unified Multi-granular Traffic Modeling Framework for Network Security

2025-03-06 · Binghui Wu, Dinil Mon Divakaran, Mohan Gurusamy

As modern networks grow increasingly complex--driven by diverse devices, encrypted protocols, and evolving threats--network traffic analysis has become critically important. Existing machine learning models often rely on…

Anomaly DetectionIoT Device Identification

UniNet: Unified Architecture Search with Convolution, Transformer, and MLP

2022-07-12 · Jihao Liu, Xin Huang, Guanglu Song, Hongsheng Li 외

Recently, transformer and multi-layer perceptron (MLP) architectures have achieved impressive results on various vision tasks. However, how to effectively combine those operators to form high-performance hybrid visual ar…

Image ClassificationNeural Architecture Search

UniNet: A Contrastive Learning-guided Unified Framework with Feature Selection for Anomaly Detection

2025-02-28 · CVPR 2025 1 · Shun Wei, Jielin Jiang, Xiaolong Xu

Anomaly detection (AD) is a crucial visual task aimed at recognizing abnormal pattern within samples. However, most existing AD methods suffer from limited generalizability, as they are primarily designed for domain-spec…

Anomaly DetectionImage ClassificationMedical Image SegmentationMulti-class Anomaly Detection+1

UniNet: Unified Architecture Search with Convolution, Transformer, and MLP

2021-10-08 · Jihao Liu, Hongsheng Li, Guanglu Song, Xin Huang 외

Recently, transformer and multi-layer perceptron (MLP) architectures have achieved impressive results on various vision tasks. A few works investigated manually combining those operators to design visual network architec…

Image Classificationobject-detectionObject DetectionSemantic Segmentation

Omni-View: Unlocking How Generation Facilitates Understanding in Unified 3D Model based on Multiview images

2025-11-10 · JiaKui Hu, Shanshan Zhao, Qing-Guo Chen, Xuerui Qiu 외 arxiv

This paper presents Omni-View, which extends the unified multimodal understanding and generation to 3D scenes based on multiview images, exploring the principle that "generation facilitates understanding". Consisting of …

Novel View SynthesisScene UnderstandingScene Generation