paper-with-me

Papers

Spatial Sampling Network for Fast Scene Understanding

2019-05-22 · Davide Mazzini, Raimondo Schettini

We propose a network architecture to perform efficient scene understanding. This work presents three main novelties: the first is an Improved Guided Upsampling Module that can replace in toto the decoder part in common semantic segmentation networks. Our second contribution is the introduction of a new module based on spatial sampling to perform Instance Segmentation. It provides a very fast instance segmentation, needing only thresholding as post-processing step at inference time. Finally, we propose a novel efficient network design that includes the new modules and test it against different datasets for outdoor scene understanding. To our knowledge, our network is one of the themost efficient architectures for scene understanding published to date, furthermore being 8.6% more accurate than the fastest competitor on semantic segmentation and almost five times faster than the most efficient network for instance segmentation.

📄 PDF Abstract BibTeX arXiv:1905.09033

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderInstance SegmentationScene UnderstandingSegmentationSemantic Segmentation

Similar Papers 제목 키워드 기반

Generating Visual Spatial Description via Holistic 3D Scene Understanding

2023-05-19 · Yu Zhao, Hao Fei, Wei Ji, Jianguo Wei 외

Visual spatial description (VSD) aims to generate texts that describe the spatial relations of the given objects within images. Existing VSD work merely models the 2D geometrical vision features, thus inevitably falling …

Scene UnderstandingText Generation

Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding

2024-11-30 · CVPR 2025 1 · Duo Zheng, Shijia Huang, LiWei Wang

The rapid advancement of Multimodal Large Language Models (MLLMs) has significantly impacted various multimodal tasks. However, these models face challenges in tasks that require spatial understanding within 3D environme…

3D Question Answering (3D-QA)PositionScene Understanding

Sceniris: A Fast Procedural Scene Generation Framework

2025-12-18 · Jinghuan Shang, Harsh Patel, Ran Gong, Karl Schmeckpeper arxiv

Synthetic 3D scenes are essential for developing Physical AI and generative models. Existing procedural generation methods often have low output throughput, creating a significant bottleneck in scaling up dataset creatio…

Scene Generation

DreamScene: 3D Gaussian-based End-to-end Text-to-3D Scene Generation

2025-07-18 · Haoran Li, Yuli Tian, Kun Lan, Yong Liao 외 arxiv

Generating 3D scenes from natural language holds great promise for applications in gaming, film, and design. However, existing methods struggle with automation, 3D consistency, and fine-grained control. We present DreamS…

Scene Generation

Understanding Dynamic Scenes in Ego Centric 4D Point Clouds

2025-08-10 · Junsheng Huang, Shengyu Hao, Bocheng Hu, Hongwei Wang 외 arxiv

Understanding dynamic 4D scenes from an egocentric perspective-modeling changes in 3D spatial structure over time-is crucial for human-machine interaction, autonomous navigation, and embodied intelligence. While existing…

Trajectory PredictionScene UnderstandingPoint Clouds