Loci-Segmented: Improving Scene Segmentation Learning
Current slot-oriented approaches for compositional scene segmentation from images and videos rely on provided background information or slot assignments. We present a segmented location and identity tracking system, Loci-Segmented (Loci-s), which does not require either of this information. It learns to dynamically segment scenes into interpretable background and slot-based object encodings, separating rgb, mask, location, and depth information for each. The results reveal largely superior video decomposition performance in the MOVi datasets and in another established dataset collection targeting scene segmentation. The system's well-interpretable, compositional latent encodings may serve as a foundation model for downstream tasks.
Code (1)
Tasks
Scene SegmentationSegmentationSimilar Papers 제목 키워드 기반
ONeRF: Unsupervised 3D Object Segmentation from Multiple Views
We present ONeRF, a method that automatically segments and reconstructs object instances in 3D from multi-view RGB images without any additional manual annotations. The segmented 3D objects are represented using separate…
3D scene EditingObjectSemantic SegmentationLearning Segmented 3D Gaussians via Efficient Feature Unprojection for Zero-shot Neural Scene Segmentation
Zero-shot neural scene segmentation, which reconstructs 3D neural segmentation field without manual annotations, serves as an effective way for scene understanding. However, existing models, especially the efficient 3D G…
DecoderPanoptic SegmentationScene SegmentationScene Understanding+4Unsupervised Temporal Segmentation of Repetitive Human Actions Based on Kinematic Modeling and Frequency Analysis
In this paper, we propose a method for temporal segmentation of human repetitive actions based on frequency analysis of kinematic parameters, zero-velocity crossing detection, and adaptive k-means clustering. Since the h…
ClusteringSegmentationNeuromorphic Vision-based Motion Segmentation with Graph Transformer Neural Network
Moving object segmentation is critical to interpret scene dynamics for robotic navigation systems in challenging environments. Neuromorphic vision sensors are tailored for motion perception due to their asynchronous natu…
BenchmarkingMotion SegmentationSegmentationSemantic Segmentation3D Instance Segmentation of MVS Buildings
We present a novel 3D instance segmentation framework for Multi-View Stereo (MVS) buildings in urban scenes. Unlike existing works focusing on semantic segmentation of urban scenes, the emphasis of this work lies in dete…
3D Instance SegmentationInstance SegmentationSegmentationSemantic Segmentation