Searching Scenes by Abstracting Things
In this paper we propose to represent a scene as an abstraction of 'things'. We start from 'things' as generated by modern object proposals, and we investigate their immediately observable properties: position, size, aspect ratio and color, and those only. Where the recent successes and excitement of the field lie in object identification, we represent the scene composition independent of object identities. We make three contributions in this work. First, we study simple observable properties of 'things', and call it things syntax. Second, we propose translating the things syntax in linguistic abstract statements and study their descriptive effect to retrieve scenes. Thirdly, we propose querying of scenes with abstract block illustrations and study their effectiveness to discriminate among different types of scenes. The benefit of abstract statements and block illustrations is that we generate them directly from the images, without any learning beforehand as in the standard attribute learning. Surprisingly, we show that even though we use the simplest of features from 'things' layout and no learning at all, we can still retrieve scenes reasonably well.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeDescriptiveObjectSimilar Papers 제목 키워드 기반
GS-RoadPatching: Inpainting Gaussians via 3D Searching and Placing for Driving Scenes
This paper presents GS-RoadPatching, an inpainting method for driving scene completion by referring to completely reconstructed regions, which are represented by 3D Gaussian Splatting (3DGS). Unlike existing 3DGS inpaint…
Multi-scale Motion-Aware Module for Video Action Recognition
Due to the lengthy computing time for optical flow, recent works have proposed to use the correlation operation as an alternative approach to extracting motion features. Although using correlation operations shows signi…
Action RecognitionGPUOptical Flow EstimationTemporal Action LocalizationSearching Neural Architectures for Sensor Nodes on IoT Gateways
This paper presents an automatic method for the design of Neural Networks (NNs) at the edge, enabling Machine Learning (ML) access even in privacy-sensitive Internet of Things (IoT) applications. The proposed method runs…
Fault DiagnosisDQFormer: Towards Unified LiDAR Panoptic Segmentation with Decoupled Queries
LiDAR panoptic segmentation, which jointly performs instance and semantic segmentation for things and stuff classes, plays a fundamental role in LiDAR perception tasks. While most existing methods explicitly separate the…
DecoderInstance SegmentationPanoptic SegmentationSegmentation+1CommonScenes: Generating Commonsense 3D Indoor Scenes with Scene Graph Diffusion
Controllable scene synthesis aims to create interactive environments for various industrial use cases. Scene graphs provide a highly suitable interface to facilitate these applications by abstracting the scene context in…
DiversityObjectScene Generation