paper-with-me

Papers

Compositional Factorization of Visual Scenes with Convolutional Sparse Coding and Resonator Networks

2024-04-29 · Christopher J. Kymn, Sonia Mazelet, Annabel Ng, Denis Kleyko, Bruno A. Olshausen

We propose a system for visual scene analysis and recognition based on encoding the sparse, latent feature-representation of an image into a high-dimensional vector that is subsequently factorized to parse scene content. The sparse feature representation is learned from image statistics via convolutional sparse coding, while scene parsing is performed by a resonator network. The integration of sparse coding with the resonator network increases the capacity of distributed representations and reduces collisions in the combinatorial search space during factorization. We find that for this problem the resonator network is capable of fast and accurate vector factorization, and we develop a confidence-based metric that assists in tracking the convergence of the resonator network.

📄 PDF Abstract BibTeX arXiv:2404.19126

Code (0)

등록된 구현이 없습니다.

Tasks

Scene Parsing

Similar Papers 제목 키워드 기반

Evaluating Compositional Scene Understanding in Multimodal Generative Models

2025-03-29 · Shuhao Fu, Andrew Jun Lee, Anna Wang, Ida Momennejad 외

The visual world is fundamentally compositional. Visual scenes are defined by the composition of objects and their relations. Hence, it is essential for computer vision systems to reflect and exploit this compositionalit…

Scene Understanding

Towards Generalizable Robotic Data Flywheel: High-Dimensional Factorization and Composition

2026-03-26 · Yuyang Xiao, Yifei Zhou, Haoran Wang, Wenxuan Ou 외 arxiv

The lack of sufficiently diverse data, coupled with limited data efficiency, remains a major bottleneck for generalist robotic models, yet systematic strategies for collecting and curating such data are not fully explore…

Sparse Factorization Layers for Neural Networks with Limited Supervision

2016-12-14 · Parker Koch, Jason J. Corso

Whereas CNNs have demonstrated immense progress in many vision problems, they suffer from a dependence on monumental amounts of labeled training data. On the other hand, dictionary learning does not scale to the size of …

DenoisingDictionary Learning

Edge Data Based Trailer Inception Probabilistic Matrix Factorization for Context-Aware Movie Recommendation

2022-02-16 · Honglong Chen, Zhe Li, Zhu Wang, Zhichen Ni 외

The rapid growth of edge data generated by mobile devices and applications deployed at the edge of the network has exacerbated the problem of information overload. As an effective way to alleviate information overload, r…

Movie RecommendationRecommendation Systems

A Benchmark for Compositional Visual Reasoning

2022-06-11 · Aimen Zerroug, Mohit Vaishnav, Julien Colin, Sebastian Musslick 외

A fundamental component of human vision is our ability to parse complex visual scenes and judge the relations between their constituent objects. AI benchmarks for visual reasoning have driven rapid progress in recent yea…

Visual Reasoning