paper-with-me

홈 › Papers

Parallelized Spatiotemporal Binding

2024-02-26 · Gautam Singh, Yue Wang, Jiawei Yang, Boris Ivanovic, Sungjin Ahn, Marco Pavone, Tong Che

While modern best practices advocate for scalable architectures that support long-range interactions, object-centric models are yet to fully embrace these architectures. In particular, existing object-centric models for handling sequential inputs, due to their reliance on RNN-based implementation, show poor stability and capacity and are slow to train on long sequences. We introduce Parallelizable Spatiotemporal Binder or PSB, the first temporally-parallelizable slot learning architecture for sequential inputs. Unlike conventional RNN-based approaches, PSB produces object-centric representations, known as slots, for all time-steps in parallel. This is achieved by refining the initial slots across all time-steps through a fixed number of layers equipped with causal attention. By capitalizing on the parallelism induced by our architecture, the proposed model exhibits a significant boost in efficiency. In experiments, we test PSB extensively as an encoder within an auto-encoding framework paired with a wide variety of decoder options. Compared to the state-of-the-art, our architecture demonstrates stable training on longer sequences, achieves parallelization that results in a 60% increase in training speed, and yields performance that is on par with or better on unsupervised 2D and 3D object-centric scene decomposition and understanding.

📄 PDF Abstract BibTeX arXiv:2402.17077

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderObject

Similar Papers 제목 키워드 기반

Scalable Spatiotemporal Graph Neural Networks

2022-09-14 · Andrea Cini, Ivan Marisca, Filippo Maria Bianchi, Cesare Alippi

Neural forecasting of spatiotemporal time series drives both research and industrial innovation in several relevant application domains. Graph neural networks (GNNs) are often the core component of the forecasting archit…

Temporal SequencesTime SeriesTime Series Analysis

View while Moving: Efficient Video Recognition in Long-untrimmed Videos

2023-08-09 · Ye Tian, Mengyu Yang, Lanshan Zhang, Zhizhen Zhang 외

Recent adaptive methods for efficient video recognition mostly follow the two-stage paradigm of "preview-then-recognition" and have achieved great success on multiple video benchmarks. However, this two-stage paradigm in…

Video Recognition

Transient motion classification through turbid volumes via parallelized single-photon detection and deep contrastive embedding

2022-04-04 · Shiqi Xu, Wenhui Liu, Xi Yang, Joakim Jönsson 외

Fast noninvasive probing of spatially varying decorrelating events, such as cerebral blood flow beneath the human skull, is an essential task in various scientific and clinical settings. One of the primary optical techni…

Contrastive Learning

Exemplar-based Video Colorization with Long-term Spatiotemporal Dependency

2023-03-27 · Siqi Chen, Xueming Li, Xianlin Zhang, Mingdao Wang 외

Exemplar-based video colorization is an essential technique for applications like old movie restoration. Although recent methods perform well in still scenes or scenes with regular movement, they always lack robustness i…

Colorization

Decentralized Data Fusion and Active Sensing with Mobile Sensors for Modeling and Predicting Spatiotemporal Traffic Phenomena

2014-08-09 · Jie Chen, Kian Hsiang Low, Colin Keng-Yan Tan, Ali Oran 외

The problem of modeling and predicting spatiotemporal traffic phenomena over an urban road network is important to many traffic applications such as detecting and forecasting congestion hotspots. This paper presents a de…