paper-with-me

홈 › Papers

StreetForward: Perceiving Dynamic Street with Feedforward Causal Attention

2026-03-20 · Zhongrui Yu, Zhao Wang, Yijia Xie, Yida Wang, Xueyang Zhang, Yifei Zhan, Kun Zhan arxiv

Feedforward reconstruction is crucial for autonomous driving applications, where rapid scene reconstruction enables efficient utilization of large-scale driving datasets in closed-loop simulation and other downstream tasks, eliminating the need for time-consuming per-scene optimization. We present StreetForward, a pose-free and tracker-free feedforward framework for dynamic street reconstruction. Building upon the alternating attention mechanism from Visual Geometry Grounded Transformer (VGGT), we propose a simple yet effective temporal mask attention module that captures dynamic motion information from image sequences and produces motion-aware latent representations. Static content and dynamic instances are represented uniformly with 3D Gaussian Splatting, and are optimized jointly by cross-frame rendering with spatio-temporal consistency, allowing the model to infer per-pixel velocities and produce high-fidelity novel views at new poses and times. We train and evaluate our model on the Waymo Open Dataset, demonstrating superior performance on novel view synthesis and depth estimation compared to existing methods. Furthermore, zero-shot inference on CARLA and other datasets validates the generalization capability of our approach. More visualizations are available on our project page: https://streetforward.github.io.

📄 PDF Abstract BibTeX arXiv:2603.19552

Code (0)

등록된 구현이 없습니다.

Tasks

Novel View SynthesisAutonomous DrivingDepth Estimation

Similar Papers 제목 키워드 기반

Pixel-wise Segmentation of Street with Neural Networks

2015-11-02 · Sebastian Bittel, Vitali Kaiser, Marvin Teichmann, Martin Thoma

Pixel-wise street segmentation of photographs taken from a drivers perspective is important for self-driving cars and can also support other object recognition tasks. A framework called SST was developed to examine the a…

Object RecognitionregressionSelf-Driving Cars

Learning nonlinear feedforward: a Gaussian Process Approach Applied to a Printer with Friction

2021-12-07 · Max van Meer, Maurice Poot, Jim Portegies, Tom Oomen

Feedforward control is essential to achieving good tracking performance in positioning systems. The aim of this paper is to develop an identification strategy for inverse models of systems with nonlinear dynamics of unkn…

Friction

Semantic4Safety: Causal Insights from Zero-shot Street View Imagery Segmentation for Urban Road Safety

2025-10-17 · Huan Chen, Ting Han, Siyu Chen, Zhihao Guo 외 arxiv

Street-view imagery (SVI) offers a fine-grained lens on traffic risk, yet two fundamental challenges persist: (1) how to construct street-level indicators that capture accident-related features, and (2) how to quantify t…

Zero-Shot Semantic SegmentationCausal Inference

Adjustment formulas for learning causal steady-state models from closed-loop operational data

2022-11-10 · Kristian Løvland, Bjarne Grimstad, Lars Struen Imsland

Steady-state models which have been learned from historical operational data may be unfit for model-based optimization unless correlations in the training data which are introduced by control are accounted for. Using rec…

CLEVRER: CoLlision Events for Video REpresentation and Reasoning

2019-10-03 · ICLR 2020 1 · Kexin Yi, Chuang Gan, Yunzhu Li, Pushmeet Kohli 외

The ability to reason about temporal and causal events from videos lies at the core of human intelligence. Most video reasoning benchmarks, however, focus on pattern recognition from complex visual and language input, in…

counterfactualDescriptiveDiagnosticVisual Reasoning