paper-with-me

홈 › Papers

DriveWeaver: Point-Conditioned Video Inpainting for Controllable Vehicle Insertion in Autonomous Driving Simulation

2026-06-30 · Junzhe Jiang, Zipei Ma, Zijie Pan, Li Zhang arxiv

A pivotal step in autonomous driving simulation involves inserting foreground vehicles with predefined trajectories into simulated scenes. This process enhances scene diversity and facilitates the creation of various corner cases for testing and improving autonomous driving models. However, existing methods often rely on pre-reconstructed 3D assets, which frequently lead to lighting inconsistencies between the inserted foreground and the background. Moreover, the reliance on limited, manually-curated 3D assets hinders large-scale deployment. To address these challenges, we propose DriveWeaver, a novel framework for controllable vehicle insertion in autonomous driving simulation. Specifically, for a masked target insertion area, DriveWeaver performs video inpainting conditioned on vehicle point clouds to generate high-quality, temporally consistent vehicles. This video-inpainting-based approach ensures seamless blending between the foreground and background, while the readily available point cloud conditions enable superior generalization. To support long-term generation, we further design a global-to-local hierarchical inpainting strategy, ensuring the consistent identity and appearance of the inserted vehicles. Meanwhile, we extract explicit 3D Gaussian representations of the inserted vehicles through an urban reconstruction pipeline to enable real-time rendering for autonomous driving simulation. Extensive experiments across diverse datasets demonstrate that our method outperforms existing baselines in visual realism and geometric consistency, providing a robust tool for scalable autonomous driving scene augmentation.

📄 PDF Abstract BibTeX arXiv:2606.31918

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingVideo InpaintingPoint Clouds

Similar Papers 제목 키워드 기반

Replace Anyone in Videos

2024-09-30 · Xiang Wang, Shiwei Zhang, Haonan Qiu, Ruihang Chu 외

The field of controllable human-centric video generation has witnessed remarkable progress, particularly with the advent of diffusion models. However, achieving precise and localized control over human motion in videos, …

Video GenerationVideo Inpainting

Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors

2026-03-19 · Jiatong Xia, Zicheng Duan, Anton van den Hengel, Lingqiao Liu arxiv

Recent progress in 3D generation has been driven largely by models conditioned on images or text, while readily available 3D priors are still underused. In many real-world scenarios, the visible-region point cloud are ea…

Scene Generation3D GenerationPoint Clouds

A Temporally-Aware Interpolation Network for Video Frame Inpainting

2018-03-20 · Ximeng Sun, Ryan Szeto, Jason J. Corso

We propose the first deep learning solution to video frame inpainting, a challenging instance of the general video inpainting problem with applications in video editing, manipulation, and forensics. Our task is less ambi…

DecoderPredictionVideo EditingVideo Inpainting+1

InpaintDPO: Mitigating Spatial Relationship Hallucinations in Foreground-conditioned Inpainting via Diverse Preference Optimization

2025-12-16 · Qirui Li, Yizhe Tang, Ran Yi, Guangben Lu 외 arxiv

Foreground-conditioned inpainting, which aims at generating a harmonious background for a given foreground subject based on the text prompt, is an important subfield in controllable image generation. A common challenge i…

Image Generation

See4D: Pose-Free 4D Generation via Auto-Regressive Video Inpainting

2025-10-30 · Dongyue Lu, Ao Liang, Tianxin Huang, Xiao Fu 외 arxiv

Immersive applications call for synthesizing spatiotemporal 4D content from casual videos without costly 3D supervision. Existing video-to-4D methods typically rely on manually annotated camera poses, which are labor-int…

Trajectory PredictionVideo GenerationVideo Inpainting