paper-with-me

홈 › Papers

Composing Driving Worlds through Disentangled Control for Adversarial Scenario Generation

2026-03-13 · Yifan Zhan, Zhengqing Chen, Qingjie Wang, Zhuo He, Muyao Niu, Xiaoyang Guo, Wei Yin, Weiqiang Ren, Qian Zhang, Yinqiang Zheng arxiv

A major challenge in autonomous driving is the "long tail" of safety-critical edge cases, which often emerge from unusual combinations of common traffic elements. Synthesizing these scenarios is crucial, yet current controllable generative models provide incomplete or entangled guidance, preventing the independent manipulation of scene structure, object identity, and ego actions. We introduce CompoSIA, a compositional driving video simulator that disentangles these traffic factors, enabling fine-grained control over diverse adversarial driving scenarios. To support controllable identity replacement of scene elements, we propose a noise-level identity injection, allowing pose-agnostic identity generation across diverse element poses, all from a single reference image. Furthermore, a hierarchical dual-branch action control mechanism is introduced to improve action controllability. Such disentangled control enables adversarial scenario synthesis-systematically combining safe elements into dangerous configurations that entangled generators cannot produce. Extensive comparisons demonstrate superior controllable generation quality over state-of-the-art baselines, with a 17% improvement in FVD for identity editing and reductions of 30% and 47% in rotation and translation errors for action control. Furthermore, downstream stress-testing reveals substantial planner failures: across editing modalities, the average collision rate of 3s increases by 173%.

📄 PDF Abstract BibTeX arXiv:2603.12864

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Similar Papers 제목 키워드 기반

WorldSplat: Gaussian-Centric Feed-Forward 4D Scene Generation for Autonomous Driving

2025-09-27 · Ziyue Zhu, Zhanqian Wu, Zhenxin Zhu, Lijun Zhou 외 arxiv

Recent advances in driving-scene generation and reconstruction have demonstrated significant potential for enhancing autonomous driving systems by producing scalable and controllable training data. Existing generation me…

Autonomous DrivingScene Generation

FactorPortrait: Controllable Portrait Animation via Disentangled Expression, Pose, and Viewpoint

2025-12-12 · Jiapeng Tang, Kai Li, Chengxiang Yin, Liuhao Ge 외 arxiv

We introduce FactorPortrait, a video diffusion method for controllable portrait animation that enables lifelike synthesis from disentangled control signals of facial expressions, head movement, and camera viewpoints. Giv…

Novel View Synthesis

Learning mixture of domain-specific experts via disentangled factors for autonomous driving

2022-06-28 · AAAI 2022 6 · Inhan Kim, Joonyeong Lee, Daijin Kim

Since human drivers only consider the driving-related factors that affect vehicle control depending on the situation, they can drive safely even in diverse driving environments. To mimic this behavior, we propose an auto…

Autonomous DrivingRepresentation Learning

HybridWorldSim: A Scalable and Controllable High-fidelity Simulator for Autonomous Driving

2025-11-27 · Qiang Li, Yingwenqi Jiang, Tuoxi Li, Duyu Chen 외 arxiv

Realistic and controllable simulation is critical for advancing end-to-end autonomous driving, yet existing approaches often struggle to support novel view synthesis under large viewpoint changes or to ensure geometric c…

Novel View SynthesisAutonomous Driving

DeX-Portrait: Disentangled and Expressive Portrait Animation via Explicit and Latent Motion Representations

2025-12-17 · Yuxiang Shi, Zhe Li, Yanwen Wang, Hao Zhu 외 arxiv

Portrait animation from a single source image and a driving video is a long-standing problem. Recent approaches tend to adopt diffusion-based image/video generation models for realistic and expressive animation. However,…

Video Generation