paper-with-me

Papers

Controllable Diverse Sampling for Diffusion Based Motion Behavior Forecasting

2024-02-06 · Yiming Xu, Hao Cheng, Monika Sester

In autonomous driving tasks, trajectory prediction in complex traffic environments requires adherence to real-world context conditions and behavior multimodalities. Existing methods predominantly rely on prior assumptions or generative models trained on curated data to learn road agents' stochastic behavior bounded by scene constraints. However, they often face mode averaging issues due to data imbalance and simplistic priors, and could even suffer from mode collapse due to unstable training and single ground truth supervision. These issues lead the existing methods to a loss of predictive diversity and adherence to the scene constraints. To address these challenges, we introduce a novel trajectory generator named Controllable Diffusion Trajectory (CDT), which integrates map information and social interactions into a Transformer-based conditional denoising diffusion model to guide the prediction of future trajectories. To ensure multimodality, we incorporate behavioral tokens to direct the trajectory's modes, such as going straight, turning right or left. Moreover, we incorporate the predicted endpoints as an alternative behavioral token into the CDT model to facilitate the prediction of accurate trajectories. Extensive experiments on the Argoverse 2 benchmark demonstrate that CDT excels in generating diverse and scene-compliant trajectories in complex urban settings.

📄 PDF Abstract BibTeX arXiv:2402.03981

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous DrivingDenoisingDiversityPredictionTrajectory Prediction

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Proposal-Conditioned Latent Diffusion for Closed-Loop Traffic Scenario Generation

2026-06-25 · Shubham Vaijanath Phoolari, Aleyna Kara, Christoph Lauer, Steven Peters arxiv

Closed-loop traffic simulation remains challenging because it must generate interactive multi-agent behaviors that are scene-consistent and controllable throughout rollout. Prior diffusion-based approaches achieve strong…

Controllable Motion Generation via Diffusion Modal Coupling

2025-03-04 · Luobin Wang, Hongzhan Yu, Chenning Yu, Sicun Gao 외

Diffusion models have recently gained significant attention in robotics due to their ability to generate multi-modal distributions of system states and behaviors. However, a key challenge remains: ensuring precise contro…

DenoisingDiversityMotion GenerationMotion Planning+2

Controllable Expressive 3D Facial Animation via Diffusion in a Unified Multimodal Space

2025-04-14 · Kangwei Liu, Junwu Liu, Xiaowei Yi, Jinlin Guo 외

Audio-driven emotional 3D facial animation encounters two significant challenges: (1) reliance on single-modal control signals (videos, text, or emotion labels) without leveraging their complementary strengths for compre…

Contrastive LearningDiversity

EmoDiff: Intensity Controllable Emotional Text-to-Speech with Soft-Label Guidance

2022-11-17 · Yiwei Guo, Chenpeng Du, Xie Chen, Kai Yu

Although current neural text-to-speech (TTS) models are able to generate high-quality speech, intensity controllable emotional TTS is still a challenging task. Most existing methods need external optimizations for intens…

Denoisingtext-to-speechText to Speech

Controllable Motion Synthesis and Reconstruction with Autoregressive Diffusion Models

2023-04-03 · Wenjie Yin, Ruibo Tu, Hang Yin, Danica Kragic 외

Data-driven and controllable human motion synthesis and prediction are active research areas with various applications in interactive media and social robotics. Challenges remain in these fields for generating diverse mo…

DecoderMotion Synthesis