paper-with-me

홈 › Papers

FLARE: Learning Future-Aware Latent Representations from Vision-Language Models for Autonomous Driving

2026-01-09 · Chengen Xie, Chonghao Sima, Tianyu Li, Bin Sun, Junjie Wu, Zhihui Hao, Hongyang Li arxiv

While Vision-Language Models (VLMs) offer rich world knowledge for end-to-end autonomous driving, current approaches heavily rely on labor-intensive language annotations (e.g., VQA) to bridge perception and control. This paradigm suffers from a fundamental mismatch between discrete linguistic tokens and continuous driving trajectories, often leading to suboptimal control policies and inefficient utilization of pre-trained knowledge. To address these challenges, we propose FLARE (Future-aware LAtent REpresentation), a novel framework that activates the visual-semantic capabilities of pre-trained VLMs without requiring language supervision. Instead of aligning with text, we introduce a self-supervised future feature prediction objective. This mechanism compels the model to anticipate scene dynamics and ego-motion directly in the latent space, enabling the learning of robust driving representations from large-scale unlabeled trajectory data. Furthermore, we integrate Group Relative Policy Optimization (GRPO) into the planning process to refine decision-making quality. Extensive experiments on the NAVSIM benchmark demonstrate that FLARE achieves state-of-the-art performance, validating the effectiveness of leveraging VLM knowledge via predictive self-supervision rather than explicit language generation.

📄 PDF Abstract BibTeX arXiv:2601.05611

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous Driving

Similar Papers 제목 키워드 기반

FLARE: Robot Learning with Implicit World Modeling

2025-05-21 · Ruijie Zheng, Jing Wang, Scott Reed, Johan Bjorck 외

We introduce $\textbf{F}$uture $\textbf{LA}$tent $\textbf{RE}$presentation Alignment ($\textbf{FLARE}$), a novel framework that integrates predictive latent world modeling into robot policy learning. By aligning features…

Imitation LearningVision-Language-Action

Semi-LAR: Semi-supervised Contrastive Learning with Linear Attention for Removal of Nighttime Flares

2026-05-18 · Xiyu Zhu, Wei Wang, Kui Jiang, Zhengguo Li arxiv

Lens flare removal is challenging due to the large spatial extent of flare artifacts and their entanglement with scene structures, while existing methods heavily rely on large-scale paired data. We propose a semi-supervi…

Contrastive LearningFlare Removal

Difflare: Removing Image Lens Flare with Latent Diffusion Model

2024-07-20 · Tianwen Zhou, Qihao Duan, Zitong Yu

The recovery of high-quality images from images corrupted by lens flare presents a significant challenge in low-level vision. Contemporary deep learning methods frequently entail training a lens flare removing model from…

Flare Removal

FLARe: Forecasting by Learning Anticipated Representations

2019-04-17 · Surya Teja Devarakonda, Joie Yeahuay Wu, Yi Ren Fung, Madalina Fiterau

Computational models that forecast the progression of Alzheimer's disease at the patient level are extremely useful tools for identifying high risk cohorts for early intervention and treatment planning. The state-of-the-…

Flare7K++: Mixing Synthetic and Real Datasets for Nighttime Flare Removal and Beyond

2023-06-07 · Yuekun Dai, Chongyi Li, Shangchen Zhou, Ruicheng Feng 외

Artificial lights commonly leave strong lens flare artifacts on the images captured at night, degrading both the visual quality and performance of vision algorithms. Existing flare removal approaches mainly focus on remo…

Flare Removal