paper-with-me

홈 › Papers

Trajeglish: Traffic Modeling as Next-Token Prediction

2023-12-07 · Jonah Philion, Xue Bin Peng, Sanja Fidler

A longstanding challenge for self-driving development is simulating dynamic driving scenarios seeded from recorded driving logs. In pursuit of this functionality, we apply tools from discrete sequence modeling to model how vehicles, pedestrians and cyclists interact in driving scenarios. Using a simple data-driven tokenization scheme, we discretize trajectories to centimeter-level resolution using a small vocabulary. We then model the multi-agent sequence of discrete motion tokens with a GPT-like encoder-decoder that is autoregressive in time and takes into account intra-timestep interaction between agents. Scenarios sampled from our model exhibit state-of-the-art realism; our model tops the Waymo Sim Agents Benchmark, surpassing prior work along the realism meta metric by 3.3% and along the interaction metric by 9.9%. We ablate our modeling choices in full autonomy and partial autonomy settings, and show that the representations learned by our model can quickly be adapted to improve performance on nuScenes. We additionally evaluate the scalability of our model with respect to parameter count and dataset size, and use density estimates from our model to quantify the saliency of context length and intra-timestep interaction for the traffic modeling task.

📄 PDF Abstract BibTeX arXiv:2312.04535

Code (2)

horizonrobotics/gump pytorch
nvlabs/catk pytorch

Tasks

DecoderPrediction

Similar Papers 제목 키워드 기반

Next-Latent Prediction Transformers Learn Compact World Models

2025-11-08 · Jayden Teoh, Manan Tomar, Kwangjun Ahn, Edward S. Hu 외 arxiv

Transformers replace recurrence with a memory that grows with sequence length and self-attention that enables ad-hoc lookups over past tokens. Consequently, they lack an inherent incentive to compress history into compac…

Enhancing next token prediction based pre-training for jet foundation models

2025-12-03 · Joschka Birk, Anna Hallin, Gregor Kasieczka, Nikol Madzharova 외 arxiv

Next token prediction is an attractive pre-training task for jet foundation models, in that it is simulation free and enables excellent generative capabilities that can transfer across datasets. Here we study multiple im…

FlowAR: Scale-wise Autoregressive Image Generation Meets Flow Matching

2024-12-19 · Sucheng Ren, Qihang Yu, Ju He, Xiaohui Shen 외

Autoregressive (AR) modeling has achieved remarkable success in natural language processing by enabling models to generate text with coherence and contextual understanding through next token prediction. Recently, in imag…

Image GenerationPrediction

Memory-based Language Models: An Efficient, Explainable, and Eco-friendly Approach to Large Language Modeling

2025-10-25 · Antal van den Bosch, Ainhoa Risco Patón, Teun Buijse, Peter Berck 외 arxiv

We present memory-based language modeling as an efficient, eco-friendly alternative to deep neural network-based language modeling. It offers log-linearly scalable next-token prediction performance and strong memorizatio…

Mimir: Large-scale Multilingual Concept Modeling

2026-05-24 · Elio Musacchio, Lucia Siciliani, Pierpaolo Basile arxiv

Current language modeling approaches are built around tokens. Text corpora are split into tokens, and models are trained by performing computations on these tokens, such as predicting the next token given the preceding o…