paper-with-me

Papers

Jasmine: A Simple, Performant and Scalable JAX-based World Modeling Codebase

2025-10-30 · Mihir Mahajan, Alfred Nguyen, Franz Srambical, Stefan Bauer arxiv

While world models are increasingly positioned as a pathway to overcoming data scarcity in domains such as robotics, open training infrastructure for world modeling remains nascent. We introduce Jasmine, a performant JAX-based world modeling codebase that scales from single hosts to hundreds of accelerators with minimal code changes. Jasmine achieves an order-of-magnitude faster reproduction of the CoinRun case study compared to prior open implementations, enabled by performance optimizations across data loading, training and checkpointing. The codebase guarantees fully reproducible training and supports diverse sharding configurations. By pairing Jasmine with curated large-scale datasets, we establish infrastructure for rigorous benchmarking pipelines across model families and architectural ablations.

📄 PDF Abstract BibTeX arXiv:2510.27002

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

JASMINE: Arabic GPT Models for Few-Shot Learning

2022-12-21 · El Moatez Billah Nagoudi, Muhammad Abdul-Mageed, AbdelRahim Elmadany, Alcides Alcoba Inciarte 외

Scholarship on generative pretraining (GPT) remains acutely Anglocentric, leaving serious gaps in our understanding of the whole class of autoregressive models. For example, we have little knowledge about the potential o…

Few-Shot Learning

Jasmine: A New Active Learning Approach to Combat Cybercrime

2021-08-13 · Jan Klein, Sandjai Bhulai, Mark Hoogendoorn, Rob van der Mei

Over the past decade, the advent of cybercrime has accelarated the research on cybersecurity. However, the deployment of intrusion detection methods falls short. One of the reasons for this is the lack of realistic evalu…

Active LearningIntrusion Detection

Cross Spline Net and a Unified World

2024-10-24 · Linwei Hu, Ye Jin Choi, Vijayan N. Nair

In today's machine learning world for tabular data, XGBoost and fully connected neural network (FCNN) are two most popular methods due to their good model performance and convenience to use. However, they are highly comp…

Sable: a Performant, Efficient and Scalable Sequence Model for MARL

2024-10-02 · Omayma Mahjoub, Sasha Abramowitz, Ruan de Kock, Wiem Khlifi 외

As multi-agent reinforcement learning (MARL) progresses towards solving larger and more complex problems, it becomes increasingly important that algorithms exhibit the key properties of (1) strong performance, (2) memory…

Multi-agent Reinforcement Learningreinforcement-learningReinforcement Learning

Jasmine: Harnessing Diffusion Prior for Self-supervised Depth Estimation

2025-03-20 · Jiyuan Wang, Chunyu Lin, Cheng Guan, Lang Nie 외

In this paper, we propose Jasmine, the first Stable Diffusion (SD)-based self-supervised framework for monocular depth estimation, which effectively harnesses SD's visual priors to enhance the sharpness and generalizatio…

Depth EstimationImage ReconstructionMonocular Depth EstimationZero-shot Generalization