paper-with-me

Papers

Olmo 3

2025-12-15 · Team Olmo, :, Allyson Ettinger, Amanda Bertsch, Bailey Kuehl, David Graham, David Heineman, Dirk Groeneveld, Faeze Brahman, Finbarr Timbers, Hamish Ivison, Jacob Morrison, Jake Poznanski, Kyle Lo, Luca Soldaini, Matt Jordan, Mayee Chen, Michael Noukhovitch, Nathan Lambert, Pete Walsh, Pradeep Dasigi, Robert Berry, Saumya Malik, Saurabh Shah, Scott Geng, Shane Arora, Shashank Gupta, Taira Anderson, Teng Xiao, Tyler Murray, Tyler Romero, Victoria Graf, Akari Asai, Akshita Bhagia, Alexander Wettig, Alisa Liu, Aman Rangapur, Chloe Anastasiades, Costa Huang, Dustin Schwenk, Harsh Trivedi, Ian Magnusson, Jaron Lochner, Jiacheng Liu, Lester James V. Miranda, Maarten Sap, Malia Morgan, Michael Schmitz, Michal Guerquin, Michael Wilson, Regan Huff, Ronan Le Bras, Rui Xin, Rulin Shao, Sam Skjonsberg, Shannon Zejiang Shen, Shuyue Stella Li, Tucker Wilde, Valentina Pyatkin, Will Merrill, Yapei Chang, Yuling Gu, Zhiyuan Zeng, Ashish Sabharwal, Luke Zettlemoyer, Pang Wei Koh, Ali Farhadi, Noah A. Smith, Hannaneh Hajishirzi arxiv

We introduce Olmo 3, a family of state-of-the-art, fully-open language models at the 7B and 32B parameter scales. Olmo 3 model construction targets long-context reasoning, function calling, coding, instruction following, general chat, and knowledge recall. This release includes the entire model flow, i.e., the full lifecycle of the family of models, including every stage, checkpoint, data point, and dependency used to build it. Our flagship model, Olmo 3 Think 32B, is the strongest fully-open thinking model released to-date.

📄 PDF Abstract BibTeX arXiv:2512.13961

Code (0)

등록된 구현이 없습니다.

Tasks

Instruction Following

Similar Papers 제목 키워드 기반

OLMoASR: Open Models and Data for Training Robust Speech Recognition Models

2025-08-28 · Huong Ngo, Matt Deitke, Martijn Bartelds, Sarah Pratt 외 arxiv

Improvements in training data scale and quality have led to significant advances, yet its influence in speech recognition remains underexplored. In this paper, we present a large-scale dataset, OLMoASR-Pool, and series o…

Speech Recognition

OlmoEarth: Stable Latent Image Modeling for Multimodal Earth Observation

2025-11-17 · Henry Herzog, Favyen Bastani, Yawen Zhang, Gabriel Tseng 외 arxiv

Earth observation data presents a unique challenge: it is spatial like images, sequential like video or text, and highly multimodal. We present OlmoEarth: a multimodal, spatio-temporal foundation model that employs a nov…

Self-Supervised Learning

OlmoEarth v1.2: A more efficient family of OlmoEarth models

2026-05-20 · Gabriel Tseng, Yawen Zhang, Favyen Bastani, Henry Herzog 외 arxiv

We present a set of improvements to the OlmoEarth family. These improvements allow us to cut compute costs during training ($3.0 \times$ reduction in GPU hours required to train our Base models) and inference ($2.9\times…

MolmoB0T: Large-Scale Simulation Enables Zero-Shot Manipulation

2026-03-17 · Abhay Deshpande, Maya Guru, Rose Hendrix, Snehal Jauhri 외 arxiv

A prevailing view in robot learning is that simulation alone is not enough; effective sim-to-real transfer is widely believed to require at least some real-world data collection or task-specific fine-tuning to bridge the…

OLMoTrace: Tracing Language Model Outputs Back to Trillions of Training Tokens

2025-04-09 · Jiacheng Liu, Taylor Blanton, Yanai Elazar, Sewon Min 외

We present OLMoTrace, the first system that traces the outputs of language models back to their full, multi-trillion-token training data in real time. OLMoTrace finds and shows verbatim matches between segments of langua…

Fact CheckingHallucinationLanguage ModelingLanguage Modelling