Google DeepMindAnalysisAI ModelsSeptember 24, 2026

Google Cloud reproduces AI2's OLMo 3 7B training on TPUs

Read original source →developers.googleblog.com

MaxText reran AI2's OLMo 3 7B stage-1 pre-training and stage-2 mid-training on Cloud TPUs in JAX/XLA, matching the original PyTorch-on-GPU reference on all held-out evaluations across a ~5.93T-token, 1.41M-step budget. The run hit up to 57.4% Model FLOPs Utilization.

2 sources

More stories today

Open the live feed