Learn how MaxText reproduced AI2’s OLMo 3 7B on Google Cloud TPUs, matching PyTorch GPU benchmarks across pre-training with up to 57.4% MFU.

Sign in to suggest edits

Key sources

  1. SOURCEdevelopers.googleblog.com
Markdown