Google Developers Blog(RSS)·· 16 小时前AI 评分52
Google 用 MaxText 在 TPU 上复现 Ai2 Olmo 3 7B 预训练与中期训练
Reproducing Olmo 3 7B Pre-training in MaxText: case study of large scale training on TPUs
AI 导读
Google 团队在 MaxText(JAX/XLA)中从零复现 Ai2 的 Olmo 3 7B,在 Google Cloud TPU 上完成 stage-1 预训练(约5.93T token。
来源:Google Developers Blog(RSS) · developers.googleblog.com