Google Developers Blog(RSS)· Gagik Amirkhanyan·· 14 天前AI 评分51
Google 团队在 TPU 上用 MaxText 复现 Ai2 的 Olmo 3 7B 预训练
Reproducing Olmo 3 7B Pre-training in MaxText: case study of large scale training on TPUs
AI 导读
Google 团队用 MaxText 在 Google Cloud TPU 上从头复现 Ai2 的 Olmo 3 7B,完成 stage-1 预训练(约5.93T token。
整理与数据来源:AIHOT
来源:Google Developers Blog(RSS) · developers.googleblog.com