跳到正文
原文
Google Developers Blog(RSS)· Gagik Amirkhanyan·· 14 天前AI 评分51

Google 团队在 TPU 上用 MaxText 复现 Ai2 的 Olmo 3 7B 预训练

Reproducing Olmo 3 7B Pre-training in MaxText: case study of large scale training on TPUs

AI 导读

Google 团队用 MaxText 在 Google Cloud TPU 上从头复现 Ai2 的 Olmo 3 7B,完成 stage-1 预训练(约5.93T token。

整理与数据来源:AIHOT

来源:Google Developers Blog(RSS) · developers.googleblog.com