跳到正文
HChuggingface.co·
暂不在当前实时榜单

Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs

AI 摘要

Olmo-core 3 is an open and scalable training infrastructure designed for large Mixture-of-Experts (MoEs) models. Benchmarked on NVIDIA B300 GPUs, it achieved a throughput of 858 TFLOP/s/GPU with a 1.2-trillion-parameter model, using 58.36 billion parameters active per token across 512 GPUs. These benchmarks focused on system performance using random routing, not the quality of a trained model. Further details on its design and experiments are available in its technical report and on GitHub.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年10月1日 16:00 UTC

收录
2026年10月1日 16:00
来源类型
官方发布

本站未收录正文。

前往源站阅读 →