chiprook
← AI
AIOctober 1, 2026, 22:16

Ai2 Releases Olmo-core 3, Open Training Stack for Trillion-Parameter MoEs

The Allen Institute for AI released Olmo-core 3, an open mixture-of-experts training stack benchmarked at over one trillion total parameters. On 512 NVIDIA B300 GPUs, a 1.2-trillion-parameter model with 58.36 billion active parameters reached 858 TFLOP/s per GPU, while an 8-GPU test showed 2.7x higher throughput than the earlier implementation.

Ai2 Releases Olmo-core 3, Open Training Stack for Trillion-Parameter MoEs
#Ai2#Olmo#Nvidia
Read next
AI

China Telecom open-sources Xing4.0-29B-A4B agentic MoE trained on Ascend

AI

Xiaomi Open-Sources Robotics-U0 Embodied World Model and Training Stack

AI

NaiveAI releases Naive-N0.5-Flash: 309B MoE open-weight model under MIT

AI

Jina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding for Low-Budget GPUs