Ai2 releases Olmo-core 3, open training stack for trillion-parameter MoE models
#ai2#olmo-core#mixture-of-experts#training-infrastructure
Ai2 has released Olmo-core 3, a redesigned, fully open training stack for efficiently scaling mixture-of-experts (MoE) models into the trillion-parameter range. The announcement was republished by the Hugging Face Blog.
Coverage timeline
Hugging Face Blog
Ai2 (Allen Institute for AI)
Olmo-core 3 introduces a redesigned, fully open training stack for efficiently scaling mixture-of-experts models into the trillion-parameter range.
