Topic map
daily cluster · Oct 3, 2026 — Oct 3, 2026

Scalable Mixture-of-Experts Training Infrastructure

The organizations collaborated on releasing Olmo-core 3, a redesigned open training stack. This infrastructure is optimized for scaling mixture-of-experts models into the trillion-parameter range.

Why it matters

Open infrastructure for giant models democratizes supercomputing scale training capabilities beyond closed-source providers.

Hugging FaceAllen Institute for AI

Source posts · 2

Hugging Face
blog · 2d ago

Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs

Hugging Face introduced Olmo-core 3, an open and scalable training infrastructure designed for large Mixture of Experts models.

moesinfrastructureopen sourceSource
Allen Institute for AI
blog · 2d ago

Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs — Olmo-core 3 introduces a redesigned, fully open training stack for efficiently scaling mixture-of-experts models into the trillion-parameter range.

The Allen Institute for AI has introduced Olmo-core 3, an open training infrastructure for scaling mixture-of-experts models.

open sourcemoeinfrastructureSource