Ai2 open-sources Olmo-core 3, training infrastructure scaling MoE models toward trillion parameters
allen_ai · x · 2026-10-01
The Allen Institute for AI released Olmo-core 3, the open training infrastructure behind the next generation of OLMo, designed to scale mixture-of-experts models into the trillion-parameter range. The PyTorch building-block library is on GitHub and PyPI (ai2-olmo-core), with optional integrations for flash-attn/ring-flash-attn/TransformerEngine attention backends, Liger-Kernel fused-linear loss, torchao float8 training, and groupedgemm for dropless MoE.
Related event: Ai2 Open-Sources Olmo-core 3 for Trillion-Parameter MoE Training(2 posts)→
More from Infra
- AMD shows off Helios system with OpenAI aboard amid deepening infra ties — AnushElangovan · 2026-10-01
- Qwen-Image 2.1 prompt enhancer hits 4.4x speedup in ComfyUI, now runs on 8GB VRAM — mozophe · 2026-10-01
- Running Omarchy desktop in Windows via WSL with GPU acceleration and 4K multi-monitor support — sytelus · 2026-10-01
- Trader initiates Cerebras position, betting SRAM-based inference beats HBM as agents multiply model calls — Sethwinterroth · 2026-10-01
- Cerebras bull case: OpenAI paid tier, ~750 tok/s, $20B+ potential value and $25B RPO — Sethwinterroth · 2026-10-01
- MLX-Serve 26.10.1 ships with up to 66% faster Qwen3.8 27B inference on Apple Silicon — TheMoonMidas · 2026-10-01