ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)
Hugging Face released Olmo-core 3, a redesigned open mixture-of-experts training framework for large language models. It scales to 1.2 trillion parameters across 512 Nvidia B300 GPUs, and a 47-billion-parameter MoE hit 52,000 tokens per second per GPU, about 2.7 times its earlier implementation. The stack underpins the next-generation Olmo model.
Tap to vote and see what everyone thinks.
Summary by ByteBrief