ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)

AMD released Instella-MoE-16B-A3B, a fully open Mixture-of-Experts model with 16B total parameters and 2.8B active per token, trained on Instinct MI300X and MI325X GPUs. Weights from every training stage, data mixtures, configs, and inference code are published under a research-only ResearchRAIL license.
Tap to vote and see what everyone thinks.
Summary by ByteBrief