ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)

DeepSeek has released DeepSeek-V4.1-Flash, a 552B-parameter Mixture-of-Experts model with a 1-million-token context window. It uses a new Causal Encoder-Decoder architecture to activate only 8B parameters during prefill and 16B during decoding. The model reduces its global KV cache to 890 bytes per token, aiming to make long-context AI inference significantly cheaper and more efficient for running agents.
Tap to vote and see what everyone thinks.
Summary by ByteBrief