ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)
DeepSeek-V4-Flash-0731-Latent-Reasoning is now a self-contained model with NVFP4 quantization, a 35.7M-param latent reasoning head, and a production vLLM runtime. It scores 0.880 on BBH, with strongest multi-step tracking and weakest bracket matching at 0.26.
Tap to vote and see what everyone thinks.
Summary by ByteBrief
DeepSeek V4-Flash Launches as Cheapest Major AI Model