ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)

Speculative decoding uses a small drafter model to guess tokens, letting the main model verify batches in one pass. Google's Vivek Kumar shipped this on Pixel 9 and 10 devices, yielding 50% or more speedups and nearly two additional tokens per inference pass.
Tap to vote and see what everyone thinks.
Summary by ByteBrief