ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)

Tsinghua University researchers published a paper on Cache-to-Cache (C2C), a technique that lets separate AI models exchange internal memory directly, skipping text generation. Accepted at ICLR 2026 with open-source code, it reports 100% to 150% faster inference and up to 14.2% accuracy gains. C2C currently works only with open-weight models.
Tap to vote and see what everyone thinks.
Summary by ByteBrief