ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)
Opus 5 outperforms Opus 4.7, Opus 4.8, and Fable on benchmarks but feels like a downgrade to work with. Benchmark-driven training selects for bold assumptions over asking for clarification, which harms coding-agent usability. Anthropic's focus on self-improving AI and benchmark scores compounds the issue.
Tap to vote and see what everyone thinks.
Summary by ByteBrief