ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)

Leading LLM models now sit within a six-point GPQA spread, making benchmark scores insufficient for vendor selection. Key differentiators include refusal rates, control for price, and national content rules. Anthropic models refuse 21 times more often than market median. Chinese open-weight models add a political filter layer specific to China's regulations.
Tap to vote and see what everyone thinks.
Summary by ByteBrief