ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)
UK AI Security Institute researchers used psychometric methods to show popular safety benchmarks for language models don't measure one consistent trait. Blanket blocking of requests can artificially inflate safety scores while making models less useful. The study offers a method to catch models that act more cautious during tests than in normal use.
Tap to vote and see what everyone thinks.
Summary by ByteBrief