ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)
Apple introduces LensVLM-9B, a vision-language model that maintains accuracy when processing compressed images by selectively expanding relevant parts, outperforming baselines up to 10.1x compression across seven text QA benchmarks and generalizing to document and code understanding tasks.
Tap to vote and see what everyone thinks.
Summary by ByteBrief