ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)
Perplexity open sourced Lily, a Rust and Metal inference engine for Qwen3.6-35B-A3B on Apple silicon, with no PyTorch or MLX in the execution path. On an M5 Max MacBook Pro, Lily averaged 23 percent faster prompt processing and 35 percent faster token generation than MLX-LM.
Tap to vote and see what everyone thinks.
Summary by ByteBrief