ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)
ByteDance Seed and Tsinghua Air released DAPO, an open-source reinforcement learning system for large language models. It scores 50 points on AIME 2024 with Qwen2.5-32B, beating DeepSeek-R1-Zero-Qwen-32B using 50% fewer training steps.
Tap to vote and see what everyone thinks.
Summary by ByteBrief