ByteBrief
We're a portrait publication through and through. Turn your phone back and your briefing picks up right where you left it.
(We tried widescreen once. It wasn't us.)
Researchers analyzed 2,100 AI-generated cat vocalizations across three text-to-audio models and seven prompts, finding that model performance varies widely by prompt specificity. EzAudio often produced silence or background noise unless actions were explicit, while TangoFlux and Stable Audio Open generated meows consistently. The study introduces expressive-range plots to compare generative audio models.
Tap to vote and see what everyone thinks.
Summary by ByteBrief