2025/08/06 by Antoine Caillon, Lyria Team, Caillon, Antoine +68 · 2 voices · 2 citations
Computer Science · #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG) #Sound (cs.SD) #cs.HC #cs.LG #cs.SD
paper · pdf · doi:10.48550/arxiv.2508.04651
arxiv published 2025/08/06 · arxiv updated 2025/11/05
We introduce a new class of generative models for music called live music models that produce a continuous stream of music in real-time with synchronized user control. We release Magenta RealTime, an open-weights live music model that can be steered using text or audio prompts to control acoustic style. On automatic metrics of music quality, Magenta RealTime outperforms other open-weights music generation models, despite using fewer parameters and offering first-of-its-kind live generation capabilities. We also release Lyria RealTime, an API-based model with extended controls, offering access to our most powerful model with wide prompt coverage. These models demonstrate a new paradigm for AI-assisted music creation that emphasizes human-in-the-loop interaction for live music performance.
t5x and seqio