Lyria 3.5 Brings 3-Minute AI Song Generation to Flow Music
Google DeepMind's Lyria 3.5 model introduces variable three-minute track generation, granular creative controls, and multi-language vocals to Flow Music.
Google DeepMind has launched its Lyria 3.5 generative audio model within the Flow Music workspace. The latent diffusion architecture introduces support for variable song lengths up to three minutes and brings granular tempo and duration controls to the interface. For developers handling audio synthesis, the release establishes a new baseline for structural coherence and multi-language vocal generation in production environments.
Generation Capabilities
Lyria 3.5 utilizes a time-audio latent space to construct complete tracks with defined musical structures. The model now adheres strictly to standard compositional formats, generating distinct intro, verse, chorus, bridge, and outro sections while strictly following prompt constraints.
Vocal rendering features precise pronunciation and natural transitions across languages. The model processes mixed-language lyrics, such as combined English-Chinese or Japanese phrasing, without breaking melodic continuity or audio fidelity.
Platform Integration
Flow Music, acquired by Google in early 2026 under the name Producer AI, serves as a comprehensive vibe coding workspace powered entirely by Lyria 3.5. A July 24 update introduced Spaces, a collaborative environment for building custom instruments and music applications.
The workspace integrates direct stem splitting to isolate vocals from instrumentals and supports real-time remixing. Users can control rhythm patterns manually using Drum Machine Pro. The platform also connects with Google’s Veo video model to synthesize matching visuals for completed audio tracks.
API Access and Pricing
Beyond the consumer workspace, developers can route workloads to Lyria models through the Gemini API and Vertex AI. Consumer access is handled via Google AI Subscriptions, featuring the AI Pro plan at $19.99 per month for 1,000 credits and AI Ultra at $200 per month for 25,000 credits.
| Model | Output Duration | API Cost (Approximate) |
|---|---|---|
| Lyria 3 | 30-second clips | $0.04 per clip |
| Lyria 3 Pro / 3.5 | 3-minute songs | $0.08 per song |
Audio Provenance
Every track generated by Lyria 3.5 contains a SynthID watermark. The digital signal is embedded directly into the audio waveform at an inaudible frequency. The provenance marker remains detectable by scanning tools even after users apply compression, adjust playback speed, or introduce background noise.
If you integrate audio generation into production applications, the three-minute threshold and strict structural adherence alter the baseline for standalone music services. Evaluate your compliance posture regarding the ongoing March 2026 copyright litigation before deploying Lyria 3.5 endpoints in commercial environments.
Get Insanely Good at AI
The book for developers who want to understand how AI actually works. LLMs, prompt engineering, RAG, AI agents, and production systems.
Keep Reading
gr.Workflow Turns AI Pipelines Into Deployable APIs
Learn how to model, debug, expose, and deploy multi-step AI pipelines with Gradio Workflow and daggr.
DeepMind's AlphaGenome Atlas Predicts Effects of All 9 Billion DNA Variants
Google DeepMind released the AlphaGenome Atlas, a petabyte-scale catalog predicting the regulatory effects of every possible single-letter DNA change, built on the AlphaGenome model published in Nature.
Google Ships Gemini 3.7 Flash With Adjustable Reasoning Tiers
Google DeepMind's Gemini 3.7 Flash introduces tunable thinking levels and hits 65.3% on DeepSWE, targeting complex agentic workflows.
DeepMind SL2T Natively Translates Sign Language on Pixel 11
Google DeepMind has launched SL2T, an on-device direct-to-text sign language translation model debuting as a core accessibility feature on the Pixel 11.
WeatherNext Cuts Cyclone Intensity Prediction Errors by 28%
Google DeepMind's new Multi-Scale Mesh GNN generates 10-day global forecasts in under 45 seconds while drastically improving rapid intensification predictions.