Lyria 3.5 doubles as a TTS model: one prompt yields eerily natural 60-second speech
fofrAI · x · 2026-09-06
- fofrAI found that Google's music-generation model Lyria 3.5 can be prompted as a text-to-speech model, producing raw sound bites that "seam real."
- Example prompt: "just a british millennial woman talking with a feeling of being jaded, candid and natural, no lyrics, no music, no beats, no percussion, just 60 seconds long."
- The output is conversational speech rather than singing, suggesting the model's underlying capability extends well beyond its official music focus — a non-obvious trick worth testing for voice and audio content workflows.
More from Multimodal
- GPT-6 'Astra' demos exploding 3D view of human arm anatomy — nickbaumann_ · 2026-09-06
- SplatBox demo converts millions of Gaussian Splats into interactive destructible Unity voxel worlds — jonstephens85 · 2026-09-06
- Open-sourced VoxCPM 1.5 TTS beats Orpheus on cost-per-concurrency by 4.5x — TheMoonMidas · 2026-09-06
- Video References Beat Text Prompts: A MiniMax Art-Making Workflow — Alive-Tomatillo5303 · 2026-09-06
- GPT 6 Astra turns an empty room into a fully editable 3D concept in 30 seconds — SimplyAnnisa · 2026-09-06
- MiniMax-H3 Lip Sync on 8GB VRAM: Multishot Renders in 26 Minutes — big-boss_97 · 2026-09-06