AI music maker Suno now generates spoken words - The Verge
AI music maker Suno now generates spoken words The new Speech feature provides synthetic voiceovers embellished with AI background music. The new Speech feature provides synthetic voiceovers embellished with AI background music.

- AI music maker Suno now generates spoken words The new Speech feature provides synthetic voiceovers embellished with AI background music.
- The new Speech feature provides synthetic voiceovers embellished with AI background music.
- Suno is branching out from the world of AI music, launching a new feature that generates spoken voices based on scripts or prompted descriptions.
AI music maker Suno now generates spoken words The new Speech feature provides synthetic voiceovers embellished with AI background music. The new Speech feature provides synthetic voiceovers embellished with AI background music. If you buy something from a link, The Verge may earn a commission. See our ethics statement. Jess Weatherbed is a news writer focused on creative industries, computing, and internet culture. Jess started her career at TechRadar, covering news and hardware reviews. Suno is branching out from the world of AI music, launching a new feature that generates spoken voices based on scripts or prompted descriptions. Speech is now available in public beta across Suno’s web and mobile platforms, and allows you to simultaneously generate voiceovers and background music to accompany them. “Music will always be at the heart of Suno and what we build. At the same time, our vision has always extended to other forms of human expression,” Suno chief product officer, Jack Brody, said in the announcement. “Today, we’re expanding what’s possible in Suno with Speech: the first audio model that generates voice and music together as one cohesive track.” AI-generated speech is hardly new — DeepMind has been experimenting with deep learning speech synthesis for a decade, Adobe has a text-to-speech tool, and ElevenLabs has become one of the most recognizable platforms for it since launching in 2023. Suno is just throwing its hat into the ring — likely in an attempt to diversify the platform, given its music generator has attracted so many lawsuits . Pairing AI music with generated voices is Suno’s spin on text-to-speech tools. It’s optional, meaning you can easily turn off the background music with a toggle if you just want clean speech, but the idea is that it’ll compliment certain use cases for generative spoken word — such as having a calming soundtrack for poems, or something more energetic for dramatic voiceovers and encouraging speeches. To use the feature, select the “Create” tab, and navigate to the Speech option. There are two modes: Simple, which allows you to describe what you want to create via the provided prompt box (such as “a pirate captain rallying his crew”), or the Advanced mode that lets you add a custom script if you already know exactly what you want it to say. Advanced settings also let you adjust the gender of the AI voice, speech style, and how much variety each voice generation will have. Speech has a maximum duration of around eight minutes. Suno admits that the feature is far from perfect, but says it’ll keep improving Speech around user feedback. “Beta really does mean beta,” said Brody. “Occasionally, British accents can wander off to Australia and back. Dramatic pauses may be very dramatic. You will almost certainly discover uses for this that never occurred to us.” Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates. Jess Weatherbed The rise, fall, and rise of portable MP3 players The iPad Mini is slightly cheaper again during Prime Day The MacBook Air M5 is $200 off for the first time in months Nacon’s new PS5 controller can mix audio from your phone and console Keurig’s new machine uses plastic-free compressed coffee pucks A free daily digest of the news that matters most. By providing your information, you agree to our Terms of Use and our Privacy Policy . We use vendors that may also process your information to help provide our services. This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.
Sources
Related stories

ElevenLabs valuation doubles to $22B – four times that of AI music rival Suno - musicbusinessworldwide.com
ElevenLabs founders Piotr Dąbkowski and Mati Staniszewski ElevenLabs , the AI audio company behind music generation platform ElevenMusic , has completed a USD $300 million employee tender offer at a valuation of $22 billion .

Audio AI firm Modulate, which runs an AI music detection tool, raises $25M - musicbusinessworldwide.com
Modulate , the Boston audio AI company behind a gaming voice moderation tool called ToxMod , has raised USD $25 million in new funding.

Eleven v4 and Eleven v4 Turbo Text to Speech models - ElevenLabs
Eleven v4 and Eleven v4 Turbo Text to Speech models &)]:tw-overflow-x-hidden"> strong]:tw-font-normal tw-text-gray-600"> Introducing Eleven v4 Meet Eleven v4, our most emotive model yet. With 3x credits included on Creator+ until October 12 Create controllable, expressive speech layered with emotion, audio events, and immersive soundscapes.

Making global data easier to explore
UN System Data Commons is an open, AI-ready platform integrating critical global statistics into a single searchable resource. Your browser does not support the audio element.