Suno launches Speech feature for AI voiceovers with music
Suno launched Speech in public beta, letting users generate spoken voiceovers with optional AI background music.
“Music will always be at the heart of Suno and what we build. At the same time, our vision has always extended to other forms of human expression,” Suno chief product officer Jack Brody said in the announcement. “Today, we’re expanding what’s possible in Suno with Speech: the first audio model that generates voice and music together as one cohesive track.”
Speech allows users to generate voiceovers and background music at the same time. The music is optional, and a toggle lets users turn it off if they want clean speech. Suno says the combination is meant to complement uses such as a calming soundtrack for poems or more energetic music for dramatic voiceovers and encouraging speeches.
To use Speech, users select the Create tab and navigate to the Speech option. There are two modes: Simple, which lets users describe what they want through a prompt box, such as “a pirate captain rallying his crew,” and Advanced, which lets users add a custom script. Advanced settings also allow adjustments to the AI voice’s gender, speech style, and how much variety each voice generation will have. Speech has a maximum duration of about eight minutes.
AI-generated speech is not new. The Verge reported that DeepMind has been experimenting with deep learning speech synthesis for a decade, Adobe has a text-to-speech tool, and ElevenLabs has become one of the most recognizable platforms since launching in 2023. The Verge noted that Suno is entering an established market and that the move is likely an effort to diversify its platform, as its music generator has attracted lawsuits.
Suno acknowledges that Speech is far from perfect and says it will keep improving the feature based on user feedback. “Beta really does mean beta,” Brody said. “Occasionally, British accents can wander off to Australia and back. Dramatic pauses may be very dramatic. You will almost certainly discover uses for this that never occurred to us.”
Editor's Summary Suno has expanded beyond AI music with Speech, a public beta tool that generates spoken voiceovers and can add AI background music. The feature offers Simple and Advanced modes, with a maximum duration of about eight minutes. Suno acknowledges current limitations and says it will refine Speech based on user feedback.