RELEASE · STARTUPS · #1431
Suno launches 'Speech' public beta to generate voiceovers with AI background music
Suno has released Speech in public beta on its web and mobile apps, a feature that generates synthetic voiceovers together with AI‑created background music. Speech offers Simple and Advanced modes (including custom scripts, voice gender/style controls, and an ~8‑minute maximum duration) and includes a toggle to disable music for clean speech; Suno says the model is still in early beta and will be improved with user feedback.
KEY POINTS
- Suno has released Speech in public beta on its web and mobile apps, a feature that generates synthetic voiceovers together with AI‑created background music.
- Speech offers Simple and Advanced modes (including custom scripts, voice gender/style controls, and an ~8‑minute maximum duration) and includes a toggle to disable music for clean speech; Suno says the model is still in early beta and will be improved with user feedback.
- Combining speech synthesis and music in one model extends Suno beyond music generation, positions it to compete with dedicated text‑to‑speech providers, and enables new creative use cases for voiceover plus soundtrack in a single output.
WHY IT MATTERS
Combining speech synthesis and music in one model extends Suno beyond music generation, positions it to compete with dedicated text‑to‑speech providers, and enables new creative use cases for voiceover plus soundtrack in a single output.
SOURCES & TIMELINE
2The new Speech feature provides synthetic voiceovers embellished with AI background music. If you buy something from a link, The Verge may earn a commission. See our ethics statement. Suno is branching out from the world of AI music, launching a new feature that generates spoken voices based on scripts or prompted descriptions. Speech is now available in public beta across Suno’s web and mobile platforms, and allo…
Suno, best known as an AI music generator, is adding a new feature called Speech. It produces spoken text and matching background music together as a single audio track. Users type in an idea or written text and describe the voice and music style they want. The model then generates both the voice and the sound. Suno product chief Jack Brody says the company tested Speech with a small group for a month. Suno says the…