AI music maker Suno now generates spoken words
Sunoās new beta lets users generate spoken word tracks with optional AIāgenerated background music, streamlining narration and audio production.

Sunoās New Speech Feature: AIāGenerated Voiceovers with BuiltāIn Music
Suno, the AI music platform that has been making headlines for its ability to compose songs from simple prompts, has just added a new Speech feature. The beta lets users generate spoken word tracksācomplete with optional background musicādirectly from a script or a short description. The move signals Sunoās ambition to become a oneāstop shop for all kinds of audio content, not just music.
How the Feature Works
Sunoās interface keeps the familiar Create tab, but now includes a Speech option. Users can choose between two modes:
- Simple ā type a prompt like "a pirate captain rallying his crew" and let the model decide the voice, style, and accompanying music.
- Advanced ā paste a full script and tweak settings such as gender, speech style, and voice variety. This mode also lets you toggle the background music on or off.
The result is a single audio file that blends the spoken words with a music track generated by Sunoās core model. The maximum length is about eight minutes, and the system currently supports a handful of voice styles and accents.
What Makes Sunoās Approach Different
AI textātoāspeech has been around for years. DeepMind, Adobe, and ElevenLabs have all released robust solutions that can read any text in a naturalāsounding voice. Sunoās twist is the integration of music generation and speech synthesis into one cohesive track. Rather than layering a separate music file over a TTS output, the model learns to compose music that matches the rhythm, emotion, and pacing of the spoken words.
This can be especially useful for creators who want a readyāmade soundtrack for podcasts, audiobooks, or video narrations. For example, a user could generate a calming background for a spoken poem or an energetic beat for a motivational speechāall in one step.
Current Limitations and Future Plans
Suno acknowledges that the beta is far from perfect. Accents can slip between British and Australian, dramatic pauses may feel exaggerated, and the voice library is still limited. The company says it will iterate on the model based on user feedback and expand the range of voices and styles.
The feature is available on both the web and mobile apps, and Suno has positioned it as a complement to its existing music generation tools rather than a replacement. Users can still choose to generate clean speech without music if they prefer.
Who Should Try It
- Content creators looking for a quick way to add professionalāsounding narration to videos or podcasts.
- Marketing teams that need voiceāover tracks for ads or explainer videos.
- Educators who want engaging audio lessons with background music.
If youāre curious, the beta is free to try. Just head to the Create tab, select Speech, and start experimenting with prompts or scripts.
Bottom Line
Sunoās Speech feature is a bold step into the broader AI audio space. While it still has rough edges, the ability to generate voice and music together could streamline production workflows for a wide range of creators.
---
TL;DR: Sunoās new beta feature lets users generate spoken word tracks with optional AIāgenerated background music, offering a oneāstop solution for narration and audio production.
Source
Read original sourceWhy we picked this
Sunoās new voice generation feature is a meaningful AI product release.