Suno Speech (Beta): How to Make AI Voiceovers, Narration and Spoken Word With Background Music in One Track
Suno Speech is a new beta in Suno's Create page that generates a spoken voice and its background music together. How Simple and Advanced mode work, the Vocal Gender, Background music and Variety controls, prompt examples, the beta's known limits, and what is still unannounced.

On October 1, 2026 Suno opened Speech (beta) to everyone on the web, iOS and Android. It is the first Suno feature that is not about singing: you type an idea or paste a script, describe the voice and the music you want, and Suno generates a spoken performance and its soundtrack together, as one track. Suno calls it “the first audio model that generates voice and music together as one cohesive track.”
The short version: Suno Speech is an AI voiceover tool with a built-in composer. Simple mode writes the words for you from a one-line idea; Advanced mode reads your own script in the style you describe. Background music is on by default and can be switched off for clean narration. A take can run to about eight minutes. Suno has not yet published a credit price, a language list or an API for it, and it says plainly that the beta still has rough edges.
Sources for everything below: Suno’s launch post by Chief Product Officer Jack Brody, the October 1 release note, Suno’s official tutorial video How to Use Suno Speech (Beta) (the interface screenshots in this article are taken from it), The Verge’s launch coverage and the reaction on X. Where Suno has not said something, we say so rather than guess. This site is an independent Suno download tool and is not affiliated with Suno.

Suno Speech at a glance
| Suno Speech (beta) | |
|---|---|
| Launched | October 1, 2026, as an open beta after a month of testing with a small group |
| Where | suno.com and the Suno apps for iOS and Android, under Create → Speech |
| Who can use it | Everyone, according to Suno; no plan restriction has been announced |
| What it makes | A spoken voice and original background music in a single track; music can be switched off |
| Modes | Simple (describe the idea) and Advanced (custom script + style of speech) |
| Controls | Vocal Gender (Male / Female), Background music (On / Off), Variety slider |
| Length | Up to about 8 minutes per take |
| Not announced yet | Credit cost per take, supported languages, an API, use of your own Voice as the narrator |
What Suno Speech actually does
Most AI voiceover work today is two jobs. A text-to-speech tool reads the script, a music tool or a stock library supplies the bed, and someone lines the two up in an editor and fixes the levels. Suno Speech does both in one pass. Because the score is generated together with the voice, it can swell when the delivery gets dramatic and drop back when the narrator slows down, rather than playing underneath at one level.
Suno pitches it with playful examples. The release note offers “bedtime stories over soft piano, hype speeches over stadium drums, ASMR grocery lists and more.” In the launch post Brody lists what the team made while building it: friends’ texts turned into “wildly overproduced dramatic readings”, ordinary voice notes with “unnecessarily epic scores”, meditations, poems, pep talks and bedtime stories for their kids. Suno’s tutorial adds more practical uses: a museum audio guide, a train station announcement, a spoken-word piece and a study explainer.
It is also not Suno’s first voice model. Before it went all-in on songs, Suno released Bark, an open-source text-to-speech model, in 2023. Speech is the first voice model built into the Suno app itself.
How to use Suno Speech on the web
Open suno.com/create. The left-hand panel now has three tabs: Songs, Speech and Sounds. Click Speech. The panel carries a red BETA badge, and you choose between Simple and Advanced at the top.
Simple mode: describe the finished idea

In Simple mode you write one prompt that covers what is being said, who is saying it and what plays underneath. You do not write the script; Suno does. In the tutorial the prompt is a museum audio guide for an obviously fake exhibit about the invention of the snooze button — “calm, serious narrator, slightly dramatic, with elegant orchestral music underneath” — and Suno returns a full narration (“Welcome to Gallery 7. Please stand a respectful distance from the glass…”) with a matching orchestral score. The result in the tutorial runs a little over five minutes, and the request comes back as two takes, the same way a song generation does.
The Verge’s example prompt, “a pirate captain rallying his crew”, shows how little you need to type.
Advanced mode: your script, your direction

Switch to Advanced when you already know exactly what should be said. You get two boxes:
- Custom script — “Write or paste your speech.” This is the text Suno performs.
- Style of speech — “Describe the delivery: tone, pacing, mood and setting.” This is where you direct the performance and, if music is on, the score.
The tutorial’s best demonstration keeps one short announcement script and changes only the style. “Calm and professional train station announcer, slow, measured pacing, completely serious delivery” produces a deadpan read of “Attention passengers, the 6:42 train to nowhere in particular is now boarding on platform 9.” Change the style to “increasingly frantic announcer trying to remain professional, starts controlled, gets faster and more exasperated with every sentence” and the same words come back as a completely different performance.
Suno’s formatting advice for the script is simple: write it like normal speech, in plain paragraphs. You do not need lyric formatting or bracketed section tags such as [Verse] or [Chorus] the way you would for a song.
The Advanced controls: Vocal Gender, Background music and Variety

- Vocal Gender — Male or Female. Neither is selected in the default view, so you only need it when the gender matters.
- Background music — On by default. Switch it Off when you need clean spoken audio: narration for a video you will score yourself, dialogue, instructions, or anything where you already have the music.
- Variety — how much the takes are allowed to vary. It sits at Normal by default; move it left for more predictable reads, right for more surprising ones.
Suno Speech on iPhone and Android

Speech is in the mobile apps too; Suno’s announcement on X tells users to “update your app for the latest” if it does not appear. In the app you pick Speech from the menu at the top of Create, and the Advanced fields become chips: Backing music on (tap the × to remove it), Script and Tone. The tutorial’s mobile example is a study aid — “I have a test tomorrow. Explain osmosis to me like a patient tutor. Start with a simple analogy, then give me a short explanation I could remember for the test” — which turns revision notes into something you can listen to on the way to class.
Suno Speech prompt examples
These follow the pattern Suno uses in its own examples: say what the piece is, then the voice, then the music.
Bedtime story (Simple mode)
A two-minute bedtime story about a fox who learns to share fireflies. Soft, warm storyteller voice with a gentle smile in it, slow pace. Quiet piano and soft night pads underneath, lullaby tempo.
Hype speech (Simple mode)
A coach’s half-time pep talk to a team that is losing 2–0. Loud, rallying, building to a shout. Stadium drums and brass that grow with every line.
Your own script (Advanced mode)
Custom script: paste the exact words in short paragraphs.
Style of speech: Warm adult narrator, intimate and conversational, steady pace, clear pronunciation. Sparse felt piano underneath, voice forward in the mix, gentle ending.
Clean voiceover, no music (Advanced mode)
Background music: Off. Variety: far left.
Style of speech: Neutral, friendly product-demo narrator, medium pace, clear diction. No background music.
Tips for better results
- Direct the delivery, not just the content. Tone, pacing, mood and setting are the four things Suno’s own placeholder asks for. “Starts calm, speeds up, ends breathless” is more useful than “dramatic”.
- Keep scripts speech-shaped. Short sentences, punctuation where a speaker would breathe, paragraph breaks between thoughts. Rhymes and repeated refrains can pull the model toward a rhythmic, half-sung read.
- Test the style on a short script first. The tutorial’s train announcement shows that the same words can sound completely different; settle the style on a few lines before you paste a five-minute script.
- Listen against your text. Check names, numbers and any word that has to be exact before you publish. This is a creative tool, not a word-perfect text-to-speech engine.
- For voice only, stack the signals. Background music Off is the main switch. An early community tip, reported by aireiter, adds moving Variety fully left and writing “No background music.” in the style or tone field, because music occasionally crept in anyway. Treat that as a workaround, not a guarantee.
- Do not imitate real people. Describe a voice by age, texture, accent and attitude rather than naming someone whose voice you do not have permission to use.
What Suno says the beta still gets wrong
Suno is unusually frank about this. “Beta really does mean beta,” Brody writes. “Occasionally, British accents can wander off to Australia and back. Dramatic pauses may be very dramatic.” Expect a share of takes where the accent drifts mid-script, a pause runs long, or a word is read differently from how you meant it. Generating two takes and picking the better one is part of the workflow.
A few things Suno has not published yet:
- Credit cost. Neither the blog post nor the release note gives a price per take or says whether Free and paid plans get different limits. The Create screen shows your balance; check it before a batch of long scripts.
- Languages. No supported-language list has been published.
- Your own voice. Suno’s separate Voices feature lets you sing with your own verified voice in songs. Nothing Suno has published says a saved Voice can be used as a Speech narrator.
- An API. Speech is in the website and the apps only, like the rest of Suno. Our Suno API explainer covers what that means for developers.
- Speech-specific rights. Suno has not issued separate terms for Speech. Its general rule is that commercial use comes with output made on a paid plan, and Free-plan output is for personal use; our legal guide goes through the details.
Suno Speech vs a text-to-speech tool
Suno Speech and tools like ElevenLabs solve different problems. Pick Suno Speech when the performance and the music belong together: bedtime stories, poems, meditations, trailers, toasts, character pieces, dramatic readings of the group chat. Pick a dedicated TTS tool when you need the same voice across hundreds of clips, exact timing, word-perfect pronunciation or an API — product tutorials, audiobooks, accessibility, apps. The two can also work together: draft the scored version in Suno, and keep a dry TTS read for anything that has to be precise.
What people are saying on X
Suno’s own announcement on X was short: “Speech is now in Beta. Try the first model ever that creates spoken audio with matching background music. Update your app for the latest.” AI news accounts picked it up within hours; the Chinese-language AI commentator Gorden Sun (@Gorden_Sun) summed it up as a natural step for a company that already generates sung vocals, with the selling point that “the generated speech fits the background music perfectly.”
Not everyone was convinced. On r/SunoAI one user called Suno’s own demo “worryingly bad” (quoted in the aireiter guide linked above), and The Verge pointed out that AI speech is a crowded field, from ElevenLabs to Adobe, and read the launch partly as Suno diversifying beyond music. On the same day, CEO Mikey Shulman told Bloomberg that Suno is now “far beyond” the 2 million subscribers and $300 million in revenue it last reported, though he gave no new figure.
Downloading Suno Speech tracks
Speech takes appear in your workspace next to your songs, labelled SPEECH, with the same like, pin and share buttons. To keep a file, use Suno’s own download on your plan — the download limits that apply to songs are the ones to plan around until Suno says otherwise.
Suno Music Downloader works with public Suno links. If you publish a Speech track and its share link opens a normal suno.com/song/… page, paste it in and you can save a listening copy as M4A, MP3, WAV or MP4 for personal use, without signing in to Suno and without using your download quota. Speech is still in beta, so if a Speech link does not load in the tool, tell us and we will look at it.
Frequently asked questions
What is Suno Speech? Suno Speech is a beta feature in Suno’s Create page that generates spoken audio — narration, voiceovers, stories, speeches — together with original background music in a single track. Suno launched it on October 1, 2026 and calls it the first audio model that makes voice and music together.
Is Suno Speech free? Suno says the beta is open to everyone on web and mobile, but it has not published how many credits a Speech take costs or whether limits differ by plan. Your credit balance is shown on the Create screen.
How do I use Suno Speech? Open Create on suno.com or in the app and choose Speech. In Simple mode, describe the piece, the voice and the music in one prompt. In Advanced mode, paste your own text into Custom script and describe the delivery in Style of speech, then set Vocal Gender, Background music and Variety if you need to.
Can Suno Speech make a voiceover without music? Yes. Turn Background music off in Advanced mode, or remove the Backing music chip on mobile. Some early users add “No background music.” to the style and lower Variety as well, because music occasionally slipped in during the beta.
How long can a Suno Speech track be? Up to about eight minutes per take, according to Suno’s tutorial.
Can I use my own voice in Suno Speech? Suno has not said so. Personal voices are part of Suno’s separate Voices feature for songs, and nothing Suno has published says a saved Voice can narrate a Speech track.
Does Suno Speech have an API? No. Like the rest of Suno, Speech is available only on the website and in the iOS and Android apps.
Is Suno Speech good enough to replace text-to-speech? Not for everything. It is strongest for short, expressive pieces where music belongs in the result. For long, word-perfect or repeatable narration, a dedicated text-to-speech tool is still more predictable, and Suno itself warns that accents and pauses can drift in the beta.
Can I download Suno Speech tracks with Suno Music Downloader? The tool reads public Suno links. A published Speech track that opens as a normal suno.com/song page can be saved like a song; Speech is new and in beta, so tell us if a Speech link does not work.
We will update this article when Suno publishes pricing, languages or other changes to Speech.


