Suno Speech (Beta): How to Make AI Voiceovers, Narration and Spoken Word With Background Music in One Track

Suno Speech is a new beta in Suno's Create page that generates a spoken voice and its background music together. How Simple and Advanced mode work, the Vocal Gender, Background music and Variety controls, prompt examples, the beta's known limits, and what is still unannounced.

Suno Speech (Beta): How to Make AI Voiceovers, Narration and Spoken Word With Background Music in One Track

On October 1, 2026 Suno opened Speech (beta) to everyone on the web, iOS and Android. It is the first Suno feature that is not about singing: you type an idea or paste a script, describe the voice and the music you want, and Suno generates a spoken performance and its soundtrack together, as one track. Suno calls it “the first audio model that generates voice and music together as one cohesive track.”

The short version: Suno Speech is an AI voiceover tool with a built-in composer. Simple mode writes the words for you from a one-line idea; Advanced mode reads your own script in the style you describe. Background music is on by default and can be switched off for clean narration. A take can run to about eight minutes. Suno has not yet published a credit price, a language list or an API for it, and it says plainly that the beta still has rough edges.

Sources for everything below: Suno’s launch post by Chief Product Officer Jack Brody, the October 1 release note, Suno’s official tutorial video How to Use Suno Speech (Beta) (the interface screenshots in this article are taken from it), The Verge’s launch coverage and the reaction on X. Where Suno has not said something, we say so rather than guess. This site is an independent Suno download tool and is not affiliated with Suno.

Suno's Speech launch graphic: cards labelled SPEECH for Sunset Calmness (meditation, soft voice), Go Team! (energetic cheerleader), Digitale Roboto (Italian robot, 90s beep bop), Scottish Catter (Scottish banter, pub, bagpipes) and Grocery Lists (Victorian English with a harpsichord)
Suno's launch graphic. Each card is one Speech track: a title, then the voice and the music in a few words. Image: Suno.

Suno Speech at a glance

Suno Speech (beta)
Launched October 1, 2026, as an open beta after a month of testing with a small group
Where suno.com and the Suno apps for iOS and Android, under Create → Speech
Who can use it Everyone, according to Suno; no plan restriction has been announced
What it makes A spoken voice and original background music in a single track; music can be switched off
Modes Simple (describe the idea) and Advanced (custom script + style of speech)
Controls Vocal Gender (Male / Female), Background music (On / Off), Variety slider
Length Up to about 8 minutes per take
Not announced yet Credit cost per take, supported languages, an API, use of your own Voice as the narrator

What Suno Speech actually does

Most AI voiceover work today is two jobs. A text-to-speech tool reads the script, a music tool or a stock library supplies the bed, and someone lines the two up in an editor and fixes the levels. Suno Speech does both in one pass. Because the score is generated together with the voice, it can swell when the delivery gets dramatic and drop back when the narrator slows down, rather than playing underneath at one level.

Suno pitches it with playful examples. The release note offers “bedtime stories over soft piano, hype speeches over stadium drums, ASMR grocery lists and more.” In the launch post Brody lists what the team made while building it: friends’ texts turned into “wildly overproduced dramatic readings”, ordinary voice notes with “unnecessarily epic scores”, meditations, poems, pep talks and bedtime stories for their kids. Suno’s tutorial adds more practical uses: a museum audio guide, a train station announcement, a spoken-word piece and a study explainer.

It is also not Suno’s first voice model. Before it went all-in on songs, Suno released Bark, an open-source text-to-speech model, in 2023. Speech is the first voice model built into the Suno app itself.

How to use Suno Speech on the web

Open suno.com/create. The left-hand panel now has three tabs: Songs, Speech and Sounds. Click Speech. The panel carries a red BETA badge, and you choose between Simple and Advanced at the top.

Simple mode: describe the finished idea

The Suno Create panel with the Speech tab selected, Simple mode active, a BETA badge, and a prompt that begins: A museum audio guide explaining an obviously fake
Create → Speech → Simple. One prompt describes the content, the voice and the music.

In Simple mode you write one prompt that covers what is being said, who is saying it and what plays underneath. You do not write the script; Suno does. In the tutorial the prompt is a museum audio guide for an obviously fake exhibit about the invention of the snooze button — “calm, serious narrator, slightly dramatic, with elegant orchestral music underneath” — and Suno returns a full narration (“Welcome to Gallery 7. Please stand a respectful distance from the glass…”) with a matching orchestral score. The result in the tutorial runs a little over five minutes, and the request comes back as two takes, the same way a song generation does.

The Verge’s example prompt, “a pirate captain rallying his crew”, shows how little you need to type.

Advanced mode: your script, your direction

Suno Speech in Advanced mode with a Custom script box reading Write or paste your speech and a Style of speech box reading Describe the delivery: tone, pacing, mood and setting; two Speech takes of The Snooze Button, 5:13 and 5:19, appear in the library on the right
Advanced mode splits the prompt in two: Custom script for the words, Style of speech for the delivery.

Switch to Advanced when you already know exactly what should be said. You get two boxes:

  • Custom script — “Write or paste your speech.” This is the text Suno performs.
  • Style of speech — “Describe the delivery: tone, pacing, mood and setting.” This is where you direct the performance and, if music is on, the score.

The tutorial’s best demonstration keeps one short announcement script and changes only the style. “Calm and professional train station announcer, slow, measured pacing, completely serious delivery” produces a deadpan read of “Attention passengers, the 6:42 train to nowhere in particular is now boarding on platform 9.” Change the style to “increasingly frantic announcer trying to remain professional, starts controlled, gets faster and more exasperated with every sentence” and the same words come back as a completely different performance.

Suno’s formatting advice for the script is simple: write it like normal speech, in plain paragraphs. You do not need lyric formatting or bracketed section tags such as [Verse] or [Chorus] the way you would for a song.

The Advanced controls: Vocal Gender, Background music and Variety

The Advanced section of Suno Speech with three rows: Vocal Gender with Male and Female options, Background music with Off and On (On selected), and a Variety slider set to Normal
The three switches under Advanced. Background music is On by default.
  • Vocal Gender — Male or Female. Neither is selected in the default view, so you only need it when the gender matters.
  • Background music — On by default. Switch it Off when you need clean spoken audio: narration for a video you will score yourself, dialogue, instructions, or anything where you already have the music.
  • Variety — how much the takes are allowed to vary. It sits at Normal by default; move it left for more predictable reads, right for more surprising ones.

Suno Speech on iPhone and Android

The Suno mobile app with Speech selected at the top, a Backing music on chip, Script and Tone chips, and the prompt: I have a test tomorrow. Explain osmosis to me like a p
On mobile the same options appear as chips: Backing music, Script and Tone.

Speech is in the mobile apps too; Suno’s announcement on X tells users to “update your app for the latest” if it does not appear. In the app you pick Speech from the menu at the top of Create, and the Advanced fields become chips: Backing music on (tap the × to remove it), Script and Tone. The tutorial’s mobile example is a study aid — “I have a test tomorrow. Explain osmosis to me like a patient tutor. Start with a simple analogy, then give me a short explanation I could remember for the test” — which turns revision notes into something you can listen to on the way to class.

Suno Speech prompt examples

These follow the pattern Suno uses in its own examples: say what the piece is, then the voice, then the music.

Bedtime story (Simple mode)

A two-minute bedtime story about a fox who learns to share fireflies. Soft, warm storyteller voice with a gentle smile in it, slow pace. Quiet piano and soft night pads underneath, lullaby tempo.

Hype speech (Simple mode)

A coach’s half-time pep talk to a team that is losing 2–0. Loud, rallying, building to a shout. Stadium drums and brass that grow with every line.

Your own script (Advanced mode)

Custom script: paste the exact words in short paragraphs.

Style of speech: Warm adult narrator, intimate and conversational, steady pace, clear pronunciation. Sparse felt piano underneath, voice forward in the mix, gentle ending.

Clean voiceover, no music (Advanced mode)

Background music: Off. Variety: far left.

Style of speech: Neutral, friendly product-demo narrator, medium pace, clear diction. No background music.

Tips for better results

  • Direct the delivery, not just the content. Tone, pacing, mood and setting are the four things Suno’s own placeholder asks for. “Starts calm, speeds up, ends breathless” is more useful than “dramatic”.
  • Keep scripts speech-shaped. Short sentences, punctuation where a speaker would breathe, paragraph breaks between thoughts. Rhymes and repeated refrains can pull the model toward a rhythmic, half-sung read.
  • Test the style on a short script first. The tutorial’s train announcement shows that the same words can sound completely different; settle the style on a few lines before you paste a five-minute script.
  • Listen against your text. Check names, numbers and any word that has to be exact before you publish. This is a creative tool, not a word-perfect text-to-speech engine.
  • For voice only, stack the signals. Background music Off is the main switch. An early community tip, reported by aireiter, adds moving Variety fully left and writing “No background music.” in the style or tone field, because music occasionally crept in anyway. Treat that as a workaround, not a guarantee.
  • Do not imitate real people. Describe a voice by age, texture, accent and attitude rather than naming someone whose voice you do not have permission to use.

What Suno says the beta still gets wrong

Suno is unusually frank about this. “Beta really does mean beta,” Brody writes. “Occasionally, British accents can wander off to Australia and back. Dramatic pauses may be very dramatic.” Expect a share of takes where the accent drifts mid-script, a pause runs long, or a word is read differently from how you meant it. Generating two takes and picking the better one is part of the workflow.

A few things Suno has not published yet:

  • Credit cost. Neither the blog post nor the release note gives a price per take or says whether Free and paid plans get different limits. The Create screen shows your balance; check it before a batch of long scripts.
  • Languages. No supported-language list has been published.
  • Your own voice. Suno’s separate Voices feature lets you sing with your own verified voice in songs. Nothing Suno has published says a saved Voice can be used as a Speech narrator.
  • An API. Speech is in the website and the apps only, like the rest of Suno. Our Suno API explainer covers what that means for developers.
  • Speech-specific rights. Suno has not issued separate terms for Speech. Its general rule is that commercial use comes with output made on a paid plan, and Free-plan output is for personal use; our legal guide goes through the details.

Suno Speech vs a text-to-speech tool

Suno Speech and tools like ElevenLabs solve different problems. Pick Suno Speech when the performance and the music belong together: bedtime stories, poems, meditations, trailers, toasts, character pieces, dramatic readings of the group chat. Pick a dedicated TTS tool when you need the same voice across hundreds of clips, exact timing, word-perfect pronunciation or an API — product tutorials, audiobooks, accessibility, apps. The two can also work together: draft the scored version in Suno, and keep a dry TTS read for anything that has to be precise.

What people are saying on X

Suno’s own announcement on X was short: “Speech is now in Beta. Try the first model ever that creates spoken audio with matching background music. Update your app for the latest.” AI news accounts picked it up within hours; the Chinese-language AI commentator Gorden Sun (@Gorden_Sun) summed it up as a natural step for a company that already generates sung vocals, with the selling point that “the generated speech fits the background music perfectly.”

Not everyone was convinced. On r/SunoAI one user called Suno’s own demo “worryingly bad” (quoted in the aireiter guide linked above), and The Verge pointed out that AI speech is a crowded field, from ElevenLabs to Adobe, and read the launch partly as Suno diversifying beyond music. On the same day, CEO Mikey Shulman told Bloomberg that Suno is now “far beyond” the 2 million subscribers and $300 million in revenue it last reported, though he gave no new figure.

Downloading Suno Speech tracks

Speech takes appear in your workspace next to your songs, labelled SPEECH, with the same like, pin and share buttons. To keep a file, use Suno’s own download on your plan — the download limits that apply to songs are the ones to plan around until Suno says otherwise.

Suno Music Downloader works with public Suno links. If you publish a Speech track and its share link opens a normal suno.com/song/… page, paste it in and you can save a listening copy as M4A, MP3, WAV or MP4 for personal use, without signing in to Suno and without using your download quota. Speech is still in beta, so if a Speech link does not load in the tool, tell us and we will look at it.

Frequently asked questions

What is Suno Speech? Suno Speech is a beta feature in Suno’s Create page that generates spoken audio — narration, voiceovers, stories, speeches — together with original background music in a single track. Suno launched it on October 1, 2026 and calls it the first audio model that makes voice and music together.

Is Suno Speech free? Suno says the beta is open to everyone on web and mobile, but it has not published how many credits a Speech take costs or whether limits differ by plan. Your credit balance is shown on the Create screen.

How do I use Suno Speech? Open Create on suno.com or in the app and choose Speech. In Simple mode, describe the piece, the voice and the music in one prompt. In Advanced mode, paste your own text into Custom script and describe the delivery in Style of speech, then set Vocal Gender, Background music and Variety if you need to.

Can Suno Speech make a voiceover without music? Yes. Turn Background music off in Advanced mode, or remove the Backing music chip on mobile. Some early users add “No background music.” to the style and lower Variety as well, because music occasionally slipped in during the beta.

How long can a Suno Speech track be? Up to about eight minutes per take, according to Suno’s tutorial.

Can I use my own voice in Suno Speech? Suno has not said so. Personal voices are part of Suno’s separate Voices feature for songs, and nothing Suno has published says a saved Voice can narrate a Speech track.

Does Suno Speech have an API? No. Like the rest of Suno, Speech is available only on the website and in the iOS and Android apps.

Is Suno Speech good enough to replace text-to-speech? Not for everything. It is strongest for short, expressive pieces where music belongs in the result. For long, word-perfect or repeatable narration, a dedicated text-to-speech tool is still more predictable, and Suno itself warns that accents and pauses can drift in the beta.

Can I download Suno Speech tracks with Suno Music Downloader? The tool reads public Suno links. A published Speech track that opens as a normal suno.com/song page can be saved like a song; Speech is new and in beta, so tell us if a Speech link does not work.

We will update this article when Suno publishes pricing, languages or other changes to Speech.

← All posts