we_are_coded.by CODE · The world, decoded
БГ
Suno

Suno launches Speech: spoken word and background music in one track, in beta for everyone

SunoAudio

You type a text, describe the voice and the musical style, and get it read aloud with its own soundtrack. Suno calls it the first audio model that makes voice and music together as one piece. The beta has been open to the whole community since 1 October.

In short
  • The input is text: an idea, a poem or something you wrote. The model speaks it and composes the music under it in one track.
  • Suno listed the weak spots itself: British accents sometimes wander off to Australia, and dramatic pauses are very dramatic.
  • No languages and no price in the post. It was tested for a month with a small group before opening.
Checked on2 October 2026Responsible editorTsvetelin IvanovHow we workMethod · Corrections

A voice note with an unnecessarily epic score under it. That is how Suno tested it, in its own words: turning friends' texts into wildly overproduced dramatic readings, making meditations, poems and bedtime stories for their kids. Out of that game, today, comes a product, open to everyone in beta.

The facts: on 1 October 2026 Suno announced Speech (beta), signed by Jack Brody, Chief Product Officer. The model generates spoken audio with original background music as one track. It is built into Suno: you type an idea, a poem or something you have written, then describe the voice and musical style. Per Suno, it is "the first audio model that generates voice and music together as one cohesive track". It was tested for a month with a small group of users; the beta is now open to everyone. The company lists the weak spots itself: British accents can occasionally "wander off to Australia and back", and dramatic pauses may be very dramatic. Languages and price are not stated.

I know this effect from live shows. The same text read over silence and read over the right music are two different texts, and the second one holds people to the end. Until now that took two steps and two tools: read in one, compose in the other, then glue them and argue with the levels. Here it is one text field, and the music arrives with the words.

For a podcast intro, a meditation, a birthday poem or a bedtime story this is more than enough. For an ad that will go on air, I do not see it yet, and Suno itself does not claim it. Beta really does mean beta, they write, and they prove it in the next paragraph.

The music under the words decides whether you listen to the end.

The best part of the post is the paragraph on weaknesses. A company that writes, itself, that its accent drifts to Australia and back is telling you exactly where the edge of the product is. Give it something of your own and listen to the pauses. That is where you will hear how far it still has to go.

The visual is generated code art. No third-party images.
Follow usFacebookLinkedIn
Official primary sources
→Suno - Introducing Speech (beta), 01.10.2026
Original: https://wearecoded.com/en/articles/suno-speech-beta-glas-s-muzika.html
ShareFacebookXLinkedInTelegramWhatsApp
← Back to all news