Suno Speech: how to generate voiceovers with AI music

Alex da Cruz
Alex da Cruz is a full-stack developer based in São Paulo, Brazil. He works with React, TypeScript and automation, and uses AI daily to solve real problems in code and operations — not as a demo. He has run an e-commerce operation end to end, and now builds and maintains the automation pipeline behind this blog. He writes about what he actually tests.
According to The Verge, Suno expanded beyond AI-generated music by launching a new Speech feature in public beta for web and mobile platforms. The tool generates spoken voices from scripts or prompt descriptions, allowing users to create voiceovers accompanied by background music.
How does Suno Speech work for creators?
The feature offers two main modes to handle different production needs. The Simple mode lets users describe the desired audio through a prompt box, while the Advanced mode accepts custom scripts for precise text delivery.
Advanced settings also provide manual control over the AI voice's gender, speech style, and variation levels. Each generated track supports a maximum duration of around eight minutes, making it suitable for short narration tasks and podcasts.
Can you remove the background music?
Yes, the music generation is entirely optional. Users can turn off the background track using a simple toggle if they only need clean synthetic speech for their projects.
Jack Brody, Suno's chief product officer, stated that the model generates voice and music together as a cohesive track. However, the company warns that the beta version still exhibits quirks, such as unpredictable accent shifts and exaggerated pauses.
Sources
- AI music maker Suno now generates spoken words — theverge.com
Frequently asked questions
- What is Suno Speech?
- Suno Speech is a new beta feature that generates synthetic spoken voices and background music together as a cohesive audio track from scripts or text prompts.
- Can I remove the background music in Suno Speech?
- Yes, you can turn off the background music using a toggle if you only need clean spoken words for your project.
- What is the maximum duration for Suno Speech?
- The Speech feature supports a maximum duration of around eight minutes per generation.
Comments
0 comments
Be the first to comment.
Continue Lendo

Claude Code Mods: Customizing the AI Tool via Code
Anthropic released Mods for Claude Code, allowing developers to customize the coding tool using JavaScript or TypeScript.

Meta Muse: Build Custom AI Hardware With Open Source Code
Meta released open source code to run the Muse AI agent on custom hardware like Raspberry Pi and ESP32 boards.

Apple Tightens macOS Full Disk Access for AI Agents
Apple is restricting macOS Full Disk Access to protect users from autonomous AI agents reading private files and messages.