The new feature allows users to generate voiceovers from scripts or simple text prompts while simultaneously layering in musical accompaniment. According to chief product officer Jack Brody, the company intends to treat speech as a core component of human expression, positioning the model as a way to create cohesive audio tracks rather than just isolated voice files. Users retain control over the output, with a toggle available to strip away the background music for those requiring clean speech.
Suno enters text-to-speech market with integrated AI audio
Suno is pivoting beyond pure melody, launching a public beta that generates spoken word alongside background music. By integrating voice synthesis directly into its web and mobile platforms, the company aims to move past its music-only roots to capture the expanding market for automated audio production and narration.

While Suno is a newcomer to the crowded speech synthesis field—entering a space already occupied by established players like ElevenLabs and Adobe—the move serves as a strategic attempt to diversify its offerings. The platform has faced significant legal pressure regarding its primary music generation tool, and expanding into speech provides a new utility for content creators. Whether intended for poetic recitations or energetic dramatic performances, the tool is now accessible through the platform’s 'Create' tab.




Comments (0)
No comments yet. Be the first!