New Speech Capabilities
Suno is expanding beyond its core AI music generation tools by introducing a new feature that produces spoken words from scripts or text prompts. The speech functionality is available in public beta across the platform's web and mobile applications, enabling users to generate voiceovers alongside synchronized background music.
Suno Expands Into Synthetic Speech
The newly introduced speech tool generates synthetic voice tracks accompanied by automated background compositions.
According to the company, music remains central to the platform, but the underlying vision encompasses broader forms of human expression. The new model is designed to generate voice and music together as a single unified audio track.
Market Context
While text-to-speech technology is already established across various industry players and specialized platforms, Suno's integration aims to offer a distinct approach by combining spoken audio with automated musical arrangements.
Features and Customization
The background music accompaniment is optional and can be disabled via a toggle switch for users seeking isolated speech. This functionality is intended to support applications such as narrated poetry, dramatic voiceovers, or presentations requiring specific tonal backgrounds.
Platform Controls
Users can access the feature through the creation menu by selecting the speech option, which offers a simple mode for prompt-based generation and an advanced mode for custom script input. Additional parameters allow adjustments to voice gender, style, and variability, with generated tracks supporting durations of up to approximately eight minutes.
The company acknowledges that the beta release has ongoing limitations and plans to refine the system based on user feedback.




