Suno, the AI music generation platform, has added a new feature that enables the creation of "spoken audio" within its generative AI service. With this update, users can now create not only songs, but also voice tracks such as narrations and dialogues, combined with perfectly matched background music.
Announced today, the new capability generates natural spoken-word audio from text input while automatically composing and synthesizing background music optimized for the content and mood. In addition to traditional music generation, creators can now seamlessly produce audio assets tailored for podcasts, video content, storytelling, and other use cases.
Leveraging its foundational expertise in music generation AI, Suno produces background audio that maintains musical harmony with the intonation and rhythm of the spoken words. The company's high processing power in music generation streamlines the audio content production workflow, providing an environment where users can create audio works without professional recording equipment or editing skills.
To further support users' creative activities, Suno plans to continue rolling out updates that enhance the diversity of audio generation. This feature expansion is expected to propel the platform beyond the framework of a traditional music production tool, positioning it as a comprehensive audio media creation suite.