Suno opens its Speech feature in beta, generating voice and music in one track
AI music platform Suno extends from songs into spoken audio, generating voiceovers and soundtracks together; quality remains unproven.
AI music platform Suno has launched Speech, now in public beta across its web and mobile apps, generating spoken voiceovers from scripts or prompted descriptions.
The differentiator is that voice and background music are generated together as one cohesive track. Chief product officer Jack Brody called it the first audio model to combine voice and music in a single output. The music can be switched off with a toggle for clean speech.
Generation caps at about eight minutes, and an Advanced mode accepts custom scripts with controls for voice gender and style. The company concedes the feature is far from perfect — accents wander and pauses misfire — and says it will improve Speech from user feedback.
Sources:https://suno.com/blog/introducing-speech-betahttps://www.theverge.com/ai-artificial-intelligence/1003925/suno-speech-ai-voice-feature-beta-availability