Suno's new feature generates speech with matching music, training method still undisclosed
Suno released the Speech beta on October 1, generating spoken audio and matching background music together in a single track.
Users type an idea or text and describe the voice and music style; the model produces both the voice and the score. The company says it suits poems, meditations and bedtime stories, and admits the beta still has bugs — a British accent can occasionally wander off to Australia.
Suno has not said how it trained the model. The company is being sued by major record labels, and a Munich court recently rejected its fair use defense.
Quellen:suno.com