Cartesia Sonic 3.6 声音克隆实测:15秒样本即可生成多语种语音
作者在 Cartesia playground 用一段约15秒的音频克隆了自己的声音,几秒钟即完成,并让克隆声音说出他不会说的日语。该功能运行在 Cartesia 最新的语音模型 Sonic 3.6 上,支持44种语言,作者认为这项技术可用于让内容触达更多语言的受众,试用地址为 play.cartesia.ai,模型介绍见 cartesia.ai/blog/sonic-3.6。
This is impressive!
I cloned my voice in the @cartesia playground from a ~15-second clip. The clone was ready in a couple of seconds.
Then I had it speak Japanese, a language I don't speak.
My clone now says in Japanese that I can find good papers by combining AI with my experience reviewing papers.
I am impressed by how good this sounds.
This runs on Sonic 3.6, their latest voice model, which supports 44 languages. You can do the same with any of them.
This tech is getting really good, and I think it's worth exploring to make your content more accessible to more people in different languages.
Try it here: play.cartesia.ai
More info about their model here: cartesia.ai/blog/sonic-3.6
来源:omarsar0 · x.com