跳到主内容
@wquguru
精选75Google DeepMind(YouTube)产品发布/更新多源精选 ×7

Google DeepMind发布Gemini Audio文本转语音功能

Create your own voices with Gemini 3.8 text-to-speech

原文
发到 X

I need to make a promo video for our new text to speech model. Let's start with a voiceover. Introducing text to speech. No more like this. Introducing text to speech transforms simple prompts into consistent, expressive voice outputs on demand. OK, now let's add in another voice for this dialogue scene in the script. Or choose from over 1,000 ready to go voices. OK, but what if we had that voice a little bit deeper and change it however you want.

我需要为我们的新文本转语音模型制作一个宣传视频。让我们从旁白开始。介绍文本转语音。不再像这样了。介绍文本转语音,它能将简单的提示词按需转化为一致且富有表现力的语音输出。好的,现在让我们为剧本中的这个对话场景添加另一个声音。或者从超过 1,000 种现成的声音中选择。好的,但如果我们让那个声音稍微低沉一点,并且可以随意更改它呢。

OK, perfect. Let's drop these voices into the dialogue scene. You can pick and drop your favorite saved voices, and you can direct them in a multi-speaker scene. Actually, maybe it should be my voice. The lighthouse keeper watched the silver horizon as the morning fog slowly lifted. Or recreate your own voice, safely protected by a quick verbal identity check. Give your words a voice with Gemini audio.

好的,完美。让我们把这些声音放入对话场景中。你可以挑选并放置你最喜欢的已保存声音,并且可以在多说话人场景中指导它们。实际上,也许应该用我的声音。灯塔看守人注视着银色的地平线,晨雾缓缓散去。或者通过快速的身份验证安全地复刻你自己的声音。用 Gemini Audio 为你的文字赋予声音。

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →