OpenAI 展示 GPT-Live 语音模型抗噪能力
Background Robustness with GPT-Live
Voice models are new interface for
[music]
lots of new things and they're only super useful if they understand context well and let you interact with them in a very natural way.
[music]
Our new voice model chooses from the context whom or what to focus on and provides response directly [music] to that.
The model actually acquires a proper understanding of what a human conversation is.
So, one of the big challenges with voice models in the past has been their ability to understand [music] who is talking with them and who is not. It's often referred to as the cocktail party problem where [music] at a party you have trouble focusing on a single person talking because of all the noise of the other voices.
Do you want to give this this a try for a demo?
Yeah, definitely.
Let's do it.
Hey Chat, where are some good places in the city to view fireworks?
Sure, if you mean San Francisco.
Awesome. I actually live in Noe Valley and don't want to go that far. Are there good places to view fireworks that are closer to Noe Valley, maybe walkable?
If you stay near Noe Valley, you've got a couple solid nearby options. Twin Peaks is
Oh, I love Twin Peaks. That's That's close enough. I'll I'll go ahead out there. Thanks, Chat.
Okay.
Great. Have a good one.
One of the really nice things here is like how close we're standing to you and the model understands it.
I think it's particularly apparent with being able to interrupt the model even in this loud environment and it adjusts to the new things you're saying to it very quickly.
The back and forth [music] dynamics are really good.
With a new model like this, the number of use cases, the places where you can apply it, the number of situations where it can function is like greatly increased.
Yeah, super excited [music] for this to get out there in the world.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力