中国AI模型Kimmy K3震动硅谷与华盛顿
The Chinese AI Model Rattling Silicon Valley and Washington | Big Take
Bloomberg Audio Studios podcasts radio news. The battle for AI dominance between the US and China has escalated sharply over the past few weeks after the surprise release of a new Chinese AI model took the tech world and Washington by storm.
彭博音频工作室播客、广播新闻。过去几周,中美之间的人工智能主导权之争急剧升级,此前一款新推出的中国人工智能模型意外发布,令科技界和华盛顿为之震惊。
Kimmy K3 which is a very powerful model
Kimmy K3,这是一个非常强大的模型
Kimmy K3 model. This company's AI model known as Kimmy K3. For the last 3 days, I've said Kimmy K3 many times.
Kimmy K3模型。这家公司的人工智能模型被称为Kimmy K3。过去三天里,我已经多次提到Kimmy K3。
Kimmy K3 developed by Chinese startup Moonshot AI.
Kimmy K3由中国的初创公司月之暗面开发。
Kim K3 is really making waves because it stacks up so well against offerings from OpenAI and Anthropic, which are among the most powerful AI models in existence.
Kim K3之所以引起轰动,是因为它与OpenAI和Anthropic的产品相比毫不逊色,而这两家公司的模型是现存最强大的人工智能模型之一。
Mark Anderson is Bloomberg's Asia technology editor based in Hong Kong. And all of this developed by a Chinese AI company that has access to far less compute and far less resources than deep pocketed American competitors. Kimmy K3 trails only OpenAI's GPT 5.6 and Anthropics Claude Fable 5. So it's the third most powerful model on the market today. Within hours of its release, Kim K3 sent shock waves through global markets, triggering memories of the deepseek panic in early 2025.
马克·安德森是彭博社驻香港的亚洲科技编辑。而这一切都是由一家中国人工智能公司开发的,与资金雄厚的美国竞争对手相比,该公司可用的计算资源和资源要少得多。Kimmy K3仅次于OpenAI的GPT 5.6和Anthropic的Claude Fable 5。因此,它是目前市场上第三强大的模型。发布后数小时内,Kim K3就冲击了全球市场,让人想起2025年初的DeepSeek恐慌。
Remember, chip stocks fell into a bare market Friday in
还记得吗,周五芯片股跌入熊市
because investors are now really reassessing whether or not all that capex by Samsung and SKH is that really justified given that China is doing it better and cheaper. Once markets settled, the questions began about how Moonshot built Kimmy in the first place.
因为投资者现在真的在重新评估,三星和SK海力士的所有那些资本支出是否真的合理,考虑到中国做得更好且成本更低。市场稳定后,人们开始质疑月之暗面最初是如何打造出Kimmy的。
It raises the question about if this was done illegally and in a violation of those export controls.
这引发了一个问题:这是否是非法行为,是否违反了那些出口管制。
Did it use banned Nvidia chips? Did it steal proprietary data from American rival Anthropic?
它是否使用了被禁的英伟达芯片?它是否窃取了美国竞争对手Anthropic的专有数据?
This is what the director of the White House Office of Science and Technology Michael Katzios. He is saying that Moonshot allegedly developed a sophisticated internal platform to conduct large-scale distillation against US models.
白宫科技政策办公室主任迈克尔·科齐奥斯就是这样说的。他表示,月之暗面涉嫌开发了一个复杂的内部平台,对美国模型进行大规模蒸馏。
And most importantly, what is yet another lowcost Chinese AI model mean for the race between the US and China for global AI supremacy?
最重要的是,又一个低成本的中国人工智能模型对美中全球人工智能主导权之争意味着什么?
I think the K3 moment tells us that the race is much closer than people previously thought. I think that China is about 6 months behind the US.
我认为K3时刻告诉我们,这场竞赛比人们之前认为的要接近得多。我认为中国落后美国大约6个月。
That's not a lot of time.
这时间并不长。
Which is not a lot of time. But it's also bringing a challenge to this perception that only American AI, only proprietary AI can be frontier, can be the most cutting edge AI. It's challenging this idea that America owns the most advanced AI models in a really exciting way. Welcome to the Big Take Asia from Bloomberg News. I'm Juan Ha. Every week, we take you inside some of the world's biggest and most powerful economies and the markets, tycoons, and businesses that drive this evershifting region.
这不是很多时间。但它也挑战了一种观念,即只有美国的人工智能,只有专有的人工智能才能成为前沿,才能是最尖端的人工智能。它以一种令人兴奋的方式挑战了美国拥有最先进人工智能模型的想法。欢迎收听彭博新闻的《Big Take Asia》。我是Juan Ha。每周,我们带你走进世界上一些最大、最强大的经济体,以及推动这个不断变化的地区的市场、大亨和企业。
Today on the show, what we know about Moonshot's Kimmy model, what the company's stunning rise means for Silicon Valley's tech giants, and the highstakes AI arms race between the US and China. Kimmy K3's launch didn't just rattle markets. Overnight, it became clear that China is intent on rewriting the rules of the global AI race. It's also upending long-held assumptions about how this technology will evolve. The prevailing view in the AI industry used to be that the companies with the biggest budgets, largest data centers, and access to the most advanced chips would lead.
今天节目中,我们将了解月之暗面的Kimmy模型,该公司惊人的崛起对硅谷科技巨头意味着什么,以及中美之间高风险的AI军备竞赛。Kimmy K3的发布不仅震动了市场。一夜之间,很明显中国意图改写全球AI竞赛的规则。它也颠覆了关于这项技术将如何发展的长期假设。AI行业的主流观点曾经是,拥有最大预算、最大数据中心和最先进芯片的公司将领先。
But with the release of DeepSeek's R1 model last year and Moonshot AI's latest offering, that narrative is being challenged. I want to get just a little geeky with you. Can you take us under the hood about this model and what's special about it?
但随着去年DeepSeek的R1模型和月之暗面AI最新产品的发布,这种说法受到了挑战。我想和你稍微深入一点。你能带我们看看这个模型的内部,它有什么特别之处吗?
Sure, Geeky is good. Um, basically it's got a really huge brain. So, by brain we mean parameters. It's got 2.8 trillion parameters.
当然,深入一点很好。嗯,基本上它有一个非常大的大脑。所以,我们说的“大脑”是指参数。它有2.8万亿个参数。
And what are parameters?
什么是参数?
I would describe parameters as the size of the model's brain. So, it's essentially the amount of information that is stored inside an artificial intelligence model. The bigger the brain, the more information that is stored in it. the more information can be processed, you know, it's it's really as simple as bigger is better and 2.8 trillion is certainly a very large number, which is essentially on par with the latest offerings from OpenAI and Anthropic.
我会把参数描述为模型大脑的大小。所以,它本质上是存储在人工智能模型中的信息量。大脑越大,存储的信息越多,能处理的信息也越多,你知道,这真的很简单,越大越好,2.8万亿当然是一个非常大的数字,这基本上与OpenAI和Anthropic的最新产品相当。
Kim K3 isn't just notable for its brain size. It's also remarkable for the way it's being shared. While models like GPT and Claude largely operate behind closed doors, K3 was released to the public last month as an openweight model, meaning anyone, a researcher, a startup, a government can access it.
Kim K3不仅因其大脑大小而引人注目。它的分享方式也值得注意。虽然GPT和Claude等模型在很大程度上是闭门运作的,但K3上个月作为开放权重模型向公众发布,这意味着任何人——研究人员、初创公司、政府——都可以访问它。
So basically, an openweight model can be downloaded and customized. Users can look and see essentially what I would describe as the recipe of the model. It can see how the model works at a level that proprietary AI doesn't offer. Is that the same as open source?
所以基本上,开放权重模型可以被下载和定制。用户可以看到,本质上我会描述为模型的配方。它可以看到模型在专有AI不提供的层面上是如何工作的。这和开源一样吗?
Open source goes one step further. Open source, I would describe it as a kitchen. Open source allows you to see all of the steps behind the recipe that the entire process by which the model was built. There are other openweight AI models on the market like those from Deepseek. But right now, Kim K3 is the biggest and size isn't its only selling point. Kim K3 comes with something called a massive context window. A feature Mark says is becoming increasingly important as AI models take on more complex tasks.
开源更进一步。开源,我将其描述为一个厨房。开源让你能看到食谱背后的所有步骤,即构建模型的整个过程。市场上还有其他开放权重的AI模型,比如Deepseek的那些。但目前,Kim K3是最大的,而且尺寸并不是它唯一的卖点。Kim K3配备了一个所谓的巨大上下文窗口。Mark表示,随着AI模型承担更复杂的任务,这一功能正变得越来越重要。
The context window is essentially the temporary memory of an AI model. So it's the AI model's ability to retrieve information in a timely manner. I think we've all probably used Gemini or Chat GPT and tried to ask for things like give me this recipe or what time should I put my toddler to bed. But when it comes to running a really complex task like build an entire app from scratch, it can take much, much longer. Most AI models have limited working memory.
上下文窗口本质上是AI模型的临时记忆。所以它是AI模型及时检索信息的能力。我想我们可能都用过Gemini或ChatGPT,并尝试询问诸如“给我这个食谱”或“我应该什么时候让我幼儿上床睡觉”之类的问题。但当涉及到运行一个真正复杂的任务,比如从头构建一个完整的应用程序时,可能需要更长的时间。大多数AI模型的工作记忆有限。
As a conversation or document gets longer, they can start to lose track of what came earlier. Imagine trying to talk to someone who can only remember the last 10 sentences you said. Kimmy K3, on the other hand, has a context window big enough to process nearly 800,000 words in a single prompt. That's like several fulllength novels at once. This makes it extremely powerful for long legal documents, code bases, or research papers.
随着对话或文档变长,它们可能会开始忘记之前的内容。想象一下,试图与一个只能记住你最后10句话的人交谈。另一方面,Kimmy K3拥有一个足够大的上下文窗口,可以在单个提示中处理近80万词。这相当于同时处理几部长篇小说。这使得它在处理长法律文件、代码库或研究论文时极为强大。
And K3 isn't just powerful, it's also remarkably cheap to run. According to a benchmarking site that is widely regarded throughout the industry called artificial analysis, the cost of running Kimmy K3 is far cheaper than running the latest offering from Anthropic or OpenAI. The most expensive model at the moment is Anthropic's Clawed Fable 5 uh which costs about $2.75 per task according to artificial analysis. If you look at Kimmy K3, I think it's about a dollar, something like 95, but it offers a very similar performance, and I think that is a key reason for its popularity.
而且K3不仅强大,运行成本也异常低廉。根据业内广泛认可的基准测试网站Artificial Analysis的数据,运行Kimmy K3的成本远低于运行Anthropic或OpenAI的最新产品的成本。目前最昂贵的模型是Anthropic的Clawed Fable 5,根据Artificial Analysis的数据,每个任务成本约为2.75美元。如果你看看Kimmy K3,我认为大约是一美元,差不多95美分,但它提供了非常相似的性能,我认为这是它受欢迎的关键原因。
K3's release has thrust its parent company, Moonshot AI, and its founder, Yang Chilling, into the spotlight. When open AI shocked the world with chat GPT, everyone was looking to Yang and Moonshot as China's best chance at matching that technology. And really, he was seen as the poster boy of Chinese AI up until Deepseek shocked the world last January. And basically, ever since then, he's toiled in their shadow. But this release of Kimmy K3 has brought him back to prominence.
K3的发布使其母公司Moonshot AI及其创始人杨植麟成为焦点。当OpenAI用ChatGPT震惊世界时,所有人都将杨植麟和Moonshot视为中国匹配该技术的最佳机会。确实,在Deepseek去年一月震惊世界之前,他一直被视为中国AI的代言人。基本上,从那时起,他就一直在他们的阴影下辛勤工作。但这次Kimmy K3的发布让他重新回到了显要位置。
It's brought him back to the forefront of the conversation about Chinese AI. Now, the last time we saw this level of attention and certainly headlines about a Chinese AI, it was Deepseek in 2025. Is this another deepseek moment? You know, deepseek 2.0 for China.
这让他重新回到了关于中国人工智能的讨论前沿。现在,我们上一次看到这种程度的关注,以及关于中国人工智能的头条新闻,是在2025年的Deepseek。这是另一个Deepseek时刻吗?你知道,对中国来说,这是Deepseek 2.0。
For me, it's not. For me, the deepseek moment was really establishing that China was much much closer than 5 years behind the US in the global technology race. You know, before Deep Seek's R1 model was released last January, people really did believe that the US had a yearslong head start in AI and Deepseek completely upended that perception, cut it down to just a few months. I think we are now in a place with Moonshot where people are just expecting more of these breakthroughs and it's not as surprising as the Deep Seek moment was.
对我来说,这不是。对我来说,Deepseek时刻真正确立的是,中国在全球科技竞赛中远远领先于美国,而不是落后5年。你知道,在Deep Seek的R1模型去年1月发布之前,人们确实认为美国在人工智能方面领先多年,而Deepseek完全颠覆了这一认知,将其缩短到只有几个月。我认为我们现在处于Moonshot的阶段,人们只是期待更多这样的突破,而这并不像Deepseek时刻那样令人惊讶。
But it is certain
但这是肯定的。
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力