跳到主内容
@wquguru
精选72Bloomberg Television(YouTube)研究与分析

彭博:中国AI模型在成本与使用量上领先美国

Chinese AI Models Gain Ground on Price and Use

原文
发到 X

Well, it's China versus The United States in the battle for AI dominance, and China is closing the gap. In fact, when it comes to usage and cost, China is now in the lead. Of course, this is causing concern in Washington, Wall Street, and in Silicon Valley where US companies are, including and OpenAI, are chasing trillion dollar valuations. Luz Ding is on a team of Bloomberg reporters who have done a deep dive into this fierce competition from The US and China on AI.

好吧,在争夺人工智能主导权的较量中,是中国与美国对决,而中国正在缩小差距。事实上,在使用量和成本方面,中国现已领先。当然,这引发了华盛顿、华尔街和硅谷的担忧,那里的美国公司(包括 OpenAI)正追逐着万亿美元估值。彭博社记者 Luz Ding 所在团队深入调查了美中两国在人工智能领域的激烈竞争。

Luz joins us now. I wanna do justice to all the data driven research you guys did. It's really, really cool. Presented as well. But first, I wanna start with this chart, which basically shows how close these two countries are in equivalencies with these AI models. I know you broke it down into different sectors and different types and what which country is better at, but can you just give us the top lines what you found when you dug into capacity and usage of these models and cost?

Luz 现在加入我们。我想充分展现你们所做的所有数据驱动研究,真的非常酷,呈现得也很好。但首先,我想从这张图表开始,它基本上展示了这两个国家在这些 AI 模型等效性方面的接近程度。我知道你将数据按不同领域和不同类型进行了细分,并指出了哪个国家在哪方面更擅长,但你能否简要总结一下,当你深入研究这些模型的能力、使用量和成本时,你发现的主要结论是什么?

I know cost is a really big issue. Yes. Of course. Thanks for having me. So, so, yeah, as we can see from the graph that you just showed before, that's the gap between The US and Chinese top models have been narrowing. It was, it it was narrowed from over a year to now just a few months, especially after the recent breakthrough from China's moonshots k m k three model release in July. We are also seeing the usage of Chinese models around the world's increasing.

我知道成本是一个非常重要的问题。是的,当然。感谢邀请我。正如我们刚才看到的图表所示,中美顶级模型之间的差距正在缩小。这一差距已从一年多前缩短到现在仅几个月,尤其是在中国 Moonshot(月之暗面)发布的 K3 模型于七月取得突破之后。我们也看到全球范围内对中国模型的使用量正在增加。

One of the data points that we use was from Open Router, one of the major AI access platforms. We're seeing the use of Chinese models reached over 60% in July. And the and in the month before, the use of Chinese models surpassed The US models for the first time on the platform. And then it's worth noting that so you don't the users, they don't have to send data to China or Chinese models to use these Chinese to Chinese companies to use Chinese models because most of the Chinese top models are open weight, meaning that these models can be downloaded, modified, and deployed anywhere anywhere needed.

我们使用的一个数据来源是 OpenRouter,这是一个主要的 AI 访问平台。我们看到七月中国模型的使用率超过了 60%。而在前一个月,该平台上的中国模型使用量首次超过美国模型。此外值得一提的是,用户无需将数据发送到中国或使用中国公司的模型来使用这些中国模型,因为大多数中国顶级模型都是开源权重(open weight),这意味着这些模型可以在任何需要的地方下载、修改和部署。

So a lot of the use that we're talking about, especially the use of the models in The US, we're talking about Chinese models deployed in data centers in The US or Canada and run by US companies. You're putting all of these models through their paces, and and my favorite test you did was asking them to create a fictional ecommerce site called Bloomberg, a coffee site, fake ecommerce company. Just trying to explain the difference between how each of these models approach that maybe in terms of what they were able to create and also the cost associated with that when you compare kind of The US models to the Chinese models?

因此,我们讨论的许多应用场景,尤其是美国对模型的使用,指的是部署在美国或加拿大数据中心、由美国公司运营的中文模型。你们对这些模型进行了全面测试,而我最喜欢的一项测试是要求它们创建一个名为 Bloomberg 的虚构电商网站、一个咖啡网站以及一家虚假的电商公司。这是否旨在解释这些模型在处理此类任务时的差异,无论是从它们能够创建的内容来看,还是从与美国模型相比中国模型所关联的成本来看?

Yeah. So, yeah, we see a lot of benchmarks and rankings and scores, but it's difficult to see, like, how actually, they perform. So we gave, this one prompt, as you said, to build a coffee shop website to some of the leading models. And then what we find, first of all, getting a ecommerce website seems not to be a difficult task for most of the models available right now. It's not challenging for the top ones, but it's not also not challenging for the lower tier ones.

是的。确实,我们看到了大量的基准测试、排名和分数,但很难直观地了解它们的实际表现。因此,正如你所说,我们将构建咖啡店网站的提示词提供给了一些领先的模型。我们发现,首先,对于目前大多数可用的模型来说,构建一个电商网站似乎并不是一项困难的任务。这对顶级模型来说不算挑战,但对低层级模型而言也不算难。

Most of them, we have found are able to build functional websites that we need. But then now we're also seeing, for example, the prices are quite different. For example, for Anthropix model, Fable five, when we asked it to be the website, it was able to deliver a very sophisticated website within an hour, and the bill was about $50. That's already pretty good for a website. But then if you use China's Kimi k three model, then you can do the similar website for just a quarter of the price.

我们发现,大多数模型都能够构建出满足我们需求的功能性网站。但现在我们也注意到,例如价格存在显著差异。以 Anthropic 的 Claude 3.5 Sonnet 模型为例,当要求它生成网站时,它能在一个小时内交付一个非常精致的网站,账单约为 50 美元。这对于一个网站来说已经相当不错了。但如果使用中国的 Kimi K3 模型,则只需四分之一的价格即可实现类似的网站。

And then if you wanna sacrifice a little bit of the, you know, graphic design and then you wanna wait for a longer hours, you can go even below $5 for the website, and they can still work. So you're gonna have, like, this this sheen version of the website, but it sounds like the sheen version is, like, pretty okay. Yeah. Yeah. Yeah. Yeah. I think in the end, it really up to, like, how how fast you want the results in the first take and how good you want it to be and how much you wanna pay for that perfection and speed.

此外,如果你愿意牺牲一些图形设计效果,并愿意等待更长的时间,网站费用甚至可以低于 5 美元,而且它们仍然可以正常运行。因此,你会得到一个简化版的网站,但这种简化版听起来其实也还可以。是的。是的。是的。是的。我认为最终这真的取决于你想要多快的首次结果产出,你对质量的要求有多高,以及你愿意为这种完美和速度支付多少费用。

There are also places that it seems like the Chinese models are actually outdoing, The US models. Can you walk us through the different categories? You looked at coding. You looked at finance. You looked at banking support. Where are the Chinese models really excelling? Yes. So it it depends on how you use it and which industry that, you know, are more adapted to the models. We I don't think from from the graphics that you saw in our story, there are rankings of different ways.

还有一些领域似乎中国模型实际上超越了美国模型。你能带我们梳理一下不同的类别吗?你考察了编程、金融和银行支持等领域。中国模型在哪些方面真正表现出色?是的。这取决于如何使用它们,以及哪些行业更适配这些模型。我认为从我们在报道中展示的图表来看,并没有对不同方式进行的排名。

And then if I could remember it correctly, on the finance service sides, there are more Chinese models leading in the top three, whereas in some other categories, are only one Chinese models in the top top top 10. But then Chinese models, it's it's real real advantage is in the price because it can be capable enough to address a lot of the a lot of the questions and tasks that we encounter in real life and also in working places.

然后如果我记得没错的话,在金融服务方面,前三名中有更多的中国模型领先,而在其他一些类别中,前十名里只有一家中国模型。但中国模型真正的优势在于价格,因为它们的能力足以应对我们在日常生活和工作场所中遇到的许多问题和任务。

So with the capable enough capability and then also the very low price, it's being it's been it's become very irresistible for a lot of companies to to adopt, including some of the top US tech companies like Airbnb, DoorDash, and Coinbase. Lisa, I wanna go back to something that Christina mentioned at the top that this is kind of in the crucible of a couple of these AI companies poisoned to to go public, so Anthropic and and OpenAI.

因此,凭借足够强的能力和极低的价格,这对许多公司来说变得极具吸引力,包括一些顶级美国科技公司如 Airbnb、DoorDash 和 Coinbase。Lisa,我想回到 Christina 开头提到的一个话题,即这几家 AI 公司正处于冲刺 IPO 的关键阶段,比如 Anthropic 和 OpenAI。

And it makes me wonder just sort of how important this metric of being the best is. That that if you look at these in side by side, they're both quite good, the Chinese models and and The US models. How important is it from a business perspective, a financial perspective that Anthropic and OpenAI maintain this kind of vanguard position here as they go into their into their IPOs? Yeah. I mean, as of now, OpenAI and and for are still the best building the best models, most capable models, and there's still a lot of money being made by them in the current in the current market.

这让我想知道,成为“最佳”这一指标究竟有多重要。如果将它们并列比较,中国模型和美国模型都相当不错。从商业和财务角度来看,Anthropic 和 OpenAI 在进入 IPO 时维持这种先锋地位有多重要?是的。我的意思是,截至目前,OpenAI 和 Anthropic 仍然在构建最强大、能力最强的模型,并且他们在当前市场上依然赚取了大量资金。

And then also the data that we saw, they're not a full capture of the all the all of the glow global AI use because OpenRotter is just one platform. There are a lot of other platforms that offer these services and the use of Anthropic and OpenAI's API the the use of Anthropic and OpenAI's service is not trackable because it's with their company. And then I would say that it's hard to say, like, how how, for example of course, if you were the best in the category, you are always you're attracting more investors.

此外,我们看到的数据并不能完全反映全球所有 AI 的使用情况,因为 Grok(注:原文为 OpenRotter,疑为 Grok 或类似平台名称的误写,此处保留原意指代单一平台)只是其中一个平台。还有许多其他平台提供此类服务,且由于 Anthropic 和 OpenAI 的服务是封闭在公司内部的,其使用情况无法被追踪。因此我认为很难断言,例如,当然如果你在某个类别中是最优秀的,你自然会吸引更多投资者。

But at the same time, there are also another market that's to be tapped is that those companies in smaller companies or companies in developing countries, they may not be able to pay that high price that Anthropic and OPIA are charging. And then now they look at and then they can choose Chinese models where which which are, like, increasingly capable, and then they offer much lower price. And they they are open ways so you can deploy it locally.

但与此同时,另一个有待挖掘的市场是:那些小型公司或发展中国家的公司可能无法支付 Anthropic 和 OpenAI 所收取的高昂费用。于是他们转而审视并可以选择能力日益增强的中国模型,这些模型提供低得多的价格,并且以开源方式提供,允许本地部署。

So that's a market that Chinese models are trying to trying to conquer, and they're already succeeding in a lot of places.

因此,这是中国模型试图征服的市场,而且它们在许多地方已经取得了成功。

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近