跳到主内容
精选75The Decoder(RSS)模型发布/更新多源精选 ×8

智谱发布GLM-5.3-Flash:320B参数,成本降七倍,全中文芯片推理

GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia

原文
推荐理由

关注国产算力与模型性价比的从业者必看:GLM-5.3-Flash 用七分之一成本逼近旗舰性能,且全链路跑在国产芯片上,建议对比自家推理成本与硬件选型。

Z.ai releases GLM-5.3-Flash, an open-source model with 320 billion parameters that lands just three points behind the larger GLM-5.3 on Artificial Analysis's Intelligence Index, at a seventh of the cost. What's notable is that all of the inference traffic ran on Chinese AI chips instead of Nvidia hardware.

Z.ai发布了GLM-5.3-Flash,这是一款拥有3200亿参数的开源模型,在Artificial Analysis的智能指数上仅比更大的GLM-5.3落后三分,而成本仅为后者的七分之一。值得注意的是,所有推理流量均运行在中国AI芯片上,而非Nvidia硬件。

The article GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia appeared first on The Decoder.

文章《GLM-5.3-Flash以极低成本媲美顶级模型,且无需Nvidia》最初出现在The Decoder上。

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源
GLM 5.3 Flash 18B 激活参数 MoE
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)原文
智谱确认匿名模型Ox Alpha为GLM-5.3-Flash
华尔街见闻(RSS)原文
智谱发布 GLM-5.3-Flash 模型
Hacker News Best(web_list)原文
Ox Alpha 实为 GLM-5.3-Flash
Przemek Chojecki | PC原文
GLM-5.3-Flash 上线 OpenRouter
OpenRouter原文

相似阅读

另一事件,读法相近