智谱发布GLM-5.3-Flash:320B参数,成本降七倍,全中文芯片推理
GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia
关注国产算力与模型性价比的从业者必看:GLM-5.3-Flash 用七分之一成本逼近旗舰性能,且全链路跑在国产芯片上,建议对比自家推理成本与硬件选型。
Z.ai releases GLM-5.3-Flash, an open-source model with 320 billion parameters that lands just three points behind the larger GLM-5.3 on Artificial Analysis's Intelligence Index, at a seventh of the cost. What's notable is that all of the inference traffic ran on Chinese AI chips instead of Nvidia hardware.
Z.ai发布了GLM-5.3-Flash,这是一款拥有3200亿参数的开源模型,在Artificial Analysis的智能指数上仅比更大的GLM-5.3落后三分,而成本仅为后者的七分之一。值得注意的是,所有推理流量均运行在中国AI芯片上,而非Nvidia硬件。
The article GLM-5.3-Flash matches top models at a fraction of the cost, and runs without Nvidia appeared first on The Decoder.
文章《GLM-5.3-Flash以极低成本媲美顶级模型,且无需Nvidia》最初出现在The Decoder上。
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力