跳到主内容
精选85Chubby♨️模型发布/更新多源精选 ×7

GLM-5.3 Flash发布:320B MoE仅18B激活

GLM-5.3 Flash ("Ox Alpha") official: Benchmarks attached. This looks exceptional…

原文
推荐理由

做Agent和编程工具的同学必看,GLM-5.3 Flash用18B激活参数在多项Agent基准上追平甚至超过Claude Opus 4.8,服务成本还低一个量级,值得立刻拿你的任务实测一遍。

GLM-5.3 Flash ("Ox Alpha") official: Benchmarks attached. This looks exceptional for its size!

GLM-5.3-Flash might be one of the most impressive efficiency releases yet.

It is a 320B MoE with only 18B parameters active per token, yet Zai reports:

  • 84.3 on Terminal-Bench 2.1, nearly matching Claude Opus 4.8 at 85.0 - 63.4 on DeepSWE, ahead of Opus 4.8 and DeepSeek V4 Vision Exp - 48.8 on AutomationBench, ahead of Opus 4.8 and GPT-5.6 Terra - The highest GDPval-AA v2 score in its comparisonIt also beats the much larger GLM-5.2 across all six reported benchmarks while costing one-tenth as much to serve.

Open weights, MIT licensed, natively multimodal, 1M context.

Important caveat: 18B active parameters does not make it a normal local 18B model. All 320B weights still need to be stored. But in terms of intelligence per active parameter, this looks exceptional!

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近