精选85Chubby♨️模型发布/更新多源精选 ×7
GLM-5.3 Flash发布:320B MoE仅18B激活
GLM-5.3 Flash ("Ox Alpha") official: Benchmarks attached. This looks exceptional…
推荐理由
做Agent和编程工具的同学必看,GLM-5.3 Flash用18B激活参数在多项Agent基准上追平甚至超过Claude Opus 4.8,服务成本还低一个量级,值得立刻拿你的任务实测一遍。
GLM-5.3 Flash ("Ox Alpha") official: Benchmarks attached. This looks exceptional for its size!
GLM-5.3-Flash might be one of the most impressive efficiency releases yet.
It is a 320B MoE with only 18B parameters active per token, yet Zai reports:
- 84.3 on Terminal-Bench 2.1, nearly matching Claude Opus 4.8 at 85.0 - 63.4 on DeepSWE, ahead of Opus 4.8 and DeepSeek V4 Vision Exp - 48.8 on AutomationBench, ahead of Opus 4.8 and GPT-5.6 Terra - The highest GDPval-AA v2 score in its comparisonIt also beats the much larger GLM-5.2 across all six reported benchmarks while costing one-tenth as much to serve.
Open weights, MIT licensed, natively multimodal, 1M context.
Important caveat: 18B active parameters does not make it a normal local 18B model. All 320B weights still need to be stored. But in terms of intelligence per active parameter, this looks exceptional!
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力