Dari 发布 Terminal-Bench 2.1 得分79.8%
79.8% on Terminal-Bench 2.1 for $76 in total inference cost.
79.8% on Terminal-Bench 2.1 for $76 in total inference cost.
That's the number Dari @daridotdev is launching with today. They open-sourced the router model behind it: a small fine-tuned model that decides which model handles each step, sending most of them to cheaper options and only calling the expensive frontier models when they actually change the result. It factors the cache into that decision too, so it won't switch models when the cached path is the cheaper win. On the same 89-task suite, the frontier setups run into the thousands.
Runs inside Claude Code, Codex, and Pi, you can bring your own Anthropic and OpenAI subscriptions, and the weights are on Hugging Face.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力