精选70OpenRouter模型发布/更新
OpenRouter在DRACO深度研究基准测试中表现
We ran it on the DRACO deep research benchmark by Perplexity: 100 deep research…
We ran it on the DRACO deep research benchmark by Perplexity: 100 deep research tasks across 10 domains, from law and medicine to finance and product comparison.
Each task is graded against ~39 weighted criteria, and wrong answers carry negative weight. (You can't bluff your way to a high score by being verbose.)
https://t.co/RIiTdy1pFP
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力