跳到主内容
精选75Rohan Paul模型发布/更新多源精选 ×8

Grok 4.5 成 Perplexity Computer 最强编排器,性价比超 Opus 4.8

For Perplexity Computer, Grok 4.5 became the strongest orchestrator, scoring 0.3…

原文

For Perplexity Computer, Grok 4.5 became the strongest orchestrator, scoring 0.328 (WANDR benchmark score) at $4.76 per trial.

  • Opus 4.8 (high, thinking) scored 0.254 at $9.46, despite much higher cost.
  • GPT-5.6 sol (medium) cost $2.64 but reached only 0.289 on the same test.

Perplexity Computer basically lets an orchestrator assign research, coding, browsing, and document work across subagents.

The WANDR benchmark measures difficult research jobs requiring search, computation, and sustained multi-step reasoning.

Grok 4.5 beat five tested configurations, including Opus 4.8, on Perplexity’s evaluation.

The WANDR evaluates whether an agent can complete large, professional research tasks that require searching many sources, running computations, organizing results, removing duplicates, checking evidence, and producing a complete structured answer. Perplexity describes these as “wide research” tasks involving the orchestration of search, compute, and model reasoning.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源
Grok 跻身前沿模型行列
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)原文
Grok 4.5 发布,xAI 加入竞争
向阳乔木原文

相似阅读

另一事件,读法相近