Grok 4.5 成 Perplexity Computer 最强编排器,性价比超 Opus 4.8
For Perplexity Computer, Grok 4.5 became the strongest orchestrator, scoring 0.3…
For Perplexity Computer, Grok 4.5 became the strongest orchestrator, scoring 0.328 (WANDR benchmark score) at $4.76 per trial.
- Opus 4.8 (high, thinking) scored 0.254 at $9.46, despite much higher cost.
- GPT-5.6 sol (medium) cost $2.64 but reached only 0.289 on the same test.
Perplexity Computer basically lets an orchestrator assign research, coding, browsing, and document work across subagents.
The WANDR benchmark measures difficult research jobs requiring search, computation, and sustained multi-step reasoning.
Grok 4.5 beat five tested configurations, including Opus 4.8, on Perplexity’s evaluation.
The WANDR evaluates whether an agent can complete large, professional research tasks that require searching many sources, running computations, organizing results, removing duplicates, checking evidence, and producing a complete structured answer. Perplexity describes these as “wide research” tasks involving the orchestration of search, compute, and model reasoning.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力