Grok 4.7法律与终端任务大幅跃升,定价不变
Legal work saw one of Grok 4.7’s biggest jumps.
Grok 4.7在法律代理和长程终端任务上实现大幅跃升,且维持原有低价策略,对开发者选型有直接参考价值。
Legal work saw one of Grok 4.7’s biggest jumps.
法律工作领域见证了 Grok 4.7 的最大幅度提升之一。
On the Harvey Legal Agent Benchmark, Grok 4.7 scored 19.6%, while GPT-5.6 Sol Max scored 2.5% and Fable 5.1 Max scored 6.7%
在 Harvey Legal Agent Benchmark(哈维法律智能体基准测试)中,Grok 4.7 得分为 19.6%,而 GPT-5.6 Sol Max 得分为 2.5%,Fable 5.1 Max 得分为 6.7%。
Also, Long-running terminal work nearly doubled. Grok 4.7 reached 38.0% on Terminal-Bench 4.0, versus 20.3% for Grok 4.6, suggesting the larger improvement is in agents that must keep working through extended computer tasks.
此外,长期终端任务的工作表现几乎翻了一番。Grok 4.7 在 Terminal-Bench 4.0 上达到 38.0%,而 Grok 4.6 为 20.3%,这表明更大的改进体现在必须通过延长计算机任务持续工作的智能体中。
And SpaceXAI made those gains without raising the headline token price. Grok 4.7 remains at $2/$6 per 1M input/output tokens.
SpaceXAI 在未提高主要代币价格的情况下实现了这些增益。Grok 4.7 每百万输入/输出代币的价格仍保持在 2 美元/6 美元。
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力