跳到主内容
@wquguru
精选88Ars Technica AI(RSS)模型发布/更新多源精选 ×8

Google发布Gemini 4 Argon模型,暂不开放

Google announces Gemini 4 Argon AI model, but you can't use it yet

原文
发到 X
推荐理由

全新旗舰模型发布,基准测试全面超越竞品,且披露了内部大规模工程落地细节,值得开发者关注其架构演进。

Google promised Gemini 3.5 Pro back in June, but it spent the summer trotting out smaller Flash models. Now, Google is ready to take on the frontier again with Gemini 4 Argon. The company claims this new AI offers industry-leading performance in coding, knowledge work, and cybersecurity, but you aren't allowed to use it yet.

谷歌早在六月就承诺推出 Gemini 3.5 Pro,但整个夏天都在陆续发布较小的 Flash 模型。如今,谷歌准备凭借 Gemini 4 Argon 再次挑战前沿领域。该公司声称这款新 AI 在编码、知识工作和网络安全方面提供行业领先的性能,但你目前还无法使用它。

While most of us will have to wait to put Gemini 4 to the test, Google says engineers inside the company are already making extensive use of the new model. Argon reportedly used "fleet-wide telemetry data" to help Google save 300 TiB of memory across its data centers. Meanwhile, Argon agents have been working to migrate C/C++ codebases to Rust across Google, including thousands of lines in the core re2 and libgav1 libraries and more than 800,000 lines in the Fuchsia OS Zircon kernel.

虽然大多数人还得等待才能测试 Gemini 4,但谷歌表示公司内部工程师已经广泛使用该新模型。据报道,Argon 利用“全车队遥测数据”帮助谷歌在其数据中心节省了 300 TiB 的内存。与此同时,Argon 代理一直在推动谷歌内部将 C/C++ 代码库迁移到 Rust,包括核心 re2 和 libgav1 库中的数千行代码,以及 Fuchsia OS Zircon 内核中超过 800,000 行代码。

Google has also come armed with a raft of benchmarks to back up its claims. On the software engineering DeepSWE v1.1 benchmark, Gemini 4 Argon hits 77.9 percent, which is higher than GPT-6 Astra, Fable 5.1, and Opus 5.5. Google promises similar power across a range of long-horizon tasks, pointing to Argon's industry-leading score in the economic analysis Vals Index test.

谷歌还带来了一系列基准测试来支持其主张。在软件工程 DeepSWE v1.1 基准测试中,Gemini 4 Argon 得分达到 77.9%,高于 GPT-6 Astra、Fable 5.1 和 Opus 5.5。谷歌承诺在一系列长周期任务中具备同等实力,并指出 Argon 在经济分析 Vals Index 测试中取得了行业领先的分数。

Read full article

阅读全文

Comments

评论

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →