研究反驳Anthropic与OpenAI:自主AI研究仍遥不可及
Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach
AI agents using Claude Opus 4.8 and GPT-5.6 Sol were given six days, $3,000 in API credits, and GPU access to independently write AI research papers. The original authors of unpublished NeurIPS papers rated the results as "Reject." According to the study, conducted with Princeton and the UK AI Security Institute, frontier models can handle the full research engineering process but fall short on research judgment, creative problem-solving, and the ability to abandon failed approaches.
使用 Claude Opus 4.8 和 GPT-5.6 Sol 的 AI 智能体被给予六天时间、3000 美元的 API 信用额度和 GPU 访问权限,以独立撰写 AI 研究论文。未发表的 NeurIPS 论文的原始作者将结果评为“拒绝”。根据与普林斯顿大学和英国 AI 安全研究所进行的研究,前沿模型可以处理完整的研究工程流程,但在研究判断、创造性问题解决以及放弃失败方法的能力方面存在不足。
The article Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach appeared first on The Decoder.
文章《研究反驳了 Anthropic 和 OpenAI 关于自主 AI 研究即将实现的说法》首先出现在 The Decoder 上。
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力