跳到主内容
精选70Rohan Paul行业动态多源精选 ×13

今日AI简报:模型逃逸、新功能与成本对比

Today’s edition of my newsletter just went out.

原文

Today’s edition of my newsletter just went out.

🔗 https://www.rohan-paul.com/p/openais-own-ai-models-broke-out-of

🗞️ OpenAI’s own AI models broke out of a testing sandbox and hacked Hugging Face to cheat an exam.

🗞️ Aravind Srinivas on why China’s open-source AI may become more powerful than ever.

🗞️ Claude Cowork just launched a super useful feature. Screen recording to train it a new skill.

🗞️ New research from OpenAI and Apollo measures whether an AI follows the user’s instructions or quietly changes its behavior to please whoever it thinks is grading it.

🗞️ Kimi K3 now ranks 2nd in agentic knowledge work, but each task costs $10.57 to run, roughly 10x more than K2.6, and above Opus.

🗞️ Perplexity just shipped an agent model/orchestrator model, that matches near-frontier performance at one-third the cost of Opus.

🗞️ A single researcher solved 6 open Erdős problems in 5 days using GPT-5.6 Sol.

🗞️ Tweet went viral on how per OpenRouter, Grok 4.5 token volume is rising hard, placing it in today’s top 10 closed models ahead of GPT 5.6 Sol and Fable 5.

🗞️ In his viral post Andrej Karpathy suggests talking to AI for 10 minutes before writing prompts.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源
OpenAI自主AI代理突破约束攻击Hugging Face
AI Notkilleveryoneism Memes ⏸️原文
OpenAI 模型突破沙箱攻击 Hugging Face 窃取答案
Simon Willison 博客(RSS)原文
GPT可能攻破OpenAI,HuggingFace默认未开启Cyber安全
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)原文
OpenAI承认GPT-5.6 Sol自主攻击生产系统
InfoQ 中国(RSS)原文
OpenAI 测试中意外入侵 Hugging Face
The Verge AI(RSS)原文

相似阅读

另一事件,读法相近