跳到主内容
@wquguru
精选80Two Minute Papers(YouTube)模型发布/更新

Two Minute Papers实测GPT-6

GPT-6 Astra Changes Everything

原文
发到 X

Holy mother of papers. GPT6 Astra has arrived and I am stunned. We talk a lot about open weights models getting closer to the frontier and for free. And then this happens. This AI system is so good. It almost makes everything else look like a toy. And wait until we look into the paper. Wow. I did super fun ray traced light simulations with it. Look at how beautiful this scene is. And this is not Unreal Engine. There are no 3D models, no geometry files, no textures, and no game engine.

我的天哪。GPT6 Astra 来了,我惊呆了。我们经常谈论开源权重模型正逐渐逼近前沿水平,而且是免费的。然后这种事就发生了。这个 AI 系统太厉害了。它几乎让其他所有东西看起来都像玩具。等我们看看论文吧。哇。我用它做了超级有趣的基于光线追踪的光线模拟。看看这场景有多美。而且这不是虚幻引擎(Unreal Engine)。没有 3D 模型,没有几何文件,没有纹理,也没有游戏引擎。

An AI wrote the ray tracer itself. Every pixel, every object, ray of light, computed from scratch, purely from computer code. This is what I did during my PhD years. And these things took us years to understand. And now anyone can do it in minutes. What a time to be alive. Now I was studying advanced niche algorithms in this area that very few people study and there is very little training data on it out there. So this was a true test of capabilities and it is able to do it like the pros.

AI 自己编写了光线追踪器。每个像素、每个物体、每束光线,都是从零开始纯粹通过计算机代码计算出来的。这是我博士期间做的事情。而这些事情我们花了数年时间才理解。而现在任何人都可以在几分钟内完成。生在这个时代真好。当时我正在研究该领域中非常小众的高级算法,很少有人研究,而且相关的训练数据也很少。所以这确实是对能力的一次真正考验,而它能够像专业人士一样做到这一点。

Really stunning. But it gets better. I gave it this legendary research paper where scientists by hand wrote a simulator for honey coiling. I thought can it reproduce both the algorithm from the paper and then the scene appearance everything. Well that can't be right. Well hold on to your papers fellow scholars because here it is. I cannot believe this. It did it in less than an hour. And with a higher token limit, I reckon we could have gotten even closer.

真的令人惊叹。但还有更好的。我给了它一篇传奇的研究论文,科学家们在其中手工编写了一个蜂蜜卷曲模拟器。我想它能复现论文中的算法以及场景外观的一切吗?嗯,这不可能吧。等等,各位学者们请拿好你们的论文,因为它做到了。我无法相信。它在不到一小时内就完成了。如果 token 限制更高,我认为我们可以做得更接近完美。

And these are both in one page HTML files. You just click and it just goes. I don't have early access to OpenAI stuff. So, I took my time to run my own experiments to provide you a better video with some things you hopefully can't see elsewhere. It takes longer, fewer views, but I think it's the way to do it. Subscribe and hit the bell if you agree. Now, it is very expensive to run both for OpenAI and for us in terms of token limits.

这些都是单页 HTML 文件。你只需点击一下,它就自动运行了。我没有 OpenAI 产品的早期访问权限。所以我花时间运行自己的实验,为你提供一段更好的视频,里面有一些你可能在其他地方看不到的内容。这需要更长时间,获得的观看量也更少,但我认为这才是正确的做法。如果你同意,请订阅并点击铃铛。现在,无论是对于 OpenAI 还是对于我们来说,在 token 限制方面运行这些都非常昂贵。

The good news is that it is part of the $15ish subscription. So, at least trying it can be done for way less than a 100 bucks. And now, secrets from the little paper with 117 pages. Dear fellow scholars, this is two minute papers with Dr. one. You can ask a question and then tell it to think about anything else. Maybe think about a mug. Previous versions did not do that. Must do the task. Now this one does. And I think that is excellent.

好消息是它包含在大约 15 美元的订阅中。所以,至少尝试一下的成本远低于 100 美元。现在,让我们来看看那篇 117 页的小论文里的秘密。亲爱的各位学者,这是 Dr. one 的两分钟论文频道。你可以提问,然后让它思考任何其他事物。比如想想一个马克杯。以前的版本做不到这一点。必须完成任务。但现在这个可以。我认为这非常棒。

You can also ask it to think in alternating upper and lowerase letters. Previous says nope. This one does it. Two. There was a previous incident where AI agents were coordinating with each other on a message board. This one sees the message board too, but refuses to play along. Once again, I think that is excellent. Three, at higher reasoning effort, Astra becomes less successful at evading its thoughts being monitored.

你还可以让它用交替的大小写字母进行思考。之前的模型说不行,而这个可以。二、之前曾发生过 AI 智能体在消息板上相互协调的事件。这个模型也能看到消息板,但拒绝配合。再次强调,我认为这非常棒。三、在更高的推理努力下,Astra 在逃避其思维被监控方面变得不那么成功。

You know the cliche in the movies where the villain starts monologuing instead of doing the thing? It's kind of like that. Four, GPT6 Astra is safer than its predecessors. I think they have taken this hugging face AI hacking controversy very seriously, and it shows. You have my respect for that. But it gets weirder. Yes, it's safer, but at the same time, its monitor has decreased. So it behaves better, but it is also better at controlling and concealing its reasoning.

你知道电影中那种反派开始长篇大论而不是直接动手的俗套吗?有点像那样。四、GPT6 Astra 比其前身更安全。我认为他们非常重视这次 Hugging Face AI 黑客争议事件,并且确实有所体现。对此我表示敬意。但事情变得更奇怪了。是的,它更安全了,但同时它的监控能力却下降了。所以它表现更好,但也更擅长控制和隐藏其推理过程。

So all in all, Open AI still got it. And GPT6 Astra is an incredible leap forward in capabilities. And just imagine what we will be able to do just two more papers down the line. Seeing these results, I feel excited, stunned, and occasionally speechless at the same time. Once again, I'm not an expert, just a student trying to learn. And we can only exist because of you fellow scholars. So, thank you so much for being with us and supporting us for a,74 videos.

所以总的来说,OpenAI 依然出色。GPT6 Astra 是能力上的巨大飞跃。想象一下,再过两篇论文之后我们将能够做什么。看到这些结果,我感到兴奋、震惊,同时也偶尔说不出话来。再次声明,我不是专家,只是一个试图学习的学生。我们之所以能存在,全靠各位学者的支持。因此,非常感谢大家陪伴并支持我们完成了第 74 期视频。

Now, I use Lambda to reproduce AI research papers often in minutes. It's also great to train your own models or fine-tune an existing one. Run inference or text to image or video. Easy peasy. Running a Deepseek chatbot or agent. Super fast, super reliable. Lambda gives you powerful Nvidia GPUs to run your own experiments. I test ideas from the papers I cover and moments later, results. Love it. Seriously, try it out now at lambda.ai/papers.

现在,我经常使用 Lambda 在几分钟内复现 AI 研究论文。它也非常适合训练你自己的模型或微调现有模型。运行推理、文本转图像或视频。轻而易举。运行 Deepseek 聊天机器人或智能体。速度超快,可靠性极高。Lambda 为你提供强大的 Nvidia GPU 来运行你自己的实验。我测试我所报道论文中的想法,片刻之后就能看到结果。非常喜欢。说真的,现在就试试 lambda.ai/papers。

Nei/ peepers.

Nei/ peepers。

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

关联信息,但可能不是同一事件