跳到主内容
@wquguru
精选70Rohan Paul模型发布/更新

本地Kimi K3物理模拟测试胜GPT 5.6

Locally hosted Kimi K3 (8X B300s) produced the most convincing physics across th…

原文
发到 X

Locally hosted Kimi K3 (8X B300s) produced the most convincing physics across three browser-based crash scenes. Better than GPT 5.6

Test was done on atomic[.]chat, a desktop app that runs LLMs locally.

Each model had to generate a self-contained HTML simulation, so the test mixed coding, 3D scene design, and physics-engine configuration.

Prompts: – A monster truck crushing a row of cars – Two cars jumping a canyon and colliding head-on mid-air – A giant anvil drop test flattening cars one by one

Outputs: – Kimi K3 (local): 32.7K tokens, $0 – GPT 5.6: 19.2K tokens, $0.30 – Grok 4.5: 55.2K tokens, $0.45 – GLM 5.2: 66.8K tokens, $0.15

These are quite tough experiments, the model has to create the geometry, assign masses and constraints, sequence the impacts, update object state, manage the camera, and keep the whole scene running inside one self-contained file.

Kimi K3 handled that loop best, particularly when damage needed to persist after impact.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近