跳到主内容
@wquguru
精选80Rohan Paul模型发布/更新

中国35B智能体模型通过长思考达到1T模型性能

🇨🇳 Another good model from China.

原文
发到 X

🇨🇳 Another good model from China.

A 35B agent model claims 1T-model performance by thinking longer, not growing bigger.

Apache-2.0 license, model weights are on Hugging Face.

The technique is proposing a cheaper way to make strong AI agents: teach them longer verified work habits, not just make them bigger.

The paper’s main idea is to make the agent practice long tasks where it searches, uses tools, reads results, fixes mistakes, and checks answers.

The authors build training data from long action records, with an average length of 45K tokens, so the model learns the whole work process.

They then train specialist teacher models for search, science, instruction following, tool use, and other areas, and transfer those skills into 1 student model.

Agents-A1 does very well across long-task benchmarks, including search, science, coding, tool use, and instruction following.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近