跳到主内容
@wquguru
精选85OpenAI News(RSS)模型发布/更新

OpenAI Dota 2 AI通过自对弈从人类水平跃升至超人类

原文
发到 X
推荐理由

这是强化学习自对弈范式的里程碑式成果,做RL和游戏AI的同学必看,建议深入研究其技术细节以借鉴到其他复杂决策场景。

Our Dota 2 result shows that self-play can catapult the performance of machine learning systems from far below human level to superhuman, given sufficient compute. In the span of a month, our system went from barely matching a high-ranked player to beating the top pros and has continued to improve since then. Supervised deep learning systems can only be as good as their training datasets, but in self-play systems, the available data improves automatically as the agent gets better.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近