精选75OpenAI News(RSS)模型发布/更新
OpenAI 发布进化策略梯度 EPG,实现快速任务适应
OpenAI 发布进化策略梯度方法 EPG
We’re releasing an experimental metalearning approach called Evolved Policy Gradients, a method that evolves the loss function of learning agents, which can enable fast training on novel tasks. Agents trained with EPG can succeed at basic tasks at test time that were outside their training regime, like learning to navigate to an object on a different side of the room from where it was placed during training.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力