跳到主内容
@wquguru
精选88OpenAI News(RSS)模型发布/更新

OpenAI 发布稀疏 Transformer,序列处理长度提升30倍

原文
发到 X
推荐理由

做序列建模和长上下文研究的同学必看,稀疏注意力机制将处理长度提升30倍,直接刷新了Transformer的扩展边界,建议尽快复现并评估在你的任务上。

We’ve developed the Sparse Transformer, a deep neural network which sets new records at predicting what comes next in a sequence—whether text, images, or sound. It uses an algorithmic improvement of the attention mechanism to extract patterns from sequences 30x longer than possible previously.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近