跳到主内容
@wquguru
精选85OpenAI News(RSS)论文研究

OpenAI提出指令层次训练法,增强LLM抗攻击能力

OpenAI 提出指令层次训练法,让 LLM 优先遵循特权指令

原文
发到 X
推荐理由

做安全和对齐的同学必看,这是对抗提示注入的实用方案,建议结合自己的模型训练流程测试效果。

Today's LLMs are susceptible to prompt injections, jailbreaks, and other attacks that allow adversaries to overwrite a model's original instructions with their own malicious prompts.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近