跳到正文
原文
Google DeepMind·· 2026-06-16AI 评分58

Google DeepMind 发布 AI Control Roadmap 框架加固智能体安全

Securing the future of AI agents

AI 导读

Google DeepMind 发布 AI Control Roadmap,以纵深防御方式把内部智能体视为可能未对齐的系统进行安全管理。框架基于 MITRE ATT&CK 构建威胁建模,用受信 AI 作为监督者做监测、阻断和响应,并按模型能力划分 D1-D4 检测级别和 R1-R3 响应级别。

来源:Google DeepMind · deepmind.google