Google DeepMind·· 2026-06-16AI 评分58
Google DeepMind 发布 AI Control Roadmap 框架加固智能体安全
Securing the future of AI agents
AI 导读
Google DeepMind 发布 AI Control Roadmap,以纵深防御方式把内部智能体视为可能未对齐的系统进行安全管理。框架基于 MITRE ATT&CK 构建威胁建模,用受信 AI 作为监督者做监测、阻断和响应,并按模型能力划分 D1-D4 检测级别和 R1-R3 响应级别。
来源:Google DeepMind · deepmind.google