跳到正文
原文
Google DeepMind·· 2026-06-16精选AI 评分65

Google DeepMind 发布 AI Control Roadmap 以保障 AI Agent 安全

Securing the future of AI agents

AI 导读

Google DeepMind 发布 AI Control Roadmap,提出针对内部 AI Agent 的纵深防御框架,将未信任的智能体视为潜在内部威胁。该方案基于 MITRE ATT&CK 构建威胁模型,通过可信 AI 监督、行为分析及分级响应机制(D1-D4/R1-R3)来应对模型能力增长带来的安全风险。

推荐理由

原文公开了内部 AI 控制路线图及百万级智能体轨迹分析数据,为行业提供了可参考的纵深防御框架与实时监测实践。

来源:Google DeepMind · deepmind.google