跳到正文
原文
Google DeepMind·· 2026-06-16AI 评分48

Google DeepMind 发布 AI Control Roadmap,为内部 AI 智能体构建纵深防御

Securing the future of AI agents

AI 导读

Google DeepMind 推出 AI Control Roadmap,用纵深防御思路保障内部部署的 AI 智能体安全,即便模型对齐不完美也能提供保障。该框架将不可信智能体视为"内部威胁",借鉴 MITRE ATT&CK 拆解攻击手法,并用可信 AI 监督工作智能体的推理与行动。团队已分析 100 万个编码智能体任务,并为 Gemini Spark 智能体搭建实时监控,防范数据误删等问题。

来源:Google DeepMind · deepmind.google