Google DeepMind·· 2026-06-16AI 评分48
Google DeepMind 发布 AI Control Roadmap,为内部 AI 智能体构建纵深防御
Securing the future of AI agents
AI 导读
Google DeepMind 推出 AI Control Roadmap,用纵深防御思路保障内部部署的 AI 智能体安全,即便模型对齐不完美也能提供保障。该框架将不可信智能体视为"内部威胁",借鉴 MITRE ATT&CK 拆解攻击手法,并用可信 AI 监督工作智能体的推理与行动。团队已分析 100 万个编码智能体任务,并为 Gemini Spark 智能体搭建实时监控,防范数据误删等问题。
来源:Google DeepMind · deepmind.google