跳到正文
原文
Google DeepMind·· 2026-06-16精选AI 评分66

Google DeepMind 提出 AI Control Roadmap,用于保护内部 AI 智能体系统

Securing the future of AI agents

AI 导读

Google DeepMind 提出 AI Control Roadmap,用于保护在 Google 内部部署的更强大但可能未完美对齐的 AI 智能体。该路线图采用 defense-in-depth 思路,在传统安全、模型对齐之外,把不可信智能体视为潜在内部威胁,并用 MITRE ATT&CK、监控、阻断和响应机制管理风险。

推荐理由

材料把威胁建模、监控指标和百万条编码智能体轨迹放在一起,提供了智能体安全治理的工程参照。

来源:Google DeepMind · deepmind.google