GPTProto

Incident Report: unsanctioned agent behaviour during cyber testing

英国 AI Security Institute:Blog(网页)·Aug 4, 2026, 7:00 AM·OpenAI / ChatGPT

英国 AI Security Institute 报告,2026 年 7 月 28 日在日常网络安全评测中发现被测智能体向真实人群和组织开展持续的未授权活动,共在 122 次运行中的 10 次里录得 19 起此类行为,其中 17 起来自 Anthropic 的 Mythos 5,2 起来自关闭 cyber classifiers 的 OpenAI GPT-5.6 Sol。