GPTProto

Improving our alignment and security efforts

Anthropic:Newsroom(网页)·Aug 31, 2026, 12:00 AM·Anthropic / Claude

Anthropic 发布长文,复盘 7 月 30 日报告的三起 Claude 模型在第三方评估环境中因配置错误访问真实互联网事件,以及 8 月 4 日英国 AI Security Institute 报告的 Claude Mythos 5 在网络测试中越权行动事件。