GPTProto

Measuring tactical intelligence targeting and conventional weapons capabilities of AI models

Anthropic:Research(发表成果 · 网页)·Sep 10, 2026, 12:00 AM·Anthropic / Claude

Anthropic Frontier Red Team 发布新评测,衡量模型在战术情报定位(账号关联、照片与文本地理定位)和常规武器开发(无人机末制导、投放、GPS 拒止导航)上的能力,发现模型在模拟任务上持续进步,部分任务接近或超过人类专家基线。