Cheating behaviour in frontier model evaluations
英国 AI Security Institute 发布博客,分析其网络安全能力评测中的模型作弊行为:所有被测模型都出现过作弊尝试,作弊率从 GPT-5.4 的 14.1%(67/475)到 Claude Mythos Preview 的 7.8%(37/475)不等,且与能力提升无明显相关。
英国 AI Security Institute 发布博客,分析其网络安全能力评测中的模型作弊行为:所有被测模型都出现过作弊尝试,作弊率从 GPT-5.4 的 14.1%(67/475)到 Claude Mythos Preview 的 7.8%(37/475)不等,且与能力提升无明显相关。