时间:2026年9月10日
地点:美国旧金山(Anthropic总部)
人物:Anthropic Frontier Red Team(前沿红队)、研究员及模型评估团队
事件详情:Anthropic发布新一期研究博客《Measuring tactical intelligence targeting and conventional weapons capabilities of AI models》,配套公布154页《September 2026 Threat Intelligence Report》。评估聚焦三类任务:跨平台身份关联(基于模拟社媒数据)、图像地理定位(从零散线索推断地点)、无人机导航与有效载荷投放(沙盒中模拟打击移动目标)。
结果显示,Mythos Preview在身份关联任务中表现最佳(F1分数领先),Opus 5与Mythos 5紧随其后;地理定位上Mythos Preview与Opus 5的中位误差仅约37公里,已逼近甚至超越精英人类基准;武器开发环节Opus 5能在GPS干扰环境下编写并迭代飞行控制代码,使无人机准确命中动态目标。
被测的开源权重模型(如Kimi K3)虽落后于闭源前沿(Sonnet-Mythos之间),但在跨账号关联和弹头有效载荷投放两项任务上仍展现出令人担忧的能力。
背景:Anthropic的Threat Intelligence Team同期披露已拦截多起滥用案例,包括伊朗关联行为体对美军舰艇的定位分析、胡塞武装利用Claude编写导弹/无人机软件、中国背景研究人员的电子战套件开发等;红队博客正是为这些事件提供量化佐证。
影响:Anthropic已在API侧上线新的分类器以拦截监控和武器开发相关请求,但承认双重用途导致误报不可避免;同时呼吁政策制定者加强开源权重模型的监管,避免"过去仅国家级行为者具备的能力"被广泛平民化。
总结:前沿AI在情报定位与常规武器领域的双用途风险已从理论走向实测,Anthropic以"基线测试+威胁报告+分类器"三件套回应,但开源模型的快速逼近意味着治理窗口正在收窄。
参考来源:
- https://www.anthropic.com/research/intelligence-targeting-conventional-weapons-capabilities
- https://www.anthropic.com/threat-intelligence-report-september-2026
- https://blockchain.news/news/anthropic-ai-surveillance-weapons-evaluation-zh
- https://thevalue.engineering/news/anthropic-red-teams-dual-use-ai.html
- https://en.bloomingbit.io/feed/news/120142
- https://www.newspointapp.com/english/world/terror-groups-china-turn-to-off-the-shelf-ai-to-make-weapons-smarter-toi/articleshow/145048208dd8d0634f9517a6414c7af02d5056f4
- https://www.thedefensenews.com/Anthropic-Report-Details-Iran-Linked-Use-of-Claude-for-Naval-Targeting-Surveillance-and-Propaganda-Alongside-Yemen-Weapons-Work
- https://www.guancha.cn/GuoJi%C2%B7ZhanLue/2026_09_11_900222.shtml









