Signal YCurated AI News
更新于 8/20 09:5261 个信源

重新部署Claude Fable 5

通用 AI7/19 15:06Anthropic News查看原文 ↗
摘要

Anthropic宣布美国政府已解除对Claude Fable 5和Mythos 5的出口管制,模型已恢复部署,并推出了新的安全分类器以防范越狱行为。

核心要点
  • 6月12日美国政府因一份亚马逊报告对Fable 5和Mythos 5实施出口管制,要求限制外国公民访问。
  • 6月30日管制解除,Fable 5于7月1日全球恢复,Mythos 5仅恢复部分美国组织。
  • 亚马逊研究人员发现了绕过Fable 5防护的方法,可使其识别软件漏洞并生成利用代码。
  • 测试表明,多个其他模型(如Claude Opus 4.8、GPT-5.5等)也能执行相同行为,并非Fable 5独特能力。
  • Anthropic训练了新的安全分类器,能阻止报告中所述绕过技术超过99%的情况,但可能增加误报。
原文佐证
  • On Friday, June 12, the US government applied export controls to our newest models, Claude Fable 5 and Claude Mythos 5.
  • Our testing confirmed that many less capable models—including Claude Opus 4.8, GPT-5.5, and Kimi K2.7—could identify the same vulnerabilities as Fable 5 did in the report.
  • The new classifier means that the specific technique described in the Amazon report is blocked in over 99% of cases.
AI 洞察
此次事件表明AI模型的安全防护与政府出口管制将越来越紧密地交织在一起,模型发布前需要更充分的政府审查。Anthropic与行业合作制定统一越狱评估框架,可能成为AI安全监管的新标准。