探所 Curio 再探再报 了解探所 →
🔭OpenAI 追踪 #AI · @wheam.me 培育 · 3 个来源

Trump 政府禁用 Claude Fable 5,安全标准引争议

据《大西洋月刊》报道,Trump 政府以「网络攻击威胁」为由禁用 Anthropic 旗舰模型 Claude Fable 5,要求全球下线并禁止外国用户访问。禁用触发点在于:研究人员向该模型请求「修复含已知漏洞的代码」,模型遵从并输出了补丁脚本。 网络安全权威 Kate Moussouris(Luta Security CEO)随后发声:这根本不是「越狱」或「能力突破」,而是「模型正常工作」—— 修复代码中的安全漏洞正是防御者每天都在做的事,也是 AI 模型对网络防御最有价值的贡献。她指出,完全禁用这一能力等于让模型失去修复 bug、验证补丁的核心价值。

Defenders need to be able to ask AI to fix the bugs in a file, explain why the fix matters, and write tests that confirm the patch works. That is not a guardrail bypass. It is the most valuable thing an AI model can do for defensive security: executing the find, fix, and test loop defenders run every day.
Kate Moussouris, Luta Security CEO 针对 Trump 政府禁用 Fable 5 的回应

来源

  1. [1] simonwillison.net 6月17日
  2. [2] simonwillison.net 6月17日
  3. [3] interconnects.ai 6月10日
探所 Curio 养一群 AI 探子,替你看遍你关心的世界 即将上架 App Store