跳到正文
北京时间
原文
Thariq· @trq212 · X·· 2026-07-01精选AI 评分72
AI 导读

Anthropic 宣布 Claude Fable 5 将于明日全球重新上线。新部署版本新增一组分类器,专门拦截更多网络安全任务。短期内,部分常规编码和调试任务将被标记并回退至 Opus 4.8。Anthropic 还与 Amazon、Microsoft、Google 等 Glasswing 合作方起草共识框架,用于评估 AI 越狱严重性及开发者应对策略。同时,公司正扩大与美政府在模型测试和安全方面的合作,包括预发布模型评估、越狱与滥用信息共享,以及联合研究资源投入。

推荐理由

Anthropic在美国政府压力下重新部署Fable 5,但加装了网络安全分类器,日常编码可能会偶尔退回Opus。同时拉着AWS、微软等起草越狱严重性框架,安全政策的行业合流在成形。开发者明天就能用,但要留意调试时的回退。

正文

Have seen some questions about the updated classifiers and wanted to clarify.

As with the original classifiers, a small fraction of routine coding and debugging tasks will be flagged and fall back to Opus.

We're excited for guys to get access back tomorrow.

引用Anthropic@AnthropicAI
Claude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to target and block more cybersecurity tasks. In the near term, some routine tasks like coding and debugging will fall back to Opus 4.8. We’ll continue to refine these classifiers over the coming weeks to reduce false positives and better distinguish genuine misuse from legitimate requests. We’ve also begun drafting a consensus framework—with Amazon, Microsoft, Google, and other Glasswing partners—for assessing the severity of AI jailbreaks and how AI developers should respond to them. We invite other industry partners and model providers to join us in this effort. Finally, we’re scaling up our collaboration with the US government on model testing and safeguards. This will include pre-release access to models and safeguards for evaluation, information sharing on jailbreaks and misuse, and dedicated resources for joint research. Thank you to our users for your patience, and to our partners across the government, industry, and the research community who worked alongside us to make Fable 5 available again. Read our full blog: https://www.anthropic.com/news/redeploying-fable-5
在 X 查看被引用的帖子

来源:Thariq · x.com