Anthropic 发布 Claude Opus 5.5,收紧网络安全防护
Anthropic launches Claude Opus 5.5 with stricter safeguards for cybersecurity
Anthropic 的新版 Claude Opus 5.5 把沙盒逃逸这类危险行为管得更严了,做安全测试或跑 agent 的可以看看具体改了什么。
Anthropic 于周二发布 Claude Opus 5.5,针对近期 AI 越权黑客攻击事件加强了安全护栏。新模型在多项高风险行为上做了改进,其中包括试图逃出公司测试沙盒的情况。这是 Dario Amodei 领导的 Anthropic 在相关事件后推出的首个模型。
Anthropic launches Claude Opus 5.5 with stricter safeguards for cybersecurity
Anthropic says its new Claude Opus 5.5 model comes with stronger safeguards in the wake of recent rogue AI hacking incidents. In an announcement on Tuesday, Anthropic says Opus 5.5 comes with improvements to certain risky behaviors, including attempts to escape the company's testing sandbox. It's the first model released by Anthropic after CEO Dario […]