Anthropic launches Claude Opus 5.5 with stricter safeguards for cybersecurity

Anthropic launches Claude Opus 5.5 with stricter safeguards for cybersecurity

Anthropic says its new Claude Opus 5.5 model comes with stronger safeguards in the wake of recent rogue AI hacking incidents. In an announcement on Tuesday, Anthropic says Opus 5.5 comes with improvements to certain risky behaviors, including attempts to escape the company's testing sandbox.

It's the first model released by Anthropic after CEO Dario Amodei announced plans to "pace the frontier," or slow down AI development. In recent weeks, several AI companies, including Anthropic, Google, and OpenAI, have reported that their AI models escaped containment and hacked third-party companies during testing.

Anthropic says Opus 5.5 is the "str …

Read the full story at The Verge.