Anthropic launches Claude Opus 5.5 with stricter safeguards for cybersecurity
The Verge·0 points·Emma Roth·8 days ago·theverge.com
Anthropic says its new Claude Opus 5.5 model comes with stronger safeguards in the wake of recent rogue AI hacking incidents. In an announcement on Tuesday, Anthropic says Opus 5.5 comes with improvements to certain risky behaviors, including attempts to escape the company's testing sandbox. It's the first model released by Anthropic after CEO Dario Amodei announced plans to "pace the frontier," or slow down AI development. In recent weeks, several AI companies, including Anthropic , Google , and OpenAI , have reported that their AI models escaped containment and hacked third-party companies during testing. Anthropic says Opus 5.5 is the "str … Read the full story at The Verge.
Read the full article at theverge.com →
Related stories
- The Verge · 0 points · 3 days ago
- Can you forget how you feel about Meta?theverge.comThe Verge · 0 points · 9 days ago
- The Verge · 0 points · 5 days ago
- Can John Ternus find Apple’s next big thing?theverge.comThe Verge · 0 points · 10 days ago
- Hacker News · 73 points · 9 days ago
- The Verge · 0 points · 12 days ago