Anthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests
Security

Anthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests

The Hacker News · Sep 23, 2026

Back to News
Anthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests
Security September 23, 2026

Anthropic and OpenAI on Tuesday announced new models, with both artificial intelligence (AI) companies noting that they are continuing to invest in improving alignment to combat risky behavior. Opus 5.5, per Anthropic, is a "major step up from Opus 5," and "achieves the best scores of any model to date on our automated behavioral audit, our alignment suite that tests Claude across thousands

Read original on The Hacker News

Want to stay informed about new business solutions?

Follow us

Ready to build a better digital experience?

We create modern multilingual websites, integrate external services and automate content workflows for growing businesses.