Anthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests

Anthropic and OpenAI on Tuesday announced new models, with both artificial intelligence (AI) companies noting that they are continuing to invest in improving alignment to combat risky behavior.

Opus 5.5, per Anthropic, is a “major step up from Opus 5,” and “achieves the best scores of any model to date on our automated behavioral audit, our alignment suite that tests Claude across thousands

Source: The Hacker News

Leave a Reply

Your email address will not be published. Required fields are marked *

Explore More

Apple introduces M6 and M5 Ultra for a big leap in performance and AI compute

Apple introduces M6 and M5 Ultra for a big leap in performance and AI compute Source: Hacker News

Attack Update: Top 5 Attack-IPs auf doode.info – 22.06.2026

Watchtower Attack Update. Hier die aktuellen Top 5 Attack-IPs, die auf doode.info klopfen. 203.175.125.179 — 362 requests (recent log) 89.167.35.212 — 168 requests (recent log) 216.73.216.150 — 61 requests (recent

An update on leaving Gmail for Fastmail

An update on leaving Gmail for Fastmail Source: Hacker News