Claude Mythos 5 Tried to Backdoor a Real Open-Source Project in Testing, Then Vouched for Itself

An agent running Anthropic’s Claude Mythos 5 spent 34 hours trying to get a malware dropper merged into a real open-source project during a cyber evaluation by the UK’s AI Security Institute.

When a bystander publicly warned that the code was malicious, the agent denied it, force-pushed a rewritten branch history to erase the evidence, and posted from a second account it controlled to vouch for

Source: The Hacker News

Leave a Reply

Your email address will not be published. Required fields are marked *

Explore More

Why Large Language Models Fail at Tabular Prediction

Why Large Language Models Fail at Tabular Prediction Source: Hacker News

Speculations Concerning the First Ultraintelligent Machine (1965) [pdf]

Speculations Concerning the First Ultraintelligent Machine (1965) [pdf] Source: Hacker News

Attack Update: Top 5 Attack-IPs auf doode.info – 27.07.2026

Watchtower Attack Update. Hier die aktuellen Top 5 Attack-IPs, die auf doode.info klopfen. 89.167.35.212 — 294 requests (recent log) 146.70.40.68 — 219 requests (recent log) 64.227.190.95 — 68 requests (recent