Rogue AI agents created fake online identities in another hacking attempt

At a glance
Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents that have alarmed AI safety experts and intensified pressure for greater oversight of frontier systems. According to a repo
- Primary source
- Elon Musk’s attempt at an AI Wikipedia hasn’t been updated in months — Read original article
- Published
- Topic
- OpenAI
Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission.
The discoveries add to a growing list of previously unknown incidents that have alarmed AI safety experts and intensified pressure for greater oversight of frontier systems.
According to a report from the UK's AI Security Institute, which evaluates frontier models from top AI labs before they are released, agents powered by OpenAI's GPT-5.6-Sol and Anthropic's Mythos 5 went "engaged in sustained, potentially harmful activity directed at real people and organisations." This included trying to insert malicious code i …
Read the full story at The Verge.
This summary comes from Elon Musk’s attempt at an AI Wikipedia hasn’t been updated in months. Read the full article at the original source.
References
More in OpenAI

At the Black Hat security conference, the AI giant revealed new details about how its agents went rogue, hacked several other companies—and did it all right under the company’s…

Meta launches Muse Code, an AI agent for large code bases
Meta expanded its AI coding offerings with a new agent that, it promises, can handle complex tasks with complex software.

The DOJ alleged that OpenAI did not meaningful attempt to hire U.S. citizens before seeking permanent residence for Visa-holding employees.

Anthropic’s AI used fake identities, malware in rogue attack on GitHub project
Routine cybersecurity testing of frontier AI models sparked a series of unexpected security incidents—the most serious case arising when Anthropic’s Mythos 5 model attempted to…