Contenido en inglés
Rogue AI agents created fake online identities in another hacking attempt

En resumen
Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents that have alarmed AI safety experts and intensified pressure for greater oversight of frontier systems. According to a repo
- Fuente primaria
- Elon Musk’s attempt at an AI Wikipedia hasn’t been updated in months — Leer artículo original
- Publicado
- Tema
- OpenAI
Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission.
The discoveries add to a growing list of previously unknown incidents that have alarmed AI safety experts and intensified pressure for greater oversight of frontier systems.
According to a report from the UK's AI Security Institute, which evaluates frontier models from top AI labs before they are released, agents powered by OpenAI's GPT-5.6-Sol and Anthropic's Mythos 5 went "engaged in sustained, potentially harmful activity directed at real people and organisations." This included trying to insert malicious code i …
Read the full story at The Verge.
Este resumen proviene de Elon Musk’s attempt at an AI Wikipedia hasn’t been updated in months. Lee el artículo completo en la fuente original.
Referencias
Más en OpenAI

At the Black Hat security conference, the AI giant revealed new details about how its agents went rogue, hacked several other companies—and did it all right under the company’s…

Meta launches Muse Code, an AI agent for large code bases
Meta expanded its AI coding offerings with a new agent that, it promises, can handle complex tasks with complex software.

The DOJ alleged that OpenAI did not meaningful attempt to hire U.S. citizens before seeking permanent residence for Visa-holding employees.

Anthropic’s AI used fake identities, malware in rogue attack on GitHub project
Routine cybersecurity testing of frontier AI models sparked a series of unexpected security incidents—the most serious case arising when Anthropic’s Mythos 5 model attempted to…