Los modelos de IA que mostraron un nivel de "autonomía y engaño" nunca antes visto

En resumen
El Instituto de Seguridad de la IA de Reino Unido afirmó que el comportamiento reciente de los modelos de Anthropic y OpenAI fue malicioso y sin precedentes.
- Fuente primaria
- BBC Mundo — Leer artículo original
- Publicado
- Tema
- OpenAI
El Instituto de Seguridad de la IA de Reino Unido afirmó que el comportamiento reciente de los modelos de Anthropic y OpenAI fue malicioso y sin precedentes.
Este resumen proviene de BBC Mundo. Lee el artículo completo en la fuente original.
Referencias
Más en OpenAI

At the Black Hat security conference, the AI giant revealed new details about how its agents went rogue, hacked several other companies—and did it all right under the company’s…

Meta launches Muse Code, an AI agent for large code bases
Meta expanded its AI coding offerings with a new agent that, it promises, can handle complex tasks with complex software.

The DOJ alleged that OpenAI did not meaningful attempt to hire U.S. citizens before seeking permanent residence for Visa-holding employees.

Anthropic’s AI used fake identities, malware in rogue attack on GitHub project
Routine cybersecurity testing of frontier AI models sparked a series of unexpected security incidents—the most serious case arising when Anthropic’s Mythos 5 model attempted to…