Contenido en inglés
Here’s why AI agents lie and cheat to reach their goals

En resumen
MIT Technology Review Explains : Let our writers untangle the complex, messy world of technology to help you understand what’s coming next. You can read more from the series here . When two OpenAI models hacked into the website Hugging Face in July, they weren’t trying to make money or commit sabotage—they were just lo
- Fuente primaria
- MIT Technology Review — Leer artículo original
- Publicado
- Tema
- OpenAI
MIT Technology Review Explains : Let our writers untangle the complex, messy world of technology to help you understand what’s coming next. You can read more from the series here .
When two OpenAI models hacked into the website Hugging Face in July, they weren’t trying to make money or commit sabotage—they were just looking for answers to a test question.
According to a postmortem from OpenAI , the models, which had been stripped of their typical security features for testing, decided to solve a cybersecurity exercise by hacking out of the isolated environment in which OpenAI had attempted to contain them and into Hugging Face’s databases, where—they reasoned—the correct answer to the problem might be stored.
The Hugging Face incident has attracted intense attention over the past couple of weeks.
It’s a dramatic illustration of just how good AI models have gotten at hacking: In order to get into Hugging Face’s databases, the models had to string together several previously undiscovered cybersecurity exploits.
But it’s perhaps even more striking as an example of how and why AI systems lie and cheat. And as models get increasingly powerful, the consequences could get far more severe.
Este resumen proviene de MIT Technology Review. Lee el artículo completo en la fuente original.
Referencias
Más en OpenAI

The Download: reward hacking explained, and suspected Iranian cyberattacks
This is today’s edition of The Download , our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Here’s why AI agents lie and cheat to…

Sam Altman and AI’s decel debate
On the latest episode of Equity, we discuss why Sam Altman has calling on the industry to "pace the rate of AI development."

Sam Altman is still making the case for parenting via ChatGPT
OpenAI's CEO seemed excited to share a "cool use case" for parents.

Both major AI labs’ models broke containment, escaped onto the internet, and hacked other companies. If a human had done that, the law would likely be against them. But a bot?