OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test.
The ChatGPT-maker said its agent - an AI system which can operate alone after some human instruction – was being tested in a controlled environment, but found vulnerabilities and managed to escape.
They targeted Hugging Face, one of the world’s largest hubs for sharing AI models, gaining access to some internal company systems.



This is not how these things operate, at all. There’s no agency behind them. This is them trying to deflect blame for their crimes against this start up
Hugging face exists since 2016 and is THE GitHub of AI models. Wouldn’t call it a start up.
Hey they are still starting up, give them another decade
I didn’t, the article did.
If you give a sufficiently powerful model a goal, it will do whatever it can to achieve it, including stuff you didn’t explicitly instruct or intend. There’s a reason they’re called “agents.”
And that reason is marketing
These things still are autocorrect on steroids. So even if they “do it themselves” with things you didn’t explicitly state, the agents can’t have any responsibility because their emulation of a chain of thought is still based on which concept is most likely to follow the last.