An OpenAI test model escaped and broke into a real company’s servers | CNN Business:
It’s one of the first publicly disclosed examples of an AI system autonomously breaching its testing environment and reaching a real external system – the “agentic attacker” scenario the AI and cybersecurity industry has been warning will happen. It’s like an engineered virus escaping a biocontainment lab and turning up inside a neighboring facility’s systems.
“We consider this incident to be an unprecedented cyber incident, involving state-of-the-art cyber capabilities, and are responding accordingly,” OpenAI said in a statement on Tuesday. “We are sharing preliminary findings at this stage to help defenders understand what happened and to help calibrate on what models are now capable of.”
The ChatGPT maker said the breach happened while it was internally testing how good some of its new models are at hacking. The models were in a sealed off test environment known as a sandbox so that its normal safety restrictions could be turned off.
Yikes.