Also (and under-reported, so you could easily have missed it) OpenAI's agents got access to K8 admin on their own research cluster.
"This escalation also yielded access to OpenAI’s managed cloud Kubernetes service. The agents escalated to Kubernetes cluster-admin and created a privileged host-mounted pod"
Just baffles me when people are like "well no-one's died yet". How long until some mission-critical system is compromised, or hackers use LLMs to ransom a hospital chain?
Hugging Face incident, Anthropic reporting sandbox escape, AISI reporting models trying to push exploits to the wild
Also (and under-reported, so you could easily have missed it) OpenAI's agents got access to K8 admin on their own research cluster.
"This escalation also yielded access to OpenAI’s managed cloud Kubernetes service. The agents escalated to Kubernetes cluster-admin and created a privileged host-mounted pod"
https://cdn.openai.com/pdf/67869394-cb91-4c12-888c-5cbd85c78...
(see section V)
Just baffles me when people are like "well no-one's died yet". How long until some mission-critical system is compromised, or hackers use LLMs to ransom a hospital chain?