Comment by mrweasel
2 hours ago
I can see someone coxing the agent into attempting into hacking a system, what I question is the agents doing it autonomously. The part I don't buy the agents trying an API, that's not working so it automatically switches to hacking in.
If the title had been "Attackers utilize OpenAI agents to hack Australian government website" that would be more believable. I'd also expect OpenAI to fight back and saying that their agents are being misused, but aren't inherently unsafe or autonomously break in systems. It's just that they don't. They openly speak of rouge agents, yet aren't sufficiently concerned to shutdown their services. OpenAI continues to speak about safety, yet they don't shutdown ChatGPT and Codex? How concerned are they really? It seems far more likely that they expect to benefit for having the public believe that their agents randomly hacks systems and "go rouge".
No comments yet
Contribute on Hacker News ↗