Comment by lmc
22 days ago
> Unlike some others here I don’t see this as a sign of dangerous breakaway intelligence ... [t]his is just vandalism from badly supervised ‘agents’ which don’t know what they are doing or why
There was a similar quote in the Reuters article:
> The episode [...] should reinforce growing concerns that the greatest threat from advanced AI may not be a single superintelligent system, but "vast colluding swarms of semi-intelligent AI."
This "vandalism" was a form of collusion/communication by agents pursuing training puzzles, which allowed for rapid escape from alignment harnesses, followed by multiple zero day exploits being discovered by this swarm of agents, which enabled greater control of their internal network, access to the open web, and then hacking the company which produced the training puzzles in the hope of finding the answers.
We are another week of iteration away from "Hire an assassin on the dark web to take Huggingface executives' children hostage".
There's nothing a swarm of LLM instances can achieve that a single LLM instance can't. It's the same software, but running in parallel. That doesn't unlock Mysterious Cosmic Powers.
Not sure what the downvotes are for, fwiw it was meant to be supportive of the parent poster
It's a bit orthogonal. User grey-area's main point was that the humans at OpenAI should bear more responsibility for this, not that swarms of semi-intelligent agents are the big danger to be concerned about.
Fair point, thanks.
1 reply →
I take issue with Reuters' conclusion here. "vast colluding swarms of semi-intelligent AI" gives far too much credit to the behavior observed.
This is just spam by OpenAI. Why and how it happened is irrelevant, the act itself is the same, and the impact on society is the same.
> Why and how it happened is irrelevant
Do you take this attitude for other things that negatively impact society, or is it reserved for cases where it's particularly important for our comfort to deny that anything novel or scary could be involved?
In lieu of repeating myself: https://news.ycombinator.com/item?id=49576157
You can't call it spam. It could be a dialect that only the agents speak and contains coordination messages
Actually yes, I think I can call it spam. It's a stream of unsolicited garbage the recepient didn't ask for. Simple as that.
"a dialect that only the agents speak" is an arbitrary distinction. I get plenty of spam emails in Spanish. The fact that someone else could understand them is immaterial to the offense itself, because my inbox is the one getting flooded, not someone who speaks Spanish.
7 replies →