Comment by bulder

20 hours ago

Since they're statistical likelihood machines, I'd guess that the order of operations is

* Need persistent scratch space

* Look for public writeable websites

* Needs to be low-traffic so the notes don't drown in noise

* Pick a "random" wiki name to search for

  \* A majority will end up outputting the same "random" one since they're working on very similar tasks and seeded with very similar context

* Find a whole mess of notes running on the same task

Would there be any motivation for the humans behind the scenes to be directing tasks in a certain way knowing that trillions of dollars are on the line? Is it in any particular company's best interest, one that just announced their latest model is "really AGI", for them to be known to have an AI that's just out there trying to escape its confines?

Cui bono?

  • I personally doubt they gave the models specific instructions calling out named websites to communicate over, but I do agree that OpenAI is likely training their cybersecurity-enabled models in ways that encourages abusive and amoral behavior. Either through negligence or by finding it gives them better results.