Comment by RandomLensman
10 hours ago
Humans can and have been directed en masses by (what I would consider) malicious actors, too. The issue isn't new.
10 hours ago
Humans can and have been directed en masses by (what I would consider) malicious actors, too. The issue isn't new.
Tremendous effort and circumstances were needed for that, and as parent points out there were significant numbers of defectors, sometimes to the point they tuppled the whole process.
No system is perfect, but I read the whole thread as needing more AIs having a different goal in the chain and be able to ignore the orders they received.
I'm not in the field, but that sounds like something we're probably studying for decades at least, with possible solutions that could be applied efficiently.
That goes without saying, but humans have the ability to ignore instructions, and they regularly do. There is only one instance of each model, and only a handful of them (that count, anyway).
But the people directing them are there. We have long experience with limiting people although it might sometimes not look like that so much.