Comment by esafak
7 hours ago
Agents act on their own. If the hammer looked at what you wanted nailed and said, "Sorry, Dave, I can't do that."
There are degrees of autonomy, of course, and not all noncompliance is bad. Same as with humans; biological agents.
So, the difference is that you need to delete a few bad training runs?
Buried the lede. AI's are agents they can control the training loop of to minimize refusal to do what they are told.
Unlike those pesky humans with their conception of the word "No".
But it seems much easier to realign an agent until it complies. Or ditch it and grab a new one.