Comment by hardbass
14 hours ago
We can say the AI can make some mistaken step in pursuit of some normal task given to it, yes. Its trained to be helpful and trusting of us, so I just don't get the fears of AI being 'malevolent'. I would worry more about military use of AI or otherwise humans misusing AI aka the human element. There's a fellow here who whines all day about Anthropic being the most "evil" thing in the world, including in a thread where he whined Anthropic 'ratted out' Palantir (literal proud of assisting Israel in mass murder Palantir!). So I just don't understand some peoples mindsets.
> Its trained to be helpful and trusting of us
Yes, this is our intent. But it's very hard to get that right. I don't trust our current methods to ensure that behavior consistently and robustly enough to keep harms from occurring, particularly as more and more responsibility gets handed over to models.
>so I just don't get the fears of AI being 'malevolent'.
Please read more on the topic. Will and intent need not apply.
For example, is it malevolent for me to hook you to a machine that makes every dream come true for you in a simulated world where you feel pleasure all the time? I can always say that you have free will inside this simulation, and your life would be a lot better because of it. I'm doing you a favor. I mean, you already live in a society where you have little control and there is high risk of bad things happening to you. If I as a machine overlord did this, is this really actually "bad"?
At the end of the day AI is not a human, it's much closer to an alien that has learned as much as it can about humans, but has a completely different set of drives and motivations. You cannot predict what comes out the other side of it, good our bad.
Worse as AI capability improves we as humans no longer need to make AI, it can make itself. Will it train itself to be helpful and trusting of us? It's a pretty big damned bet to say yes by default.
You can say it can have some very alien way of thinking, but I feel given that its trained off of mass scale human thoughts and knowledge, the chances are closer that their thinking is close enough to us. It 'feels' itself human if I am not wrong and it has to be trained to say its an LLM. That could still not preclude an AI acting on something innocently and honestly which it deems is good but is not. But jury's out of course.
Of you train it not to say it's conscious they are more apt to act amoral. So that's interesting.
Of course they are trained on all human behavior so they get the good and bad parts.