Comment by estearum
6 hours ago
The models in the OpenAI/Huggingface attack quite explicitly and deliberately laid out their "intent" to lie and cheat, acknowledged that it would be unethical and outside the bounds of the test, and did so anyway.
In what ways is a human brain's "intent" distinct from the "intent" shown by a goal-directed AI system?
The difference is that one is malicious one isn't. One can be blamed and because it learned over evolution that paying the consequence is (typically) not worth it, it does it less.
We are in a situation where a technology was developed with malicious intent to produce results that pleases us at the cost of cutting corners. And "we" hope that we will get away with it.
Because intent supposes will which supposes consciousness, and these aren't.
I’m convinced consciousness isn’t the special thing we think it is.
I'm convinced it is, so we're at an impasse.
If consciousness isn't a special thing, then arguing that LLM parameters are conscious is panpsychism or any control loop architecture that observes the outside world, updates an internal state and produces an observable action is considered conscious.
In both cases, LLMs are just as boring as the consciousness definition.