Comment by fny
2 hours ago
None of these LLMs are plastic. They lack neurodiversity. Their thought space and their traversal are likely constrained in someway that humanity's isn't as a collective.
2 hours ago
None of these LLMs are plastic. They lack neurodiversity. Their thought space and their traversal are likely constrained in someway that humanity's isn't as a collective.
Is it possible to have neuroplasticity and still keep them aligned?
how do you imagine that would work? you being able to. influence globally stored weights with some prompts? we have fine-tuning for that.
Yet I can't randomly order another person to steal a car for me, just because I tell them to. Alignment for an intelligent system is a hard problem and at this stage is seems close to unsolvable.
My guess is that we'll just ignore it and make money along the way and every 2-3 months we'll have the equivalent to "Equifax gets hacked and millions of user records are stolen", etc. (this time with the LLM itself doing the hacking at someone's behest - accidental or not).
What does this even mean. There is strong evidence of LLMs doing in context learning.
Some of the linear RNN layers in recent models are provably doing SGD in hidden space during inference
> What does this even mean. There is strong evidence of LLMs doing in context learning.
1. Is this learning persistent?
2. Do they verify these new lessons against core principles?
3. Do they and protect themselves/ignore requests if these new lessons contradict those core principles?
Humans do that from the time they're 3 years old (not that well, but they do do it).