Comment by _heimdall
10 hours ago
> We are placing stricter requirements on alignment
This is comical. Its impossible to align a black box and that's precisely what LLMs are. It also seems impossible to align recursive text prediction algorithms, which LLMs are.
How exactly do they gate on alignment today, and how can they tighten it? Is it purely gates based on input/output pairs to check whether they're happy enough with responses regardless of how and why the response was actually chosen?
Aren't humans black boxes? Aren't humans prediction algorithms?
How do we align humans?
Well, we don't. Just look at all the wars going on right now. In light of that, does it seem like a good idea to introduce a bunch of even less aligned agents into the world?
"Grey goo" nanobots are another example of artificial agents that aren't aligned with humanity, that we should probably try to avoid creating.
We don't align humans. Just look at how often in documented human history there weren't wars going on somewhere.
We align humans via the propagation of morals and ethics, primarily through parenting and social pressure.
1 reply →
Conventionally we would call that alignment process parenting
Humans are naturally aligned with humanity.
Are they? Humans are often at war with other groups of humans. And they do absolutely terrible things to the "other" group.
1 reply →