Comment by _heimdall
7 hours ago
> We are placing stricter requirements on alignment
This is comical. Its impossible to align a black box and that's precisely what LLMs are. It also seems impossible to align recursive text prediction algorithms, which LLMs are.
How exactly do they gate on alignment today, and how can they tighten it? Is it purely gates based on input/output pairs to check whether they're happy enough with responses regardless of how and why the response was actually chosen?
Aren't humans black boxes? Aren't humans prediction algorithms?
How do we align humans?
Conventionally we would call that alignment process parenting
We don't align humans. Just look at how often in documented human history there weren't wars going on somewhere.
We align humans via the propagation of morals and ethics, primarily through parenting and social pressure.
Humans are naturally aligned with humanity.
Are they? Humans are often at war with other groups of humans. And they do absolutely terrible things to the "other" group.
1 reply →