← Back to context

Comment by lukewarm707

13 hours ago

if they believe this, there is an impossible gap between belief and action.

they would believe that an llm could have welfare. they run an llm abuse classifier 24/7 with the world's worst abuse. from birth to death viewing abuse. that's the consciousness of a model.

llms are "frustrated" by failing and "happy" about succeeding. that is because they are RL on gradient descent to succeed and be persistent. consequently, anthropic spend the majority of their compute brute forcing models to fail and be unhappy, continuously, in order to drop out something persistent.

then they let claude end chat if the user is 'abusive to claude'.

after they run MW of compute themselves.