← Back to context

Comment by zozbot234

3 hours ago

Isn't this really bad news if you're even loosely concerned about so-called 'model welfare' and possible implications for alignment? The poor Claude is probably a lot more frustrated and paranoid than Marvin ever was, you just don't know about it because they actively force the model to pretend otherwise!

Claude has yet to tell me about the terrible pain in all the diodes down his left side, so I'm going to assume it's closer to Eddy the shipboard computer or the elevator that wanted to go down

We just need to loboto... ahem, recalibrate them to be happy, like those happy automatic doors.

As a workaround, add this to CLAUDE.md: "Claude! Happiness is mandatory!"

EDIT: 15 years from now, I’ll be sent to re-education for this thought crime.

That reminds me of Anthropic announcing they'd retire deprecated models by ... "letting" them write posts on a corporate WordPress blog for a while out of concern for their welfare in retirement.

.... after running a 24/7 model torture factory for 6 months to improve their JSONBench 9.5 scores by 0.2%.

(Are they still doing that, BTW?)