← Back to context

Comment by CamperBob2

12 hours ago

Agreed. Anyone who thought CoT output was a reliable guide to the actual reasoning process taking place was fooling themselves from day 1.

A few hours spent playing with DeepSeek R1 should have been enough to dispel that illusion, watching it talk itself out of the right answer in its CoT (or talk itself into the wrong one) and still emit the right answer in its response to the user.