Comment by CamperBob2
12 hours ago
Agreed. Anyone who thought CoT output was a reliable guide to the actual reasoning process taking place was fooling themselves from day 1.
A few hours spent playing with DeepSeek R1 should have been enough to dispel that illusion, watching it talk itself out of the right answer in its CoT (or talk itself into the wrong one) and still emit the right answer in its response to the user.
No comments yet
Contribute on Hacker News ↗