← Back to context

Comment by amluto

13 hours ago

Anthropic seems like they, at least on some vague philosophical level, have the opposite view. They are very explicit about trying not to let the CoT enter directly into their RL process, and they encrypt the CoT traces (and recently further nerfed their API surface) to make it as difficult as practical for their customers to have any idea what their model is thinking.