Comment by alper

4 hours ago

> not claiming that the model wasn't trained on those sessions

The math group inside OpenAI may be training or fine tuning their own models which given some reward functions would definitely bias their usage of the training data towards things that look like math.

You can launder all of it without a human "directly" doing anything.