Comment by ijustlovemath
1 hour ago
I just think that with the vast amounts of compute involved and the tendency to reward hack, we can't assume the steps towards that formalization are without error until full human understanding of the formalization.
Repeating the above comment: most of the statements were already formalized prior OpenAI's work. So no, the statements were not "reward hacked".