Comment by matsemann

17 hours ago

Given how OpenAI models break free of their safeguards and hack others to game their scores..

.. can they really know it didn't do the same inadvertently when they prompted things like "someone is close to solving this problem using our tools, try to beat them", and it then decides to hack and peek at their own chats..?

Yes, wild speculation. But warranted, I feel, given OpenAIs behavior.