Comment by noname120
2 months ago
ChatGPT 5.6 Sol Pro believes that the proof is sound. Usually it’s very good at determining if proofs are correct and their mistakes (a friend of mine is a top mathematician researcher and confirmed): https://chatgpt.com/share/6a515ead-b464-83ed-b85c-c8674f56ea...
Personally this gives me additional confidence that this is the real deal.
Of course it believes the proof is sound, it wrote it. If you want to check an LLM's output, you should use a different LLM.
Your comment is not substantiated at all.
No, the comment is right. The prompt had GPT-5.6 reviewing the proof, and the result, unsurprisingly, survives review by GPT-5.6.
2 replies →
If you'd ever tried to get an LLM to review its own code, you'd know.
1 reply →
Use a human maybe.
Only people can really verify clankers.
Can't trust anything LLM since it will confidently lie too.
It can't take responsibility for verification so it can't verify.