Comment by rich_sasha

1 day ago

I’m a little confused - I thought their proofs were all driven by Lean proofs - is that not right? So even if the quality of the work is low in some metrics, it either passes the test or not..? No space for changing your mind either way.

No, in their original blog they wrote:

> As part of our GitHub repository, we are sharing formalizations of many of the proofs in Lean, a programming language that allows mathematical proofs to be checked by a computer. We will update the repository with more formalizations as we obtain them.

Meaning they published all results before checking all of them, and intended to add more Lean proofs later. In the linked post they state ~42% of the posted results now have formalized proofs, some were added, some verified, and I assume this means that some results turned out to be wrong.

  • This is extremely disappointing. It means they are sharing unproven work for PR, forcing the mathematicians community to do the verification job for them, while so-called "accelerationists" surf on the hype and help with the pro-AI propaganda.

    If your AI tool can help advance mathematical research, share the tool with mathematicians. Using it like this is irresponsible.

    "AI will kill us all": no. Greedy humans will kill us all. With AI.

    • Every option is going to lead to someone shitting on OpenAI for what seems to be a pretty huge accomplishment. There have been opinions written by some mathematicians that OpenAI should just share the work that they have now so that people who are working on any solved problems can know. Which seems reasonable to me.

      11 replies →

    • > they are sharing unproven work for PR, forcing the mathematicians community to do the verification job for them

      Lean 4 is relatively a new thing, last time I checked the formalization of undergraduate level mathematics isn't entirely done yet.

      example https://ai.math.uw.edu/projects/spring-2026/

      Lean itself is very hard to get rigorously correct, if you have every tried it yourself. I am not surprised if some AI even tries to benchmaxx Lean 4 by some loopholes

    • > forcing the mathematicians community to do the verification job for them

      What do you mean? They're publishing the Lean proofs themselves. Who's forced into anything?

      2 replies →

    • But that exactly how human mathematicians do things. They upload their research to preprint services like arXiv as they await it to be peer reviewed and accepted into a journal. Why is it okay for mathematicians to publish preprint papers, but when OpenAI does it, it's irresponsible?

      12 replies →

    • No, it's the exact opposite, actually. They were sharing them early on advice of mathematicians -- they were criticized for being opaque for too long with previous announcements. OpenAI is in ~bad faith, but this isn't a sound criticism.

      Also you are deeply confused about what accelerationism is, I believe. Sorry.

      5 replies →

    • I fail to see how this is disappointing. It was obvious from the start that it was what was happening, there is no disappointment to have!

    • oh come on. even andrew wiles made a mistake and his original proof was still instrumental for the final proof of fermat's last theorem

    • This is not propaganda, this is real AI progress and massivly too.

      And im completly lost on why you think sharing progress is irresponsible? Its not a recipe for building a nuclear weapon at home in 5 easy steps.

      THese are Math proofs.

      Either a Mathematican ignores it, or not. Thats the only risk.

      2 replies →

  • Generating Lean proof is much harder and time consuming so these errors made were probably discovered during Lean proof stage.

Only a subset contain Lean formalizations. And even for that subset, there's the potential that the formalization is semantically off (that is, it's a formalization for a slightly different problem).

If you wrote twenty million lines of Lean to verify something, my suspicion is you've been fuzzing the Lean solver rather than coming up with new math.

  • Well, fine - but my understanding is, if a fuzz-generated Lean proof is correct, that’s end of story. It can’t be “incorrect” if it “passes”.

    You might think this is not very useful, maybe - but that’s not a reason to retract..?

    • > Well, fine - but my understanding is, if a fuzz-generated Lean proof is correct, that’s end of story. It can’t be “incorrect” if it “passes”.

      It may be correct, but it might not be a proof of what OpenAI claims it to be a proof of.