Comment by vb-8448

8 hours ago

> Historically, many of these problems were bottlenecked by human attention. Someone had to care enough to spend hours or days reading obscure material, testing unpromising ideas, tracing references, and trying things that might go nowhere

I wonder how many of the recent results are due to the fact that very few looked at the problem to start with. Still great results, but the general impression is that it's more about the so many low-hanging fruits than the actual capability.

Let's not normalize the achievement. Just a couple years ago this would be considered science fiction. We can argue that 2026 AI can't solve the very toughest cryptograms, but the fact it can solve nontrivial ones is already magical.

Now on to the Voynich Manuscript :)

  • "AI solves niche thing you've never heard of" is a daily headline at this point. What's genuinely cool isn't that AI managed to solve some specific problem only a handful of people even cared about, it's that humanity can now cheaply clean up its backlog of such things.*

    That does not mean that specific instances of it are still very interesting though. This article is the "I had claude vibecode a thermostat for my bathtub" of cryptography.

    * And in this case I'm not sure it even meets that bar. For all we know a couple readers back when the book released had a delightful afternoon with it, solved the riddle, then forgot about it.

  • Someone wrote a prompt, that included instructions for finding the problem itself and got handed a solution by a machine trained on all available text. I don’t see any achievement for the prompter. As for the machine, we can’t keep being perpetually shocked 24x7. It’s tiring (unless if we’re being paid for it)

  • It is indeed absolutely incredible that it can solve these puzzles given plaintext instructions with very little context.

  • No, he's right. Actually, let's have a bit of sobriety when discussing the achievements of the most heavily marketed technology of all time, as published by an organisation that stands to benefit financially from the public perception of that technology. The discussion of "what made this problem low hanging fruit" is much more interesting, imo, than just breathlessly joining the hype train.

    • Thank you.

      More money than the GDP 90% of the sovereign countries around the world is hanging in the balance, and people are taking everything OpenAI and Anthropic are saying at face value as if this isn't the financial / marketing equivalent of war, assuming they they wouldn't use every legal and shady tactic, bending every truth available to them to sway the balance of public opinion in their favor. It makes me feel like I'm living in the twilight zone. People need to wake up.

      1 reply →

Yes, even many of the proofs seem to be extremely long and complicated. The Navier-Stokes proof is 57 pages of very dense math and a pretty crazy amount of code: https://github.com/openai/NavierStokesAndEuler/tree/main/Nav...

Given the close relationship between compression and intelligence, I'm somewhat surprised at how poorly the cutting edge models do with being concise.

  • You know, the first time you navigate somewhere (if you don't already have perfect directions) will probably be the longest route you'll ever take to get there

    For Earth, the proof presented for NS is just our first attempt navigating from our previously known facts to the proof.

    I expect we will be able to shorten it dramatically (most likely with human and AI insights), but I don't think we should read too much into the length. If you want a similar point of comparison, see the original proof (by humans) of Fermat's last theorem. It has been shortened significantly. This is normal.

  • >I'm somewhat surprised at how poorly the cutting edge models do with being concise.

    because they're not intelligent in the sense you're hinting at (conceptual integrity or generalization) but they are as the name suggests, large. Like comparing a forklift to a human. It's easier to bulldoze through a lot of things than tie your shoes.

    If we weren't quite as impoverished conceptually and still had the vocabulary of the Catholics we'd recognize this as ratio (discursive knowledge) vs Intellectus (apprehending knowledge)

    • What an incredibly useless comment. You state a conclusion as fact without any supportive reasoning/evidence.

      Prove that human intellect is different and that we solve problems using fundamentally different processes. I’m waiting.