← Back to context

Comment by saithound

6 hours ago

That's the ratio the widely published numbers give [1]. One does not have to believe the numbers [2], but those who do believe them are then justified to conclude that there's no moat.

Which numbers you believe is of course going to affect whether you think there's a moat or not. That's largely orthogonal to your TSMC/Samsung analogy I responded to. If you think the "moatists" are wrong because they believe the wrong numbers, that's fine, but then there's no need for the analogy.

[1] https://galileo.ai/blog/llm-model-training-cost

[2] https://medium.com/@theiand/how-can-deepseek-a-5-6-million-l...

But fundamentally, why is their cost 1/20 and is it sustainable in the next 10 years of competition?

  • Now that is a good and interesting question! Hopefully a "no-moatist" will share their reasoning.

    • Because they're distilling frontier models and that's a lot faster and cheaper than training a frontier model from scratch?

      1 reply →

    • I am not a "no-moatist" per se but one can argue their might be a plateau to how good a inference llm can become. If this is the case the playing field shifts to context, tools and harness, which are much cheaper to build an compete on.

    • Ultimately, that's what I need to be convinced. No one has put forth a good argument yet.

      Clever architecture --> Ok but OpenAI/Anthropic can use these as well and they also have very smart people with their secret clever architectures

      Distilling --> Ok but distilling means you will never be smarter than the original. Furthermore, reasoning is now hidden by private labs and they have poison pill answers for distilling if they can detect it. They will be able to detect distilling better and better.

      Cheaper electricity --> Ok this is cancelled out by their chips being much less efficient due to not having ASML EUV machine access.

      So I don't see why fundamentally their training costs are cheaper over the long term.

      I'm looking for a no-moatist to convince me.

      4 replies →