← Back to context

Comment by ethbr1

10 hours ago

My perspective is that the addition of thinking loops to models allows sufficiently advanced ones to approximate world models.

Incredibly inefficiently because of the recursive loops ("Wait, the object is on the table. I should think about this more deeply..."), and likely instantly surpassed by large world models if/when those are shipped, but effectively enough vs non-thinking models.

LeCun calling them "world models" gives a high-level description of the desired functionality. They are Joint Embedding Predictive Architectures (with SIGReg). They might produce more useful world models, but it's yet to be seen.

This sounds like a human trying to reason about quantum mechanics. We als simplify to newtonian for day to day tasks.

  • I like this analogy. Both GenRel and QM are well beyond our experience, and although there is some intuition that comes from working with the equations over time, it is bizarre and "just calculate" often gets the correct answer faster.

    Picking the right tool or model is like picking the right problem to work on. It's actually quite hard (often you can't just try them all), but without it you will be incredibly inefficient and occasionally, fundamentally wrong.

    All models are wrong, but some are useful. -Box