← Back to context

Comment by usernametaken29

8 hours ago

> You can ask a model for output directly and stop

That’s still several orders of magnitudes too slow to fit fast vs slow. Think of 30ms vs 3-4 seconds to get an idea of what we’re talking about here

That’s a function of the amount of processing power involved not the underlying architecture of decision making.

  • In terms of making an LLM faster but not in terms of meta-cognition. System 1 thinking as defined by Kahneman doesn’t have 100000x more compute than System 2, it is actually the opposite. That completely contradicts your claim

    • >System 1 thinking as defined by Kahneman doesn’t have 100000x more compute than System 2, it is actually the opposite.

      When are you measuring?

      Systems 1 thinking is closer to precomputed tables in some ways. That is by evolution or massive amounts of training your neural network has a narrow fast path it can execute with as little compute at execution as needed.

      1 reply →

    • The ratio between a single pass and multiple passes is unchanged when you throw more processing power at both.

      In a human 30ms vs 3-4 seconds is a 1:100 ratio. Single vs multiple passes with an LLM varies but a 1:100 ratio isn’t unrealistic. So with enough compute and the right workload single vs multiple pass LLM could sit in that exact same 30ms vs 3-4 second timeframe.

      2 replies →