Comment by usernametaken29
9 hours ago
> You can ask a model for output directly and stop
That’s still several orders of magnitudes too slow to fit fast vs slow. Think of 30ms vs 3-4 seconds to get an idea of what we’re talking about here
9 hours ago
> You can ask a model for output directly and stop
That’s still several orders of magnitudes too slow to fit fast vs slow. Think of 30ms vs 3-4 seconds to get an idea of what we’re talking about here
That’s a function of the amount of processing power involved not the underlying architecture of decision making.
In terms of making an LLM faster but not in terms of meta-cognition. System 1 thinking as defined by Kahneman doesn’t have 100000x more compute than System 2, it is actually the opposite. That completely contradicts your claim
>System 1 thinking as defined by Kahneman doesn’t have 100000x more compute than System 2, it is actually the opposite.
When are you measuring?
Systems 1 thinking is closer to precomputed tables in some ways. That is by evolution or massive amounts of training your neural network has a narrow fast path it can execute with as little compute at execution as needed.
1 reply →
The ratio between a single pass and multiple passes is unchanged when you throw more processing power at both.
In a human 30ms vs 3-4 seconds is a 1:100 ratio. Single vs multiple passes with an LLM varies but a 1:100 ratio isn’t unrealistic. So with enough compute and the right workload single vs multiple pass LLM could sit in that exact same 30ms vs 3-4 second timeframe.
2 replies →