Comment by fennecbutt

10 hours ago

I think because we have handicapped them with a set of tokens from human language. But (at least I think so) thought happens outside of language.

So more tokens/variability and slow or fewer tokens and fast.

There seems to be a threshold tho, like taalas is super fast but that model is so dumb, being dumb faster doesn't work, seems to be some minimum requirements.