Comment by wongarsu
9 hours ago
Is frontier scale larger than this? Kimi K3 seems to benchmark in the same range as Opus and Fable. I would have expected they are all in the 2-4T range, with quality of the training and architecture differences as the major differentiators
The number of active parameters is vastly different. Deepseek CEO hinted that he estimates it as an order of magnitude difference in one of his recent interviews.
> Seems to benchmark
yes, but in human usage the differences show up
Would you happen to have a link to that interview? Sounds like an interesting read.
It's related to this: https://news.ycombinator.com/item?id=49052912
It was posted to HN a few days back
1. https://news.ycombinator.com/item?id=49052912 (translated english [pdf])