Comment by kroaton
7 hours ago
Considering the fact that Google/Anthropic/OpenAI have WAY more compute and the race is this close, it's obvious that DeepSeek/GLM/Qwen teams are better or we're approaching a wall in terms of progress.
7 hours ago
Considering the fact that Google/Anthropic/OpenAI have WAY more compute and the race is this close, it's obvious that DeepSeek/GLM/Qwen teams are better or we're approaching a wall in terms of progress.
US gave China a gift by restricting GPU, they made them more resourceful. Too much money/resources is often a disadvantage.
Not only compute, but more money and people.
Also data. US companies are actively using ongoing conversations to further tweak their models. Possibly even stealing a SOTA math solution.
Often times these limitations for you to be creative. When you can't just throw more processing power at the problem you figure other things out.