Comment by hmate9
10 hours ago
This model is not for builders and engineers. DeepSWE score of 49% is behind gpt 5.4 and muse spark. It's clearly intended to be an efficient model for google gemini usage.
What is interesting is how this is announced before any Gemini Pro progress. From the outside it seems as though Google cannot keep up with other frontier models.
Who tested it on DeepSWE?
Edit: Oh it's in the other link
https://blog.google/innovation-and-ai/models-and-research/ge...
Everyone wants to announce as late as possible (i.e., last) to chart the highest. Google is in a position financially to take a hit for these last few months.