Comment by not_paid_by_yt

2 days ago

Probably to a degree, I have found Gemini to be the least dis-likable of the models from the big 3 on that front. I wonder if the poor English comprehension of Deepseek-v4-pro and K3 is because of alleged distillation of Claude (speaking of why doesn't anyone distill openAI, are they just dramatically more competent at stopping API use that breaks their terms?).

V4-pro in particular seems very capable, but will just dramatically completely misunderstand user intent, it seems almost like it wasn't trained at all on non LLM generated instructions mid conversation.

I really like the clean, neutral, not over-keen way Gemma writes, and I guess my unwieldy thesis is that the way it writes is in part a consequence of the culture of the team being international.

Then again I quite like the way Muse Glimmer writes (and thinks)! It's sparky without seeming forced or insincere.