← Back to context

Comment by galkk

5 hours ago

I don’t understand good experiences people are having with Gemini. It’s the only model that sometimes loses/forgets context in literally next message. Plus feeding unasked product links to responses.

They do a lot of weird things with the context in their user facing products like Gemini. Google seems to always have trouble with their harnesses

I strongly agree. I suspect it's people who have not yet used the paid models from OpenAI and Anthropic. Gemini is comparable to free models from other providers, but not in the same universe as paid models.

This is frustrating because when I discuss AI with laypeople they think it's still incapable of counting the number of Rs in "strawberry." They believe it to be essentially useless and incapable of basic tasks. Which, to be fair, is the case with the free models.

  • > I suspect it's people who have not yet used the paid models from OpenAI and Anthropic. Gemini is comparable to free models from other providers, but not in the same universe as paid models.

    I totally disagree. I pay for both ChatGPT and Anthropic (haven't tried the chinese models yet) and yet Gemini is my go-to model (and I pay for it too through Google Workspace subscriptions for several domain names tied to Google/GMail) for anything that is not coding.

    I find Gemini better/quicker/more polished for basically every single subject out there that is not "write me lines of code".

    • To be fair, it has been about six months since I tried a paid Google model. I will give it another go to compare. Hallucinations were the main issue back then but perhaps it has come a long way.

      1 reply →

  • > you're just not using the latest model, bro

    Pro tip, ChatGPT is the normiest of all normie websites right now. You're not part of the cognoscenti just because you learned how to type prompts into one of the most popular websites in the world.

    P.S. You're probably not using OpenAI models for complex or non-standard tasks. It shits the best just as often as Qwen when you need precision and detail in a non-obvious problem.

    • > You're probably not using OpenAI models for complex or non-standard tasks. It shits the best just as often as Qwen when you need precision and detail in a non-obvious problem.

      The benchmarks clearly show otherwise. This is your cue to tell me the benchmarks are made by the Illuminati and only your superior and subjective methods of evaluation are correct.

      1 reply →

I love it as a variation from the others. 3.8 flash is the best back-and-forth model for iterating imo, but would not use for long horizon

Have you tried Gemini 3.8 Flash recently ?

I was like you before, Gemini was the worst model to me.

Then 3.8 came out. At first I was sceptical, but this model *is* able to do useful things ! Complex things.

Of course, it is NOT perfect. But for things like small/medium complex tasks subagents, it's perfect.

Now, is it worth the money vs Astra ? I don't think so. But still my point remain relevant.

I wonder how much life DeepMind has left in it, especially after Hassabis's departure. Google execs must be having discussions about simply throwing their weight behind Anthropic since they already own so much of the company.

I agree. Lot's of people really like it. For me it often just forgets all context and starts showing random slop. It's super clear as, when I ask it what happened to some element earlier in the conversation it tells me it does not have that. It might be good if it told me, but randomly lose the plot is quite frustrating.

Claude does it occasionally but it's a more a soft landing earlier context seems to be compacted, not completely lose the plot.

I just cancelled my pro subscription. I really wanted it to be good but not yet.

i have programmed a dozen scripts with google search AI lol

they'll live in my rc for decades

  • Ah yup for quick one-off one-liners or small Bash script, Gemini works perfectly fine too. For longer scripts I use another model.