← Back to context

Comment by jswelker

4 hours ago

I similarly noticed Gemini absolutely refuses to look at a url when I give it one and will instead just hallucinate based on what it thinks the url is. Here I am assuming Google will have the best web capabilities in its AI.

I searched for something, it told me that according to a YouTube video, the entire point of my search was wrong. I asked it for the source, I watched the video, it never made the claims Gemini hallucinated. I asked again and it claimed it scrubbed the video and found the point it made multiple times. I said those timecodes were wrong and it admitted it couldn't actually parse videos and just guessed.

What the fuck?

  • I don't understand how this isn't considered as active malice. Like purposefully outputting random stuff. Any other sort of computer system would get lot more flak than these are getting.

    • Throughout my career, I've almost always been close enough to the user that I hear about it quickly when something is wrong. It's a tough problem that so many Google engineers are typically so far removed from end users. Or maybe it's just that a small part of the company has long subsidized the rest of the employees to the point that it doesn't matter how good their work is because they'll get paid anyway.

      1 reply →

  • When it says it can read videos you don’t trust it.

    When it says it can’t read videos you think that’s an accurate introspection on its abilities?

    (Rather than a statistically likely continuation of a conversation where one side seems to be reading videos and the other side says the read is inaccurate)

  • Typical exchange:

    Me: "You bastard."

    Gemini: "Fair callout. I should have been more up front that [has no idea what the fuck it is talking about]."

  • I still believe this AI push out of nowhere is due to the current Government in power state side. Making everyone question themselves and each other and being uncertain about facts while being inundated with techbro fake news called hallucinations is a recipe for disaster for older populations that don't 'trust but verify' like most technology inclined people. This is all by design and we'll falling for it.

    • The more simple explanation is that the current government in power is over exposed in their AI investment. The normal conservative mainstream media was already doing a great job of propagandizing older populations.

      1 reply →

    • Yup. I'm most familiar with Jill Lepore and Quinn Slobodian (and others in their respective orbits). They're both historians who've written extensively about Musk and Muskism (et al). Wild stuff.

      Apparently the plan is to use AI slop, mediated thru social medias, to defeat the woke mind virus, perpetuated by the Anti-Christ, in order to safe guard humanity's future.

      I wish I was making this up.

  • Same experience with the interaction I described in my comment... after I finally got the correct answer, I asked where it got the faulty information from. It said it had just simply fabricated it. That's actively worse than just saying "I don't know", for something that's sold to us as an easy way to look up information.

I have one google account where gemini constantly confidently hallucinates crap, and another where it doesn't (at least to the extent that other modern LLMs don't).

I take this to mean that my one account has been mistaken for a competitor and they're trying to poison its data. But who knows.

Where ever would you get that impression from the company that made its fortune by being the best at searching the web?

Literally every experience I've had with Gemini / Google Search AI answers followed this exact pattern, often repeated several times more if I remained persistent instead of just giving up.

Typical example:

"Where can I buy <thing I'm looking for that I can't find anywhere using normal search terms>?"

> You're looking for <related but different and widely available thing>. It is sold by <sites I never heard of>.

"No, that's different. I'm looking for <that thing but with the exact differences spelled out again>."

> Ah, my mistake. You're looking for <thing I described>. It is sold by <sites I never heard of but which don't actually sell it>.

"I've checked your links and none of those sites actually sell it, one doesn't even sell products and instead only offers manufacturing - but also not for what I asked you for."

> I'm sorry, my bad. Those sites don't sell what you are looking for. Instead you should check out <more sites I've never heard of>.

"Those sites sell the thing you initially thought I was asking about but not the thing I described."

> I'm sorry for the misunderstanding. You can find the thing you actually described at <yet more sites including some of the same>.

"No. None of these sites sell anything close to what I asked you for and two of them don't actually exist."

> Oh, sorry about that. You're completely right. The thing you asked me about isn't actually being sold by anyone. However you could buy <thing it first thought I meant and that wouldn't bring me any closer to solving my problem>.

(ad nauseam)

  • The few times I've resigned myself to asking Gemini or ChatGPT something I couldn't find an answer to, my experience has been the same. 100% of the time. An LLM has never, not once, given me a correct answer or not lied to me. It's always been the same experience you describe. It goes in circles "Try X... Try Y... Try X?" Until I tell it to stop telling me X or Y, and then it goes, "LOL you can't do that at all, I was just wasting your time."

    Most recently, I was considering moving away from iterm2 on MacOS, and I wanted to know if any other terminal emulator supported gestures for switching between tabs. So I asked Gemini, and it says, "Yes Ghostty supports gestures for switching between tabs."

    "Ok I just installed Ghostty and I can't find anything about gestures."

    "You need to add foo=bar to your conf file."

    "I added foo=bar to my conf file and now it's saying the conf file is invalid."

    "Sorry bro, remove foo=bar and add baz=boo to the conf file."

    "It says baz=boo is invalid too."

    "baz=boo isn't a real option. Remove that and add foo=bar to your conf file."

    "You already told me to do that and I already told you that doesn't work."

    "You shouldn't put foo=bar or baz=boo in the Ghostty conf file. Both are invalid. Ghostty doesn't support gestures. Have you considered iterm2?"

    • Gemini is utterly useless, but every other modern model would be capable of answering this correctly. If you ran a local coding agent, I wouldn’t be surprised if you could one-shot implement gesture support in Ghostty. It’d definitely configure BTT for you.

      Here’s GPT 6’s answer to the prompt “What macOS terminal apps support gestures? Include a reference to the docs on how to enable/configure them.”:

      iTerm2 supports configurable three-finger taps and swipes for switching tabs/panes, creating splits, pasting, etc. Set them up under Settings > Pointer > Bindings. Check for conflicting macOS trackpad assignments. [1]

      The others are more limited: Ghostty supports macOS lookup/Quick Look gestures [2], while WezTerm lets you bind scroll events—for example, Ctrl+scroll to change font size. [3] Neither is equivalent to iTerm2’s gesture bindings.

      For custom gestures without switching terminals, BetterTouchTool can map app-specific trackpad gestures to the terminal’s existing keyboard shortcuts. [4]

      [1] https://iterm2.com/documentation-preferences-pointer.html

      [2] https://ghostty.org/docs/features#macos

      [3] https://wezterm.org/config/mouse.html

      [4] https://docs.folivora.ai/docs/trackpad-mouse/magic-mouse-tra...

      1 reply →

Claude also has very arbitrary and confusing rules about what web pages it allows itself to look at, and how much of the page it can read. Did you know for example you’ll get a much deeper analysis if you download a PDF yourself and upload it to Claude instead of giving it the url?

This is a key reason why I actually like Grok for factual queries based on web grounding. It’s fast and reliable. Maybe it’s ignoring robots.txt? Dunno. But it works well.