Comment by mlvljr

11 hours ago

Honest take, this is a critical CVE.

How so? All 6 of the CVEs covered in the article did not actually exist when investigated

  • The comment is making fun of a Claude-ism where it becomes super “honest” about stuff.

    It’s a joke but there is an underlying real effect where this type of language is psychologically manipulative and I would guess makes people believe LLMs output more than if it didn’t use “honest” (or “load bearing” or whatever super serious important sounding word).

    • Or maybe they didn't train it that way to be manipulative (although it's certainly a plausible explanation) but simply as an accidental artifact of trying to make it give honest answers?

      LLM-generated images sometimes includes text from the prompt as literal text in the image, so perhaps this is the same sort of artifact? If they've told it to be honest, it responds by talking about being honest instead of actually being honest, because it has no actual understanding of anything.

      2 replies →

  • You're absolutely right, I have hallucinated this. Would you like to find some real CVEs next?