Comment by Retr0id

10 hours ago

As a "proof" of their poor judgement, take a paragraph you like. Ask the LLM to rewrite it to make it better (which I think we both agree will not make it better), and then in a fresh session ask it which it thinks is best. It'll almost always pick its own writing, even when it sucks.

Right, I agree: that’s exactly not what I’m recommending.

  • You are specifically recommending asking the model which of two versions is better (the quote in my top-level comment).

    We both agree that they are poor "make it better" machines, but I also believe they are bad A/B testers and I'm using the former to demonstrate the latter.