← Back to context

Comment by PedroBatista

14 hours ago

I sometimes have that feeling too, then ask another LLM to do a code and vulnerability review and OMG: rookie mistakes, over complications and security gaps even a 1st year student would not make regularly.

So.. one more year of untreated bipolar AI psychosis I guess..

At least we are at a point where we can have AI review code and reliably find real problems. That alone is incredibly valuable.

I think these kinds of comments really need to say which LLM that is. There's an enormous difference in skill between the frontier ones and say the Google search AI.

  • Codex Luna, Terra and Sol. Claude Opus, Sonnet and sometime Fable.

    They all work, they all are "good", they all are both "smart" and commit incredible basic mistakes a fair amount of times.

    Then there's the cost situation..