← Back to context

Comment by azan_

5 days ago

Was the last time you've used LLM two years ago? Current SOTA makes good judgement calls and works well even with shitty prompts. And current SOTA is the worst these models will be.

I use unreleased models at work every day. When they work they work great. But almost every day I run into a case where they make a mistake. Also, the humans prompting them may be themselves wrong or understand the problem wrongly or not include important context. That's not a solvable problem either really.