Comment by jdlshore
6 hours ago
This is an amazing article. The problems it describes are exactly what we found when building a production system that used LLMs to (most of the time) produce reliable results. Extensive tests are necessary, and stakeholders have no idea how their suggestions fail in production. They just see the handful of times they tried something and had it work, not the long tail of cursed results. (“How hard can it be? Why don’t you just…”)
We didn’t get to the point of self-built prompts, as the article suggests, but it’s an intriguing idea.
No comments yet
Contribute on Hacker News ↗