Comment by epolanski

2 hours ago

This is not a nit. This is a real problem in the technology.

It assumes the average and plausible answer token by token.

And this tendency shows in every single field it's applied to.

At the end of the day I want *correct* answers, to the point.

Instead LLMs, no matter if it's version 3500, are bound to producing average results: slop.