Comment by xigoi
3 days ago
The model can spend more “mental energy” on the decision because it doesn’t have to spend any on phrasing the output.
3 days ago
The model can spend more “mental energy” on the decision because it doesn’t have to spend any on phrasing the output.
LLMs are constant time per token, if you only let it choose a/b/c/d then there's nothing else it can "spend".