← Back to context

Comment by kay_o

8 hours ago

on older gemini models ide have to actively give them encouragement and/or easy bait problems that they can correctively solve without issue to avoid runaway spiraling into "i'm useless and i want to kms" behaviour with complex use case.

I have not seen this in other models.

I assumed it was more because the LLM might echo an understandable human claim of "if it's been unsolved for 370 years, it's unlikely to be solved now/likely to need expert knowledge", which is probably a mindset that appears in its training data.

The LLM likely needs to be reminded of its abilities.

  • > The LLM likely needs to be reminded of its abilities

    Like when it tells you something is 3 days of work but it can do it with some degree of guidance in a couple hours

    • If it’s 3 days it’s something like 15 minutes, if it’s 3 weeks, that takes a couple hours lol. Seems like there’s some sanity to the estimates after all when you think about it, it’s just the scale it gets wrong due to estimating human time.