Comment by inciampati

15 hours ago

Isn't finding this out from an LLM somewhat... complex and non reproducible.

Everyone who survived 7 rounds of multi model reviews and they still keep finding mediums in their PRs is not in the least surprised. These things are not oracles - they miss stuff all the time even when told to look.

  • I'm convinced that the LLMs are capable of finding everything in one shot but that's not good for token usage so they only report a few at a time.

    • you're being downvoted for the paranoid tone i think. but that is correct.

      well not token usage, but revenue. their costs for this work would have been astronomical in their own service tier because i bet the context was way larger than anything they even offer.

      tweaking context size is the main, or only, "strategy" they have for cost/revenue. and is the reason new trained versions continue to generate hype: you need data in training because you cannot have it in context

Yep :) it's VERY non-reproducible - but I haven't asked it to look for Alu repeats in the first place, I can then go and reproduce the work it's doing