Comment by pfdietz

12 hours ago

Specifically: bearish on LLMs generally, not bearish on LLMs for pure math.

yes, huge for pure math and activities that look like it.

  • Doesn't really even need to look like it. If you can verify rewards, RLVR will optimize really really well. If you can't... it's a struggle. There are probably fewer fields where you can verify rewards than one might hope.

    • > There are probably fewer fields where you can verify rewards than one might hope.

      2 tasks I've done today that I believe robots are nowhere near being able to do: Cleaning my wardrobe and draining bad fuel out of my generator. As in generic use cases.