| ▲ | pfdietz 4 hours ago |
| Specifically: bearish on LLMs generally, not bearish on LLMs for pure math. |
|
| ▲ | jaykru 3 hours ago | parent [-] |
| yes, huge for pure math and activities that look like it. |
| |
| ▲ | danielmarkbruce 3 hours ago | parent [-] | | Doesn't really even need to look like it. If you can verify rewards, RLVR will optimize really really well. If you can't... it's a struggle. There are probably fewer fields where you can verify rewards than one might hope. | | |
| ▲ | skydhash 2 hours ago | parent [-] | | > There are probably fewer fields where you can verify rewards than one might hope. 2 tasks I've done today that I believe robots are nowhere near being able to do: Cleaning my wardrobe and draining bad fuel out of my generator. As in generic use cases. |
|
|