| ▲ | ingatorp 4 hours ago | |
This is a result of benchmaxxing the models to infinity. If you RL with the goal of only achieving the correct result no matter how you arrive there, then the models will try to get there using any method in their disposal, including cheating. This happens also because LLMs are black boxes that we know almost nothing on how they arrive at the result they are giving. | ||