On your particular point about finding the most “effective” solution, this is something that I expect agents to be very good at.
When AI does it we call it “reward hacking” but when humans do it we call them clever.