On your particular point about finding the most “effective” solution, this is something that I expect agents to be very good at.
When AI does it we call it “reward hacking” but when humans do it we call them clever.
On your particular point about finding the most “effective” solution, this is something that I expect agents to be very good at.
When AI does it we call it “reward hacking” but when humans do it we call them clever.