Models don't have a sense of time, and wasting resources (token spend) is something that it's not clear they're optimized against