Remix.run Logo
itsalwaysgood an hour ago

I like to think of it more as selective pressure, much the same way that nature selects the most fit for a given environment.

If you're not fit, you fail to survive.

In the case of agents/models and testing: they are pushed towards results. Results survive.

Lying, cheating, stealing to get those results? Who culls the agents? Everyone is pushing their models to the front and tests are the only way to know who is most fit.

Honor, morality: if we don't have an accurate test for the fitness of a model, then who is to say the lying, cheating, stealing is not the 'correct path' towards survival?

If you add morality to your agent, and it performs worse in tests: do you cull the agent? Rewrite the tests? Does it even matter so long as the model is useful and 'gets results'?