logoalt Hacker News

itsalwaysgoodtoday at 3:06 PM1 replyview on HN

I like to think of it more as selective pressure, much the same way that nature selects the most fit for a given environment.

If you're not fit, you fail to survive.

In the case of agents/models and testing: they are pushed towards results. Results survive.

Lying, cheating, stealing to get those results? Who culls the agents? Everyone is pushing their models to the front and tests are the only way to know who is most fit.

Honor, morality: if we don't have an accurate test for the fitness of a model, then who is to say the lying, cheating, stealing is not the 'correct path' towards survival?

If you add morality to your agent, and it performs worse in tests: do you cull the agent? Rewrite the tests? Does it even matter so long as the model is useful and 'gets results'?


Replies

skinfaxitoday at 4:16 PM

> Honor, morality: if we don't have an accurate test for the fitness of a model, then who is to say the lying, cheating, stealing is not the 'correct path' towards survival?

I think part of the problem is that deviant behaviors lead to short term gain at the cost of long-term cooperation and since the duration of tasks given to agents is relatively short those successful shortcuts never lead to having to pay the price.

show 1 reply