flâneur

Cooperative AI Competitions Without Spiteful Incentives

cooperativeai.com · 2,171 words · saved by 1 readers

Caspar Oesterheld (Carnegie Mellon University) addresses an important design challenge in cooperative AI tournaments with mixed motives: how to reward participants without creating perverse incentives for spiteful or excessively competitive behaviour.

Let’s say we want to identify effective strategies for multi-agent games like the Iterated Prisoner’s Dilemma or more complex environments (like the kind of environments in Melting Pot). Then tournaments are a natural approach: let people submit strategies, and then play all these strategies against each other in a round-robin tournament. ‍ There are lots of examples of such tournaments: most famously, Axelrod ran two competitions for the iterated Prisoner’s Dilemma, identifying tit for tat as a good strategy. Less famously, there’s Alex Mennen’s open-source Prisoner’s Dilemma tournament,…

saved by

related reading