iterated-prisoner-s-dilemma-740f06f5·1 events·first seen Aliases: Iterated Prisoner's Dilemma
OpenAI has released Learning with Opponent-Learning Awareness (LOLA), an algorithm designed for multi-agent settings where each agent accounts for the fact that other agents are also learning. LOLA discovers self-interested yet collaborative strategies such as tit-for-tat in the iterated prisoner's dilemma. The work represents an early step toward agents capable of modeling other minds and reasoning about opponent behavior.