self-play-reinforcement-learning-10ea19d7·1 events·first seen Aliases: self-play reinforcement learning
OpenAI developed a bot that defeats world-class professional players in 1v1 Dota 2 matches under standard tournament rules. The system learned entirely through self-play without imitation learning or tree search. This was presented as a milestone toward AI systems that can achieve well-defined goals in complex, real-world environments involving humans.