More on Dota 2
OpenAI demonstrates that self-play reinforcement learning can achieve superhuman performance without human training data—a fundamental breakthrough in how ML systems improve themselves. This validates a training paradigm that scales with compute, not data labeling.
Why it ranks · · Self-play system progressed from below-professional to superhuman in one month · Aug 14 – 20, 2017
Read full story