Reinforcement learning with prediction-based rewards
OpenAI's Random Network Distillation (RND) demonstrates a breakthrough in curiosity-driven exploration for RL agents, surpassing human performance on a notoriously difficult benchmark. This advances the capability of AI systems to learn in sparse-reward environments without explicit guidance—a critical foundation for real-world agent deployment.
Why it ranks · · Random Network Distillation (RND) method developed · October 2018
Read full story