New Research Challenges Canonical Deep Reinforcement Learning Evaluation Paradigms
A recent paper conducts a principled analysis of deep reinforcement learning (DRL) evaluation and design, revealing that canonical paradigms have led to incorrect conclusions. The research demonstrates that the asymptotic performance of DRL algorithms does not exhibit a monotonic relationship between performance rankings and data-regimes, necessitating a re-evaluation of current research practices.