The Effect of Multi-step Methods on Overestimation in Deep Reinforcement Learning

Jan 1, 2021·
Lingheng Meng
Lingheng Meng
,
Rob Gorbet
,
Dana Kulic
· 0 min read
Type
Publication
In 25th International Conference on Pattern Recognition (ICPR 2020), Milan, Italy
publications
Lingheng Meng
Authors
CERC Postdoctoral Fellow
Robotics and human-robot interaction researcher advancing human-centered robot learning through deep reinforcement learning, preference learning, and adaptive interactive systems. My work spans from foundational RL theory to real-world deployment, including a museum installation with 60,000+ public visitors.
Authors
Authors