The Effect of Multi-step Methods on Overestimation in Deep Reinforcement Learning
Jan 1, 2021·
,,·
0 min read
Lingheng Meng
Rob Gorbet
Dana Kulic
Type
Publication
In 25th International Conference on Pattern Recognition (ICPR 2020), Milan, Italy

Authors
CERC Postdoctoral Fellow
Robotics and human-robot interaction researcher advancing human-centered
robot learning through deep reinforcement learning, preference learning,
and adaptive interactive systems. My work spans from foundational RL theory
to real-world deployment, including a museum installation with 60,000+
public visitors.
Authors
Authors