Balance Between Efficient and Effective Learning: Dense2Sparse Reward   Shaping for Robot Manipulation with Environment Uncertainty

Yongle Luo; Kun Dong; Lili Zhao; Zhiyong Sun; Chao Zhou; Bo Song

arXiv:2003.02740·cs.LG·March 6, 2020·5 cites

Balance Between Efficient and Effective Learning: Dense2Sparse Reward Shaping for Robot Manipulation with Environment Uncertainty

Yongle Luo, Kun Dong, Lili Zhao, Zhiyong Sun, Chao Zhou, Bo Song

PDF

Open Access

TL;DR

This paper introduces Dense2Sparse, a reward shaping method for deep reinforcement learning in robot manipulation, balancing learning speed and robustness to system uncertainty.

Contribution

The study proposes a novel reward shaping technique that combines dense and sparse rewards to improve learning efficiency and effectiveness under system uncertainty.

Findings

01

Dense2Sparse outperforms standalone dense or sparse rewards in expected reward.

02

The method demonstrates higher tolerance to system uncertainty.

03

Experimental results confirm faster convergence and robustness.

Abstract

Efficient and effective learning is one of the ultimate goals of the deep reinforcement learning (DRL), although the compromise has been made in most of the time, especially for the application of robot manipulations. Learning is always expensive for robot manipulation tasks and the learning effectiveness could be affected by the system uncertainty. In order to solve above challenges, in this study, we proposed a simple but powerful reward shaping method, namely Dense2Sparse. It combines the advantage of fast convergence of dense reward and the noise isolation of the sparse reward, to achieve a balance between learning efficiency and effectiveness, which makes it suitable for robot manipulation tasks. We evaluated our Dense2Sparse method with a series of ablation experiments using the state representation model with system uncertainty. The experiment results show that the Dense2Sparse…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsReinforcement Learning in Robotics · Robot Manipulation and Learning · Advanced Memory and Neural Computing