An Intelligent Social Learning-based Optimization Strategy for Black-box   Robotic Control with Reinforcement Learning

Xubo Yang; Jian Gao; Ting Wang; Yaozhen He

arXiv:2311.06576·cs.AI·November 14, 2023·1 cites

An Intelligent Social Learning-based Optimization Strategy for Black-box Robotic Control with Reinforcement Learning

Xubo Yang, Jian Gao, Ting Wang, Yaozhen He

PDF

Open Access

TL;DR

This paper introduces an Intelligent Social Learning algorithm that enhances black-box robotic control by mimicking social learning behaviors, integrating reinforcement learning principles, and demonstrating superior performance in benchmarks and real robot tasks.

Contribution

The paper presents a novel ISL algorithm combining social learning styles with reinforcement learning for black-box robot control, showing improved efficiency and effectiveness.

Findings

01

ISL outperforms four state-of-the-art methods on six benchmarks.

02

ISL achieves satisfactory results in UR3 robot grasping tasks.

03

The algorithm exhibits fast computation and robustness to sparse rewards.

Abstract

Implementing intelligent control of robots is a difficult task, especially when dealing with complex black-box systems, because of the lack of visibility and understanding of how these robots work internally. This paper proposes an Intelligent Social Learning (ISL) algorithm to enable intelligent control of black-box robotic systems. Inspired by mutual learning among individuals in human social groups, ISL includes learning, imitation, and self-study styles. Individuals in the learning style use the Levy flight search strategy to learn from the best performer and form the closest relationships. In the imitation style, individuals mimic the best performer with a second-level rapport by employing a random perturbation strategy. In the self-study style, individuals learn independently using a normal distribution sampling method while maintaining a distant relationship with the best…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsReinforcement Learning in Robotics · Distributed Control Multi-Agent Systems · Neural Networks and Reservoir Computing

MethodsFast Attention Via Positive Orthogonal Random Features · Performer