A Talent-infused Policy-gradient Approach to Efficient Co-Design of   Morphology and Task Allocation Behavior of Multi-Robot Systems

Prajit KrisshnaKumar; Steve Paul; Souma Chowdhury

arXiv:2411.18519·cs.RO·November 28, 2024

A Talent-infused Policy-gradient Approach to Efficient Co-Design of Morphology and Task Allocation Behavior of Multi-Robot Systems

Prajit KrisshnaKumar, Steve Paul, Souma Chowdhury

PDF

Open Access

TL;DR

This paper introduces a novel talent-infused policy-gradient co-design method for optimizing robot morphology and behavior simultaneously, significantly improving multi-robot system performance in flood response tasks.

Contribution

It presents an efficient co-design framework that leverages Pareto front analysis and reinforcement learning to optimize morphology and behavior concurrently, outperforming traditional sequential approaches.

Findings

01

Co-designed systems outperform sequential design baselines.

02

Significant morphology and behavior differences between single and multi-robot systems.

03

Enhanced collective performance in flood response scenario.

Abstract

Interesting and efficient collective behavior observed in multi-robot or swarm systems emerges from the individual behavior of the robots. The functional space of individual robot behaviors is in turn shaped or constrained by the robot's morphology or physical design. Thus the full potential of multi-robot systems can be realized by concurrently optimizing the morphology and behavior of individual robots, informed by the environment's feedback about their collective performance, as opposed to treating morphology and behavior choices disparately or in sequence (the classical approach). This paper presents an efficient concurrent design or co-design method to explore this potential and understand how morphology choices impact collective behavior, particularly in an MRTA problem focused on a flood response scenario, where the individual behavior is designed via graph reinforcement…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsReinforcement Learning in Robotics