Action-Constrained Imitation Learning

Chia-Han Yeh; Tse-Sheng Nan; Risto Vuorio; Wei Hung; Hung-Yen Wu; Shao-Hua Sun; Ping-Chun Hsieh

arXiv:2508.14379·cs.RO·August 21, 2025

Action-Constrained Imitation Learning

Chia-Han Yeh, Tse-Sheng Nan, Risto Vuorio, Wei Hung, Hung-Yen Wu, Shao-Hua Sun, Ping-Chun Hsieh

PDF

Open Access 1 Video

TL;DR

This paper introduces Action-Constrained Imitation Learning (ACIL), a novel framework for safe robot control that aligns expert demonstrations with action constraints using trajectory planning and Dynamic Time Warping, improving imitation learning performance.

Contribution

It proposes DTWIL, a trajectory alignment method that generates surrogate datasets respecting action constraints, addressing occupancy measure mismatch in ACIL.

Findings

01

DTWIL improves imitation learning performance across multiple tasks.

02

The method outperforms existing benchmark algorithms in sample efficiency.

03

Trajectory alignment via Model Predictive Control effectively handles action constraints.

Abstract

Policy learning under action constraints plays a central role in ensuring safe behaviors in various robot control and resource allocation applications. In this paper, we study a new problem setting termed Action-Constrained Imitation Learning (ACIL), where an action-constrained imitator aims to learn from a demonstrative expert with larger action space. The fundamental challenge of ACIL lies in the unavoidable mismatch of occupancy measure between the expert and the imitator caused by the action constraints. We tackle this mismatch through \textit{trajectory alignment} and propose DTWIL, which replaces the original expert demonstrations with a surrogate dataset that follows similar state trajectories while adhering to the action constraints. Specifically, we recast trajectory alignment as a planning problem and solve it via Model Predictive Control, which aligns the surrogate…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

Action-Constrained Imitation Learning· slideslive

Taxonomy

TopicsHuman Pose and Action Recognition · Multimodal Machine Learning Applications · Robot Manipulation and Learning