Towards Consistent Long-Term Pose Generation

Yayuan Li; Filippos Bellos; Jason Corso

arXiv:2507.18382·cs.CV·July 25, 2025

Towards Consistent Long-Term Pose Generation

Yayuan Li, Filippos Bellos, Jason Corso

PDF

Open Access

TL;DR

This paper introduces a one-stage, direct pose generation method from minimal input that maintains temporal coherence and outperforms existing approaches, especially in long-term scenarios.

Contribution

A novel architecture that directly generates continuous poses from minimal context, eliminating intermediate representations and improving long-term pose generation consistency.

Findings

01

Outperforms existing methods on Penn Action and F-PHAB datasets.

02

Significantly better in long-term pose generation scenarios.

03

Maintains consistent distributions between training and inference.

Abstract

Current approaches to pose generation rely heavily on intermediate representations, either through two-stage pipelines with quantization or autoregressive models that accumulate errors during inference. This fundamental limitation leads to degraded performance, particularly in long-term pose generation where maintaining temporal coherence is crucial. We propose a novel one-stage architecture that directly generates poses in continuous coordinate space from minimal context - a single RGB image and text description - while maintaining consistent distributions between training and inference. Our key innovation is eliminating the need for intermediate representations or token-based generation by operating directly on pose coordinates through a relative movement prediction mechanism that preserves spatial relationships, and a unified placeholder token approach that enables single-forward…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsRobotic Mechanisms and Dynamics · Robot Manipulation and Learning · Robotic Path Planning Algorithms