TOCH: Spatio-Temporal Object-to-Hand Correspondence for Motion   Refinement

Keyang Zhou; Bharat Lal Bhatnagar; Jan Eric Lenssen; Gerard Pons-Moll

arXiv:2205.07982·cs.CV·October 30, 2023

TOCH: Spatio-Temporal Object-to-Hand Correspondence for Motion Refinement

Keyang Zhou, Bharat Lal Bhatnagar, Jan Eric Lenssen, Gerard Pons-Moll

PDF

Open Access

TL;DR

TOCH introduces a novel spatio-temporal representation and learning framework to refine 3D hand-object interaction sequences, improving realism and contact accuracy in motion sequences.

Contribution

The paper proposes TOCH fields, a new object-centric spatio-temporal representation, and a latent manifold learned via a temporal denoising auto-encoder for interaction refinement.

Findings

01

Outperforms state-of-the-art static interaction models

02

Produces smooth, realistic hand-object interactions

03

Effectively corrects erroneous sequences from existing methods

Abstract

We present TOCH, a method for refining incorrect 3D hand-object interaction sequences using a data prior. Existing hand trackers, especially those that rely on very few cameras, often produce visually unrealistic results with hand-object intersection or missing contacts. Although correcting such errors requires reasoning about temporal aspects of interaction, most previous works focus on static grasps and contacts. The core of our method are TOCH fields, a novel spatio-temporal representation for modeling correspondences between hands and objects during interaction. TOCH fields are a point-wise, object-centric representation, which encode the hand position relative to the object. Leveraging this novel representation, we learn a latent manifold of plausible TOCH fields with a temporal denoising auto-encoder. Experiments demonstrate that TOCH outperforms state-of-the-art 3D hand-object…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsHuman Pose and Action Recognition · Hand Gesture Recognition Systems · Robot Manipulation and Learning