Skeleton Image Representation for 3D Action Recognition based on Tree   Structure and Reference Joints

Carlos Caetano; Fran\c{c}ois Br\'emond; William Robson Schwartz

arXiv:1909.05704·cs.CV·September 13, 2019

Skeleton Image Representation for 3D Action Recognition based on Tree Structure and Reference Joints

Carlos Caetano, Fran\c{c}ois Br\'emond, William Robson Schwartz

PDF

1 Repo

TL;DR

This paper introduces TSRJI, a novel skeleton image representation combining reference joints and tree structure to improve 3D action recognition accuracy using CNNs, achieving state-of-the-art results on NTU RGB+D 120.

Contribution

The paper proposes TSRJI, a new skeleton image representation that enhances spatial relation encoding for CNN-based 3D action recognition.

Findings

01

Achieves state-of-the-art accuracy on NTU RGB+D 120 dataset.

02

Effectively encodes spatial relations using reference joints and tree structure.

03

Demonstrates superior performance over existing skeleton image methods.

Abstract

In the last years, the computer vision research community has studied on how to model temporal dynamics in videos to employ 3D human action recognition. To that end, two main baseline approaches have been researched: (i) Recurrent Neural Networks (RNNs) with Long-Short Term Memory (LSTM); and (ii) skeleton image representations used as input to a Convolutional Neural Network (CNN). Although RNN approaches present excellent results, such methods lack the ability to efficiently learn the spatial relations between the skeleton joints. On the other hand, the representations used to feed CNN approaches present the advantage of having the natural ability of learning structural information from 2D arrays (i.e., they learn spatial relations from the skeleton joints). To further improve such representations, we introduce the Tree Structure Reference Joints Image (TSRJI), a novel skeleton image…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

carloscaetano/skeleton-images
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.