No frame left behind: Full Video Action Recognition

Xin Liu; Silvia L. Pintea; Fatemeh Karimi Nejadasl; Olaf Booij; Jan C.; van Gemert

arXiv:2103.15395·cs.CV·March 30, 2021

No frame left behind: Full Video Action Recognition

Xin Liu, Silvia L. Pintea, Fatemeh Karimi Nejadasl, Olaf Booij, Jan C., van Gemert

PDF

Open Access 1 Repo

TL;DR

This paper introduces a full video action recognition method that processes all frames efficiently by clustering frame activations, outperforming traditional sampling techniques on multiple datasets.

Contribution

It proposes a novel end-to-end trainable approach that clusters all video frames based on activation similarity, enabling full video analysis without prohibitive computational costs.

Findings

01

Outperforms existing heuristic frame sampling methods

02

Efficient clustering-based aggregation of all frames

03

Validated on multiple benchmark datasets

Abstract

Not all video frames are equally informative for recognizing an action. It is computationally infeasible to train deep networks on all video frames when actions develop over hundreds of frames. A common heuristic is uniformly sampling a small number of video frames and using these to recognize the action. Instead, here we propose full video action recognition and consider all video frames. To make this computational tractable, we first cluster all frame activations along the temporal dimension based on their similarity with respect to the classification task, and then temporally aggregate the frames in the clusters into a smaller number of representations. Our method is end-to-end trainable and computationally efficient as it relies on temporally localized clustering in combination with fast Hamming distances in feature space. We evaluate on UCF101, HMDB51, Breakfast, and…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

L-KID/Full-Video-Action-Recognition
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsHuman Pose and Action Recognition · Anomaly Detection Techniques and Applications · Diabetic Foot Ulcer Assessment and Management