Rank Pooling for Action Recognition

Basura Fernando; Efstratios Gavves; Jose Oramas; Amir Ghodrati; Tinne; Tuytelaars

arXiv:1512.01848·cs.CV·May 17, 2016

Rank Pooling for Action Recognition

Basura Fernando, Efstratios Gavves, Jose Oramas, Amir Ghodrati, Tinne, Tuytelaars

PDF

1 Repo

TL;DR

This paper introduces a function-based temporal pooling method called rank pooling for action recognition, capturing video dynamics by learning to rank frame features, leading to improved recognition accuracy across various benchmarks.

Contribution

The paper presents a novel rank pooling technique that models temporal evolution in videos, providing a robust and interpretable representation for action recognition.

Findings

01

Rank pooling improves recognition accuracy by 7-10% over average pooling.

02

The method is compatible with existing appearance and motion features.

03

Rank pooling is fast, easy to implement, and effective across diverse action recognition tasks.

Abstract

We propose a function-based temporal pooling method that captures the latent structure of the video sequence data - e.g. how frame-level features evolve over time in a video. We show how the parameters of a function that has been fit to the video data can serve as a robust new video representation. As a specific example, we learn a pooling function via ranking machines. By learning to rank the frame-level features of a video in chronological order, we obtain a new representation that captures the video-wide temporal dynamics of a video, suitable for action recognition. Other than ranking functions, we explore different parametric models that could also explain the temporal changes in videos. The proposed functional pooling methods, and rank pooling in particular, is easy to interpret and implement, fast to compute and effective in recognizing a wide variety of actions. We evaluate our…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

https://bitbucket.org/bfernando/videodarwin
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.