Return of the Devil in the Details: Delving Deep into Convolutional Nets

Ken Chatfield; Karen Simonyan; Andrea Vedaldi; Andrew Zisserman

arXiv:1405.3531·cs.CV·November 6, 2014·670 cites

Return of the Devil in the Details: Delving Deep into Convolutional Nets

Ken Chatfield, Karen Simonyan, Andrea Vedaldi, Andrew Zisserman

PDF

Open Access 1 Repo

TL;DR

This paper provides a comprehensive evaluation of CNNs compared to shallow methods, revealing key properties, implementation details, and the impact of data augmentation, with publicly available source code.

Contribution

It offers a rigorous comparison of CNN architectures with shallow methods, highlighting properties like dimensionality reduction and the transferability of data augmentation techniques.

Findings

01

CNN output dimensionality can be reduced without performance loss

02

Data augmentation benefits shallow methods similarly to CNNs

03

Deep and shallow methods share useful properties

Abstract

The latest generation of Convolutional Neural Networks (CNN) have achieved impressive results in challenging benchmarks on image recognition and object detection, significantly raising the interest of the community in these methods. Nevertheless, it is still unclear how different CNN methods compare with each other and with previous state-of-the-art shallow representations such as the Bag-of-Visual-Words and the Improved Fisher Vector. This paper conducts a rigorous evaluation of these new techniques, exploring different deep architectures and comparing them on a common ground, identifying and disclosing important implementation details. We identify several useful properties of CNN-based representations, including the fact that the dimensionality of the CNN output layer can be reduced significantly without having an adverse effect on performance. We also identify aspects of deep and…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

tzing/t-cnn
pytorch

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Neural Network Applications · Advanced Image and Video Retrieval Techniques · Domain Adaptation and Few-Shot Learning