On Learning Where To Look

Marc'Aurelio Ranzato

arXiv:1405.5488·cs.CV·May 22, 2014·37 cites

On Learning Where To Look

Marc'Aurelio Ranzato

PDF

Open Access

TL;DR

This paper introduces a learning-based visual recognition model that processes images through sequential glimpses, reducing computational costs and improving robustness to appearance variability, demonstrated on handwritten digit datasets.

Contribution

It presents a novel glimpse-based recognition model that scales with input complexity and learns parameters from data, addressing scalability and variability challenges.

Findings

01

Reduces computational complexity compared to full-image processing.

02

Demonstrates robustness to appearance changes.

03

Shows promising results on handwritten digit datasets.

Abstract

Current automatic vision systems face two major challenges: scalability and extreme variability of appearance. First, the computational time required to process an image typically scales linearly with the number of pixels in the image, therefore limiting the resolution of input images to thumbnail size. Second, variability in appearance and pose of the objects constitute a major hurdle for robust recognition and detection. In this work, we propose a model that makes baby steps towards addressing these challenges. We describe a learning based method that recognizes objects through a series of glimpses. This system performs an amount of computation that scales with the complexity of the input rather than its number of pixels. Moreover, the proposed method is potentially more robust to changes in appearance since its parameters are learned in a data driven manner. Preliminary experiments…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Image and Video Retrieval Techniques · Advanced Neural Network Applications · Face recognition and analysis