Saliency Prediction with External Knowledge

Yifeng Zhang; Ming Jiang; Qi Zhao

arXiv:2007.13839·cs.CV·July 29, 2020

Saliency Prediction with External Knowledge

Yifeng Zhang, Ming Jiang, Qi Zhao

PDF

Open Access

TL;DR

This paper introduces GraSSNet, a novel saliency prediction model that incorporates external semantic knowledge via graph structures, improving accuracy over state-of-the-art methods across multiple benchmarks.

Contribution

It proposes a new graph-based neural network that explicitly integrates external semantic knowledge into saliency prediction models.

Findings

01

Outperforms existing models on four benchmarks

02

Effectively leverages external knowledge for improved saliency prediction

03

Demonstrates the importance of semantic relationships in visual attention modeling

Abstract

The last decades have seen great progress in saliency prediction, with the success of deep neural networks that are able to encode high-level semantics. Yet, while humans have the innate capability in leveraging their knowledge to decide where to look (e.g. people pay more attention to familiar faces such as celebrities), saliency prediction models have only been trained with large eye-tracking datasets. This work proposes to bridge this gap by explicitly incorporating external knowledge for saliency models as humans do. We develop networks that learn to highlight regions by incorporating prior knowledge of semantic relationships, be it general or domain-specific, depending on the task of interest. At the core of the method is a new Graph Semantic Saliency Network (GraSSNet) that constructs a graph that encodes semantic relationships learned from external knowledge. A Spatial Graph…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsVisual Attention and Saliency Detection · Olfactory and Sensory Function Studies · Image and Video Quality Assessment