Interpreting Context of Images using Scene Graphs

Himangi Mittal; Ajith Abraham; Anuja Arora

arXiv:1912.00501·cs.CV·December 3, 2019

Interpreting Context of Images using Scene Graphs

Himangi Mittal, Ajith Abraham, Anuja Arora

PDF

TL;DR

This paper proposes a scene graph-based model to interpret image context by representing objects and their relationships, aiding in understanding and applications like image retrieval and captioning.

Contribution

It introduces a novel approach that combines visual and semantic cues in scene graphs and uses SVMs to detect object relations, enhancing image understanding.

Findings

01

Effective scene graph representation of images.

02

Improved relation detection using combined cues.

03

Potential applications in image captioning and retrieval.

Abstract

Understanding a visual scene incorporates objects, relationships, and context. Traditional methods working on an image mostly focus on object detection and fail to capture the relationship between the objects. Relationships can give rich semantic information about the objects in a scene. The context can be conducive to comprehending an image since it will help us to perceive the relation between the objects and thus, give us a deeper insight into the image. Through this idea, our project delivers a model that focuses on finding the context present in an image by representing the image as a graph, where the nodes will the objects and edges will be the relation between them. The context is found using the visual and semantic cues which are further concatenated and given to the Support Vector Machines (SVM) to detect the relation between two objects. This presents us with the context of…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.