The Underappreciated Power of Vision Models for Graph Structural Understanding

Xinjian Zhao; Wei Pang; Zhongkai Xue; Xiangru Jian; Lei Zhang; Yaoyao Xu; Xiaozhuang Song; Shu Wu; Tianshu Yu

arXiv:2510.24788·cs.CV·October 30, 2025

The Underappreciated Power of Vision Models for Graph Structural Understanding

Xinjian Zhao, Wei Pang, Zhongkai Xue, Xiangru Jian, Lei Zhang, Yaoyao Xu, Xiaozhuang Song, Shu Wu, Tianshu Yu

PDF

TL;DR

This paper explores the underutilized potential of vision models in understanding graph structures, showing they excel at global pattern recognition and scale-invariant reasoning compared to traditional GNNs.

Contribution

It introduces GraphAbstract, a benchmark for evaluating models' ability to perceive global graph properties, highlighting vision models' superior performance in holistic structural understanding.

Findings

01

Vision models outperform GNNs on global structural tasks

02

Vision models generalize better across different graph sizes

03

GNNs struggle with global pattern abstraction and scalability

Abstract

Graph Neural Networks operate through bottom-up message-passing, fundamentally differing from human visual perception, which intuitively captures global structures first. We investigate the underappreciated potential of vision models for graph understanding, finding they achieve performance comparable to GNNs on established benchmarks while exhibiting distinctly different learning patterns. These divergent behaviors, combined with limitations of existing benchmarks that conflate domain features with topological understanding, motivate our introduction of GraphAbstract. This benchmark evaluates models' ability to perceive global graph properties as humans do: recognizing organizational archetypes, detecting symmetry, sensing connectivity strength, and identifying critical elements. Our results reveal that vision models significantly outperform GNNs on tasks requiring holistic structural…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.