Degree-Quant: Quantization-Aware Training for Graph Neural Networks

Shyam A. Tailor; Javier Fernandez-Marques; Nicholas D. Lane

arXiv:2008.05000·cs.LG·March 16, 2021·20 cites

Degree-Quant: Quantization-Aware Training for Graph Neural Networks

Shyam A. Tailor, Javier Fernandez-Marques, Nicholas D. Lane

PDF

Open Access 1 Video

TL;DR

Degree-Quant introduces a novel quantization-aware training method for GNNs that maintains high accuracy with low-precision arithmetic, enabling faster inference without sacrificing performance.

Contribution

It proposes a new architecturally-agnostic quantization-aware training method for GNNs that improves accuracy and generalization over existing baselines.

Findings

01

INT8 models match FP32 performance in most cases

02

INT4 models achieve up to 26% gains over baselines

03

up to 4.7x CPU speedup with INT8 inference

Abstract

Graph neural networks (GNNs) have demonstrated strong performance on a wide variety of tasks due to their ability to model non-uniform structured data. Despite their promise, there exists little research exploring methods to make them more efficient at inference time. In this work, we explore the viability of training quantized GNNs, enabling the usage of low precision integer arithmetic during inference. We identify the sources of error that uniquely arise when attempting to quantize GNNs, and propose an architecturally-agnostic method, Degree-Quant, to improve performance over existing quantization-aware training baselines commonly used on other architectures, such as CNNs. We validate our method on six datasets and show, unlike previous attempts, that models generalize to unseen graphs. Models trained with Degree-Quant for INT8 quantization perform as well as FP32 models in most…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

Degree-Quant: Quantization-Aware Training for Graph Neural Networks· slideslive

Taxonomy

TopicsAdvanced Graph Neural Networks · Advanced Neural Network Applications · Stochastic Gradient Optimization Techniques