We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.DC

Change to browse by:

cs

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo ScienceWISE logo

Computer Science > Distributed, Parallel, and Cluster Computing

Title: QGTC: Accelerating Quantized Graph Neural Networks via GPU Tensor Core

Abstract: Over the most recent years, quantized graph neural network (QGNN) attracts lots of research and industry attention due to its high robustness and low computation and memory overhead. Unfortunately, the performance gains of QGNN have never been realized on modern GPU platforms. To this end, we propose the first Tensor Core (TC) based computing framework, QGTC, to support any-bitwidth computation for QGNNs on GPUs. We introduce a novel quantized low-bit arithmetic design based on the low-bit data representation and bit-decomposed computation. We craft a novel TC-tailored CUDA kernel design by incorporating 3D-stacked bit compression, zero-tile jumping, and non-zero tile reuse technique to improve the performance systematically. We incorporate an effective bandwidth-optimized subgraph packing strategy to maximize the transferring efficiency between CPU host and GPU device. We integrate QGTC with Pytorch for better programmability and extensibility. Extensive experiments demonstrate that QGTC achieves on average 2.7x speedup compared with the state-of-the-art Deep Graph Library framework across diverse settings.
Subjects: Distributed, Parallel, and Cluster Computing (cs.DC)
Cite as: arXiv:2111.09547 [cs.DC]
  (or arXiv:2111.09547v5 [cs.DC] for this version)

Submission history

From: Yuke Wang [view email]
[v1] Thu, 18 Nov 2021 06:55:12 GMT (1910kb,D)
[v2] Fri, 19 Nov 2021 05:14:09 GMT (1910kb,D)
[v3] Sat, 18 Dec 2021 01:37:15 GMT (2044kb,D)
[v4] Wed, 29 Dec 2021 02:46:19 GMT (1946kb,D)
[v5] Thu, 30 Dec 2021 03:23:33 GMT (2006kb,D)

Link back to: arXiv, form interface, contact.