Hyper-Sphere Quantization: Communication-Efficient SGD for Federated Learning

Dai, Xinyan; Yan, Xiao; Zhou, Kaiwen; Yang, Han; Ng, Kelvin K. W.; Cheng, James; Fan, Yu

Full-text links:

Download:

Current browse context:

cs.IR

< prev | next >

new | recent | 1911

Computer Science > Machine Learning

Title: Hyper-Sphere Quantization: Communication-Efficient SGD for Federated Learning

Authors: Xinyan Dai, Xiao Yan, Kaiwen Zhou, Han Yang, Kelvin K. W. Ng, James Cheng, Yu Fan

(Submitted on 12 Nov 2019 (v1), last revised 25 Nov 2019 (this version, v2))

Abstract: The high cost of communicating gradients is a major bottleneck for federated learning, as the bandwidth of the participating user devices is limited. Existing gradient compression algorithms are mainly designed for data centers with high-speed network and achieve $O(\sqrt{d} \log d)$ per-iteration communication cost at best, where $d$ is the size of the model. We propose hyper-sphere quantization (HSQ), a general framework that can be configured to achieve a continuum of trade-offs between communication efficiency and gradient accuracy. In particular, at the high compression ratio end, HSQ provides a low per-iteration communication cost of $O(\log d)$, which is favorable for federated learning. We prove the convergence of HSQ theoretically and show by experiments that HSQ significantly reduces the communication cost of model training without hurting convergence accuracy.

Subjects:	Machine Learning (cs.LG); Information Retrieval (cs.IR); Machine Learning (stat.ML)
Cite as:	arXiv:1911.04655 [cs.LG]
	(or arXiv:1911.04655v2 [cs.LG] for this version)

Submission history

From: Xinyan Dai [view email]
[v1] Tue, 12 Nov 2019 03:36:09 GMT (348kb,D)
[v2] Mon, 25 Nov 2019 11:00:41 GMT (348kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:1911.04655

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Hyper-Sphere Quantization: Communication-Efficient SGD for Federated Learning

Submission history