Learning Discrete Weights Using the Local Reparameterization Trick

Shayer, Oran; Levi, Dan; Fetaya, Ethan

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 1710

Computer Science > Machine Learning

Title: Learning Discrete Weights Using the Local Reparameterization Trick

Authors: Oran Shayer, Dan Levi, Ethan Fetaya

(Submitted on 21 Oct 2017 (v1), last revised 2 Feb 2018 (this version, v3))

Abstract: Recent breakthroughs in computer vision make use of large deep neural networks, utilizing the substantial speedup offered by GPUs. For applications running on limited hardware, however, high precision real-time processing can still be a challenge. One approach to solving this problem is training networks with binary or ternary weights, thus removing the need to calculate multiplications and significantly reducing memory size. In this work, we introduce LR-nets (Local reparameterization networks), a new method for training neural networks with discrete weights using stochastic parameters. We show how a simple modification to the local reparameterization trick, previously used to train Gaussian distributed weights, enables the training of discrete weights. Using the proposed training we test both binary and ternary models on MNIST, CIFAR-10 and ImageNet benchmarks and reach state-of-the-art results on most experiments.

Comments:	ICLR 2018
Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1710.07739 [cs.LG]
	(or arXiv:1710.07739v3 [cs.LG] for this version)

Submission history

From: Oran Shayer [view email]
[v1] Sat, 21 Oct 2017 02:06:09 GMT (129kb,D)
[v2] Sat, 28 Oct 2017 00:48:21 GMT (205kb,D)
[v3] Fri, 2 Feb 2018 12:20:03 GMT (501kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:1710.07739

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Learning Discrete Weights Using the Local Reparameterization Trick

Submission history