Understanding intermediate layers using linear classifier probes

Alain, Guillaume; Bengio, Yoshua

Full-text links:

Download:

Current browse context:

stat.ML

< prev | next >

new | recent | 1610

Statistics > Machine Learning

Title: Understanding intermediate layers using linear classifier probes

Authors: Guillaume Alain, Yoshua Bengio

(Submitted on 5 Oct 2016 (v1), revised 10 Oct 2016 (this version, v2), latest version 22 Nov 2018 (v4))

Abstract: Neural network models have a reputation for being black boxes. We propose a new method to understand better the roles and dynamics of the intermediate layers. This has direct consequences on the design of such models and it enables the expert to be able to justify certain heuristics (such as the auxiliary heads in the Inception model). Our method uses linear classifiers, referred to as "probes", where a probe can only use the hidden units of a given intermediate layer as discriminating features. Moreover, these probes cannot affect the training phase of a model, and they are generally added after training. They allow the user to visualize the state of the model at multiple steps of training. We demonstrate how this can be used to develop a better intuition about a known model and to diagnose potential problems.

Subjects:	Machine Learning (stat.ML); Machine Learning (cs.LG)
Cite as:	arXiv:1610.01644 [stat.ML]
	(or arXiv:1610.01644v2 [stat.ML] for this version)

Submission history

From: Guillaume Alain [view email]
[v1] Wed, 5 Oct 2016 20:59:01 GMT (4314kb,D)
[v2] Mon, 10 Oct 2016 02:33:57 GMT (10356kb,D)
[v3] Fri, 14 Oct 2016 18:47:19 GMT (4453kb,D)
[v4] Thu, 22 Nov 2018 23:40:00 GMT (4238kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> stat > arXiv:1610.01644v2

Download:

Current browse context:

Change to browse by:

References & Citations

Bookmark

Statistics > Machine Learning

Title: Understanding intermediate layers using linear classifier probes

Submission history