Generalization and Regularization in DQN

Farebrother, Jesse; Machado, Marlos C.; Bowling, Michael

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 1810

Computer Science > Machine Learning

Title: Generalization and Regularization in DQN

Authors: Jesse Farebrother, Marlos C. Machado, Michael Bowling

(Submitted on 29 Sep 2018 (this version), latest version 17 Jan 2020 (v3))

Abstract: Deep reinforcement learning (RL) algorithms have shown an impressive ability to learn complex control policies in high-dimensional environments. However, despite the ever-increasing performance on popular benchmarks like the Arcade Learning Environment (ALE), policies learned by deep RL algorithms can struggle to generalize when evaluated in remarkably similar environments. These results are unexpected given the fact that, in supervised learning, deep neural networks often learn robust features that generalize across tasks. In this paper, we study the generalization capabilities of DQN in order to aid in understanding this mismatch between generalization in deep RL and supervised learning methods. We provide evidence suggesting that DQN overspecializes to the domain it is trained on. We then comprehensively evaluate the impact of traditional methods of regularization from supervised learning, $\ell_2$ and dropout, and of reusing learned representations to improve the generalization capabilities of DQN. We perform this study using different game modes of Atari 2600 games, a recently introduced modification for the ALE which supports slight variations of the Atari 2600 games used for benchmarking in the field. Despite regularization being largely underutilized in deep RL, we show that it can, in fact, help DQN learn more general features. These features can then be reused and fine-tuned on similar tasks, considerably improving the sample efficiency of DQN.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Machine Learning (stat.ML)
Cite as:	arXiv:1810.00123 [cs.LG]
	(or arXiv:1810.00123v1 [cs.LG] for this version)

Submission history

From: Marlos C. Machado [view email]
[v1] Sat, 29 Sep 2018 00:52:34 GMT (1784kb,D)
[v2] Wed, 30 Jan 2019 17:59:21 GMT (7371kb,D)
[v3] Fri, 17 Jan 2020 23:25:22 GMT (7157kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:1810.00123v1

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Generalization and Regularization in DQN

Submission history