We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

stat.ML

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Statistics > Machine Learning

Title: Learning a Generative Model for Validity in Complex Discrete Structures

Abstract: Deep generative models have been successfully used to learn representations for high-dimensional discrete spaces by representing discrete objects as sequences and employing powerful sequence-based deep models. Unfortunately, these sequence-based models often produce invalid sequences: sequences which do not represent any underlying discrete structure; invalid sequences hinder the utility of such models. As a step towards solving this problem, we propose to learn a deep recurrent validator model, which can estimate whether a partial sequence can function as the beginning of a full, valid sequence. This validator provides insight as to how individual sequence elements influence the validity of the overall sequence, and can be used to constrain sequence based models to generate valid sequences -- and thus faithfully model discrete objects. Our approach is inspired by reinforcement learning, where an oracle which can evaluate validity of complete sequences provides a sparse reward signal. We demonstrate its effectiveness as a generative model of Python 3 source code for mathematical expressions, and in improving the ability of a variational autoencoder trained on SMILES strings to decode valid molecular structures.
Comments: Conference paper at ICLR 2018. Link to online code in paper
Subjects: Machine Learning (stat.ML); Machine Learning (cs.LG)
Cite as: arXiv:1712.01664 [stat.ML]
  (or arXiv:1712.01664v3 [stat.ML] for this version)

Submission history

From: Jos van der Westhuizen [view email]
[v1] Tue, 5 Dec 2017 14:36:23 GMT (261kb,D)
[v2] Wed, 11 Apr 2018 19:19:59 GMT (706kb,D)
[v3] Mon, 16 Apr 2018 17:55:48 GMT (706kb,D)
[v4] Fri, 2 Nov 2018 02:19:00 GMT (706kb,D)

Link back to: arXiv, form interface, contact.