We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.CL

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Computation and Language

Title: Learning Spoken Language Representations with Neural Lattice Language Modeling

Abstract: Pre-trained language models have achieved huge improvement on many NLP tasks. However, these methods are usually designed for written text, so they do not consider the properties of spoken language. Therefore, this paper aims at generalizing the idea of language model pre-training to lattices generated by recognition systems. We propose a framework that trains neural lattice language models to provide contextualized representations for spoken language understanding tasks. The proposed two-stage pre-training approach reduces the demands of speech data and has better efficiency. Experiments on intent detection and dialogue act recognition datasets demonstrate that our proposed method consistently outperforms strong baselines when evaluated on spoken inputs. The code is available at this https URL
Comments: Published in ACL 2020 as a short paper
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as: arXiv:2007.02629 [cs.CL]
  (or arXiv:2007.02629v2 [cs.CL] for this version)

Submission history

From: Chao-Wei Huang [view email]
[v1] Mon, 6 Jul 2020 10:38:03 GMT (762kb,D)
[v2] Mon, 2 Nov 2020 06:53:56 GMT (762kb,D)

Link back to: arXiv, form interface, contact.