We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.CL

Change to browse by:

cs

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Computation and Language

Title: Retrofitting Structure-aware Transformer Language Model for End Tasks

Abstract: We consider retrofitting structure-aware Transformer-based language model for facilitating end tasks by proposing to exploit syntactic distance to encode both the phrasal constituency and dependency connection into the language model. A middle-layer structural learning strategy is leveraged for structure integration, accomplished with main semantic task training under multi-task learning scheme. Experimental results show that the retrofitted structure-aware Transformer language model achieves improved perplexity, meanwhile inducing accurate syntactic phrases. By performing structure-aware fine-tuning, our model achieves significant improvements for both semantic- and syntactic-dependent tasks.
Comments: Accepted as long paper in EMNLP2020 main proceeding
Subjects: Computation and Language (cs.CL)
Cite as: arXiv:2009.07408 [cs.CL]
  (or arXiv:2009.07408v1 [cs.CL] for this version)

Submission history

From: Hao Fei [view email]
[v1] Wed, 16 Sep 2020 01:07:07 GMT (456kb,D)

Link back to: arXiv, form interface, contact.