Language Modelling as a Multi-Task Problem

Weber, Lucas; Jumelet, Jaap; Bruni, Elia; Hupkes, Dieuwke

Full-text links:

Download:

Current browse context:

cs.CL

< prev | next >

new | recent | 2101

Computer Science > Computation and Language

Title: Language Modelling as a Multi-Task Problem

Authors: Lucas Weber, Jaap Jumelet, Elia Bruni, Dieuwke Hupkes

(Submitted on 27 Jan 2021)

Abstract: In this paper, we propose to study language modelling as a multi-task problem, bringing together three strands of research: multi-task learning, linguistics, and interpretability. Based on hypotheses derived from linguistic theory, we investigate whether language models adhere to learning principles of multi-task learning during training. To showcase the idea, we analyse the generalisation behaviour of language models as they learn the linguistic concept of Negative Polarity Items (NPIs). Our experiments demonstrate that a multi-task setting naturally emerges within the objective of the more general task of language modelling.We argue that this insight is valuable for multi-task learning, linguistics and interpretability research and can lead to exciting new findings in all three domains.

Comments:	Accepted for publication at EACL 2021
Subjects:	Computation and Language (cs.CL); Machine Learning (cs.LG)
Cite as:	arXiv:2101.11287 [cs.CL]
	(or arXiv:2101.11287v1 [cs.CL] for this version)

Submission history

From: Lucas Weber [view email]
[v1] Wed, 27 Jan 2021 09:47:42 GMT (983kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2101.11287

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computation and Language

Title: Language Modelling as a Multi-Task Problem

Submission history