Semi-Supervised Disentangled Framework for Transferable Named Entity Recognition

Hao, Zhifeng; Lv, Di; Li, Zijian; Cai, Ruichu; Wen, Wen; Xu, Boyan

doi:10.1016/j.neunet.2020.11.017

Full-text links:

Download:

Current browse context:

cs.CL

< prev | next >

new | recent | 2012

Computer Science > Computation and Language

Title: Semi-Supervised Disentangled Framework for Transferable Named Entity Recognition

Authors: Zhifeng Hao, Di Lv, Zijian Li, Ruichu Cai, Wen Wen, Boyan Xu

(Submitted on 22 Dec 2020)

Abstract: Named entity recognition (NER) for identifying proper nouns in unstructured text is one of the most important and fundamental tasks in natural language processing. However, despite the widespread use of NER models, they still require a large-scale labeled data set, which incurs a heavy burden due to manual annotation. Domain adaptation is one of the most promising solutions to this problem, where rich labeled data from the relevant source domain are utilized to strengthen the generalizability of a model based on the target domain. However, the mainstream cross-domain NER models are still affected by the following two challenges (1) Extracting domain-invariant information such as syntactic information for cross-domain transfer. (2) Integrating domain-specific information such as semantic information into the model to improve the performance of NER. In this study, we present a semi-supervised framework for transferable NER, which disentangles the domain-invariant latent variables and domain-specific latent variables. In the proposed framework, the domain-specific information is integrated with the domain-specific latent variables by using a domain predictor. The domain-specific and domain-invariant latent variables are disentangled using three mutual information regularization terms, i.e., maximizing the mutual information between the domain-specific latent variables and the original embedding, maximizing the mutual information between the domain-invariant latent variables and the original embedding, and minimizing the mutual information between the domain-specific and domain-invariant latent variables. Extensive experiments demonstrated that our model can obtain state-of-the-art performance with cross-domain and cross-lingual NER benchmark data sets.

Subjects:	Computation and Language (cs.CL); Machine Learning (cs.LG)
DOI:	10.1016/j.neunet.2020.11.017
Cite as:	arXiv:2012.11805 [cs.CL]
	(or arXiv:2012.11805v1 [cs.CL] for this version)

Submission history

From: Zijian Li [view email]
[v1] Tue, 22 Dec 2020 02:55:04 GMT (1070kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2012.11805

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computation and Language

Title: Semi-Supervised Disentangled Framework for Transferable Named Entity Recognition

Submission history