Evaluating NLP Systems On a Novel Cloze Task: Judging the Plausibility of Possible Fillers in Instructional Texts

Hu, Zizhao; Chanumolu, Ravikiran; Lin, Xingyu; Ayaz, Nayela; Chi, Vincent

Full-text links:

Download:

Current browse context:

cs.CL

< prev | next >

new | recent | 2112

Change to browse by:

Computer Science > Computation and Language

Title: Evaluating NLP Systems On a Novel Cloze Task: Judging the Plausibility of Possible Fillers in Instructional Texts

Authors: Zizhao Hu, Ravikiran Chanumolu, Xingyu Lin, Nayela Ayaz, Vincent Chi

(Submitted on 3 Dec 2021)

Abstract: Cloze task is a widely used task to evaluate an NLP system's language understanding ability. However, most of the existing cloze tasks only require NLP systems to give the relative best prediction for each input data sample, rather than the absolute quality of all possible predictions, in a consistent way across the input domain. Thus a new task is proposed: predicting if a filler word in a cloze task is a good, neutral, or bad candidate. Complicated versions can be extended to predicting more discrete classes or continuous scores. We focus on subtask A in Semeval 2022 task 7, explored some possible architectures to solve this new task, provided a detailed comparison of them, and proposed an ensemble method to improve traditional models in this new task.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2112.01867 [cs.CL]
	(or arXiv:2112.01867v1 [cs.CL] for this version)

Submission history

From: Zizhao Hu [view email]
[v1] Fri, 3 Dec 2021 12:02:52 GMT (209kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2112.01867

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computation and Language

Title: Evaluating NLP Systems On a Novel Cloze Task: Judging the Plausibility of Possible Fillers in Instructional Texts

Submission history