We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.CL

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Computation and Language

Title: Rethinking the Objectives of Extractive Question Answering

Abstract: This work demonstrates that using the objective with independence assumption for modelling the span probability $P(a_s,a_e) = P(a_s)P(a_e)$ of span starting at position $a_s$ and ending at position $a_e$ has adverse effects. Therefore we propose multiple approaches to modelling joint probability $P(a_s,a_e)$ directly. Among those, we propose a compound objective, composed from the joint probability while still keeping the objective with independence assumption as an auxiliary objective. We find that the compound objective is consistently superior or equal to other assumptions in exact match. Additionally, we identified common errors caused by the assumption of independence and manually checked the counterpart predictions, demonstrating the impact of the compound objective on the real examples. Our findings are supported via experiments with three extractive QA models (BIDAF, BERT, ALBERT) over six datasets and our code, individual results and manual analysis are available online.
Comments: camera-ready version accepted to MRQA'21
Subjects: Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as: arXiv:2008.12804 [cs.CL]
  (or arXiv:2008.12804v4 [cs.CL] for this version)

Submission history

From: Martin Fajčík [view email]
[v1] Fri, 28 Aug 2020 18:22:19 GMT (130kb,D)
[v2] Thu, 31 Dec 2020 15:04:42 GMT (144kb,D)
[v3] Mon, 19 Jul 2021 13:24:01 GMT (110kb,D)
[v4] Tue, 12 Oct 2021 07:43:45 GMT (113kb,D)

Link back to: arXiv, form interface, contact.