How Many and Which Training Points Would Need to be Removed to Flip this Prediction?

Yang, Jinghan; Jain, Sarthak; Wallace, Byron C.

Full-text links:

Download:

Current browse context:

cs.CL

< prev | next >

new | recent | 2302

Computer Science > Machine Learning

Title: How Many and Which Training Points Would Need to be Removed to Flip this Prediction?

Authors: Jinghan Yang, Sarthak Jain, Byron C. Wallace

(Submitted on 4 Feb 2023 (v1), last revised 9 Feb 2023 (this version, v2))

Abstract: We consider the problem of identifying a minimal subset of training data $\mathcal{S}_t$ such that if the instances comprising $\mathcal{S}_t$ had been removed prior to training, the categorization of a given test point $x_t$ would have been different. Identifying such a set may be of interest for a few reasons. First, the cardinality of $\mathcal{S}_t$ provides a measure of robustness (if $|\mathcal{S}_t|$ is small for $x_t$, we might be less confident in the corresponding prediction), which we show is correlated with but complementary to predicted probabilities. Second, interrogation of $\mathcal{S}_t$ may provide a novel mechanism for contesting a particular model prediction: If one can make the case that the points in $\mathcal{S}_t$ are wrongly labeled or irrelevant, this may argue for overturning the associated prediction. Identifying $\mathcal{S}_t$ via brute-force is intractable. We propose comparatively fast approximation methods to find $\mathcal{S}_t$ based on influence functions, and find that -- for simple convex text classification models -- these approaches can often successfully identify relatively small sets of training examples which, if removed, would flip the prediction.

Comments:	Accepted to EACL 2023
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
Cite as:	arXiv:2302.02169 [cs.LG]
	(or arXiv:2302.02169v2 [cs.LG] for this version)

Submission history

From: Jinghan Yang [view email]
[v1] Sat, 4 Feb 2023 13:55:12 GMT (1234kb,D)
[v2] Thu, 9 Feb 2023 04:23:22 GMT (1235kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2302.02169

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: How Many and Which Training Points Would Need to be Removed to Flip this Prediction?

Submission history