Testing the effectiveness of saliency-based explainability in NLP using randomized survey-based experiments

Rahimi, Adel; Jain, Shaurya

Full-text links:

Download:

Current browse context:

cs.CL

< prev | next >

new | recent | 2211

Computer Science > Computation and Language

Title: Testing the effectiveness of saliency-based explainability in NLP using randomized survey-based experiments

Authors: Adel Rahimi, Shaurya Jain

(Submitted on 25 Nov 2022)

Abstract: As the applications of Natural Language Processing (NLP) in sensitive areas like Political Profiling, Review of Essays in Education, etc. proliferate, there is a great need for increasing transparency in NLP models to build trust with stakeholders and identify biases. A lot of work in Explainable AI has aimed to devise explanation methods that give humans insights into the workings and predictions of NLP models. While these methods distill predictions from complex models like Neural Networks into consumable explanations, how humans understand these explanations is still widely unexplored. Innate human tendencies and biases can handicap the understanding of these explanations in humans, and can also lead to them misjudging models and predictions as a result. We designed a randomized survey-based experiment to understand the effectiveness of saliency-based Post-hoc explainability methods in Natural Language Processing. The result of the experiment showed that humans have a tendency to accept explanations with a less critical view.

Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2211.15351 [cs.CL]
	(or arXiv:2211.15351v1 [cs.CL] for this version)

Submission history

From: Adel Rahimi [view email]
[v1] Fri, 25 Nov 2022 08:49:01 GMT (819kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2211.15351

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computation and Language

Title: Testing the effectiveness of saliency-based explainability in NLP using randomized survey-based experiments

Submission history