Fewer Errors, but More Stereotypes? The Effect of Model Size on Gender Bias

Tal, Yarden; Magar, Inbal; Schwartz, Roy

Full-text links:

Download:

Current browse context:

cs.CL

< prev | next >

new | recent | 2206

Change to browse by:

Computer Science > Computation and Language

Title: Fewer Errors, but More Stereotypes? The Effect of Model Size on Gender Bias

Authors: Yarden Tal, Inbal Magar, Roy Schwartz

(Submitted on 20 Jun 2022)

Abstract: The size of pretrained models is increasing, and so is their performance on a variety of NLP tasks. However, as their memorization capacity grows, they might pick up more social biases. In this work, we examine the connection between model size and its gender bias (specifically, occupational gender bias). We measure bias in three masked language model families (RoBERTa, DeBERTa, and T5) in two setups: directly using prompt based method, and using a downstream task (Winogender). We find on the one hand that larger models receive higher bias scores on the former task, but when evaluated on the latter, they make fewer gender errors. To examine these potentially conflicting results, we carefully investigate the behavior of the different models on Winogender. We find that while larger models outperform smaller ones, the probability that their mistakes are caused by gender bias is higher. Moreover, we find that the proportion of stereotypical errors compared to anti-stereotypical ones grows with the model size. Our findings highlight the potential risks that can arise from increasing model size.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2206.09860 [cs.CL]
	(or arXiv:2206.09860v1 [cs.CL] for this version)

Submission history

From: Yarden Tal [view email]
[v1] Mon, 20 Jun 2022 15:52:40 GMT (1179kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2206.09860

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computation and Language

Title: Fewer Errors, but More Stereotypes? The Effect of Model Size on Gender Bias

Submission history