References & Citations
Computer Science > Computation and Language
Title: Predicting Sentence-Level Factuality of News and Bias of Media Outlets
(Submitted on 27 Jan 2023 (v1), last revised 28 Jun 2023 (this version, v3))
Abstract: Automated news credibility and fact-checking at scale require accurately predicting news factuality and media bias. This paper introduces a large sentence-level dataset, titled "FactNews", composed of 6,191 sentences expertly annotated according to factuality and media bias definitions proposed by AllSides. We use FactNews to assess the overall reliability of news sources, by formulating two text classification problems for predicting sentence-level factuality of news reporting and bias of media outlets. Our experiments demonstrate that biased sentences present a higher number of words compared to factual sentences, besides having a predominance of emotions. Hence, the fine-grained analysis of subjectivity and impartiality of news articles provided promising results for predicting the reliability of media outlets. Finally, due to the severity of fake news and political polarization in Brazil, and the lack of research for Portuguese, both dataset and baseline were proposed for Brazilian Portuguese.
Submission history
From: Francielle Alves Vargas [view email][v1] Fri, 27 Jan 2023 16:56:24 GMT (207kb,D)
[v2] Wed, 26 Apr 2023 17:33:00 GMT (256kb,D)
[v3] Wed, 28 Jun 2023 21:11:39 GMT (259kb,D)
Link back to: arXiv, form interface, contact.