We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:


Current browse context:


Change to browse by:

References & Citations

DBLP - CS Bibliography


(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo ScienceWISE logo

Computer Science > Social and Information Networks

Title: Wiki-based Communities of Interest: Demographics and Outliers

Abstract: In this paper, we release data about demographic information and outliers of communities of interest. Identified from Wiki-based sources, mainly Wikidata, the data covers 7.5k communities, such as members of the White House Coronavirus Task Force, and 345k subjects, e.g., Deborah Birx. We describe the statistical inference methodology adopted to mine such data. We release subject-centric and group-centric datasets in JSON format, as well as a browsing interface. Finally, we forsee three areas this research can have an impact on: in social sciences research, it provides a resource for demographic analyses; in web-scale collaborative encyclopedias, it serves as an edit recommender to fill knowledge gaps; and in web search, it offers lists of salient statements about queried subjects for higher user engagement.
Comments: Accepted to ICWSM 2023. For demo, see this https URL and for dataset see this https URL
Subjects: Social and Information Networks (cs.SI); Computers and Society (cs.CY)
Cite as: arXiv:2303.09189 [cs.SI]
  (or arXiv:2303.09189v2 [cs.SI] for this version)

Submission history

From: Hiba Arnaout [view email]
[v1] Thu, 16 Mar 2023 09:58:11 GMT (512kb,D)
[v2] Fri, 17 Mar 2023 08:34:30 GMT (512kb,D)

Link back to: arXiv, form interface, contact.