• Medientyp: E-Artikel
  • Titel: Wiki-Based Communities of Interest: Demographics and Outliers
  • Beteiligte: Arnaout, Hiba; Razniewski, Simon; Pan, Jeff Z.
  • Erschienen: Association for the Advancement of Artificial Intelligence (AAAI), 2023
  • Erschienen in: Proceedings of the International AAAI Conference on Web and Social Media
  • Sprache: Nicht zu entscheiden
  • DOI: 10.1609/icwsm.v17i1.22206
  • ISSN: 2334-0770; 2162-3449
  • Entstehung:
  • Anmerkungen:
  • Beschreibung: <jats:p>In this paper, we release data about demographic information and outliers of communities of interest. Identified from Wiki-based sources, mainly Wikidata, the data covers 7.5k communities, e.g., members of the White House Coronavirus Task Force, and 345k subjects, e.g., Deborah Birx. We describe the statistical inference methodology adopted to mine such data. We release subject-centric and group-centric datasets in JSON format, as well as a browsing interface. Finally, we forsee three areas where this dataset can be useful: in social sciences research, it provides a resource for demographic analyses; in web-scale collaborative encyclopedias, it serves as an edit recommender to fill knowledge gaps; and in web search, it offers lists of salient statements about queried subjects for higher user engagement. The dataset can be accessed at: https://doi.org/10.5281/zenodo.7410436</jats:p>