arXiv Open Access 2023

Wiki-based Communities of Interest: Demographics and Outliers

Hiba Arnaout Simon Razniewski Jeff Z. Pan
Lihat Sumber

Abstrak

In this paper, we release data about demographic information and outliers of communities of interest. Identified from Wiki-based sources, mainly Wikidata, the data covers 7.5k communities, such as members of the White House Coronavirus Task Force, and 345k subjects, e.g., Deborah Birx. We describe the statistical inference methodology adopted to mine such data. We release subject-centric and group-centric datasets in JSON format, as well as a browsing interface. Finally, we forsee three areas this research can have an impact on: in social sciences research, it provides a resource for demographic analyses; in web-scale collaborative encyclopedias, it serves as an edit recommender to fill knowledge gaps; and in web search, it offers lists of salient statements about queried subjects for higher user engagement.

Topik & Kata Kunci

Penulis (3)

H

Hiba Arnaout

S

Simon Razniewski

J

Jeff Z. Pan

Format Sitasi

Arnaout, H., Razniewski, S., Pan, J.Z. (2023). Wiki-based Communities of Interest: Demographics and Outliers. https://arxiv.org/abs/2303.09189

Akses Cepat

Lihat di Sumber
Informasi Jurnal
Tahun Terbit
2023
Bahasa
en
Sumber Database
arXiv
Akses
Open Access ✓