Discussion paper

DP15825 Human Biographical Record (HBR)

We construct a new dataset of more than seven million notable individuals across recorded human history, the Human Biographical Record (HBR). With Wikidata as the backbone, HBR adds further information from various digital sources, including Wikipedia in all 292 languages. Machine learning and text analysis combine the sources and extract information on date and place of birth and death, gender, occupation, education, and family background. This paper discusses HBR's construction and its completeness, coverage, accuracy, and also its strength and weakness relative to prior datasets. HBR is the first part of a larger project, the human record project that we briefly introduce.

£6.00
Citation

Nekoei, A and F Sinn (2021), ‘DP15825 Human Biographical Record (HBR)‘, CEPR Discussion Paper No. 15825. CEPR Press, Paris & London. https://cepr.org/publications/dp15825