Record linkage in the Cape of Good Hope Panel*

A. Rijpma, Jeanne Cilliers, J. Fourie
{"title":"Record linkage in the Cape of Good Hope Panel*","authors":"A. Rijpma, Jeanne Cilliers, J. Fourie","doi":"10.1080/01615440.2018.1517030","DOIUrl":null,"url":null,"abstract":"Abstract In this article, we describe the record linkage procedure to create a panel from Cape Colony census returns, or opgaafrolle, for 1787–1828, a dataset of 42,354 household-level observations. Based on a subset of manually linked records, we first evaluate statistical models and deterministic algorithms to best identify and match households over time. By using household-level characteristics in the linking process and near-annual data, we are able to create high-quality links for 84% of the dataset. We compare basic analyses on the linked panel dataset to the original cross-sectional data, evaluate the feasibility of the strategy when linking to supplementary sources, and discuss the scalability of our approach to the full Cape panel.","PeriodicalId":154465,"journal":{"name":"Historical Methods: A Journal of Quantitative and Interdisciplinary History","volume":"53 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2019-02-14","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"12","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Historical Methods: A Journal of Quantitative and Interdisciplinary History","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1080/01615440.2018.1517030","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 12

Abstract

Abstract In this article, we describe the record linkage procedure to create a panel from Cape Colony census returns, or opgaafrolle, for 1787–1828, a dataset of 42,354 household-level observations. Based on a subset of manually linked records, we first evaluate statistical models and deterministic algorithms to best identify and match households over time. By using household-level characteristics in the linking process and near-annual data, we are able to create high-quality links for 84% of the dataset. We compare basic analyses on the linked panel dataset to the original cross-sectional data, evaluate the feasibility of the strategy when linking to supplementary sources, and discuss the scalability of our approach to the full Cape panel.
记录连接在好望角面板*
在本文中,我们描述了从开普殖民地1787-1828年的人口普查报告(opgaafrolle)中创建面板的记录链接程序,该数据集包含42,354个家庭观测数据。基于手动链接记录的子集,我们首先评估统计模型和确定性算法,以最佳地识别和匹配家庭。通过在链接过程中使用家庭级特征和近年度数据,我们能够为84%的数据集创建高质量的链接。我们将链接面板数据集的基本分析与原始横截面数据进行了比较,在链接到补充资源时评估了该策略的可行性,并讨论了我们的方法对整个Cape面板的可扩展性。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
求助全文
约1分钟内获得全文 求助全文
来源期刊
自引率
0.00%
发文量
0
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
确定
请完成安全验证×
copy
已复制链接
快去分享给好友吧!
我知道了
右上角分享
点击右上角分享
0
联系我们:info@booksci.cn Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。 Copyright © 2023 布克学术 All rights reserved.
京ICP备2023020795号-1
ghs 京公网安备 11010802042870号
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术官方微信