The census place project: A method for geolocating unstructured place names

IF 2.6 1区 历史学 Q1 ECONOMICS Explorations in Economic History Pub Date : 2023-01-01 DOI:10.1016/j.eeh.2022.101477
Enrico Berkes , Ezra Karger , Peter Nencka
{"title":"The census place project: A method for geolocating unstructured place names","authors":"Enrico Berkes ,&nbsp;Ezra Karger ,&nbsp;Peter Nencka","doi":"10.1016/j.eeh.2022.101477","DOIUrl":null,"url":null,"abstract":"<div><p>Researchers use microdata to study the economic development of the United States and the causal effects of historical policies. Much of this research focuses on county- and state-level patterns and policies because comprehensive sub-county data is not consistently available. We describe a new method that geocodes and standardizes the towns and cities of residence for individuals and households in decennial census microdata from 1790–1940. We release public crosswalks linking individuals and households to consistently-defined place names, longitude-latitude pairs, counties, and states. Our method dramatically increases the number of individuals and households assigned to a sub-county location relative to standard publicly available data: we geocode an average of 83% of the individuals and households in 1790–1940 census microdata, compared to 23% in widely-used crosswalks. In years with individual-level microdata (1850–1940), our average match rate is 94% relative to 33% in widely-used crosswalks. To illustrate the value of our crosswalks, we measure place-level population growth across the United States between 1870 and 1940 at a sub-county level, confirming predictions of Zipf’s Law and Gibrat’s Law for large cities but rejecting similar predictions for small towns. We describe how our approach can be used to accurately geocode other historical datasets.</p></div>","PeriodicalId":47413,"journal":{"name":"Explorations in Economic History","volume":null,"pages":null},"PeriodicalIF":2.6000,"publicationDate":"2023-01-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"6","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Explorations in Economic History","FirstCategoryId":"98","ListUrlMain":"https://www.sciencedirect.com/science/article/pii/S0014498322000559","RegionNum":1,"RegionCategory":"历史学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q1","JCRName":"ECONOMICS","Score":null,"Total":0}
引用次数: 6

Abstract

Researchers use microdata to study the economic development of the United States and the causal effects of historical policies. Much of this research focuses on county- and state-level patterns and policies because comprehensive sub-county data is not consistently available. We describe a new method that geocodes and standardizes the towns and cities of residence for individuals and households in decennial census microdata from 1790–1940. We release public crosswalks linking individuals and households to consistently-defined place names, longitude-latitude pairs, counties, and states. Our method dramatically increases the number of individuals and households assigned to a sub-county location relative to standard publicly available data: we geocode an average of 83% of the individuals and households in 1790–1940 census microdata, compared to 23% in widely-used crosswalks. In years with individual-level microdata (1850–1940), our average match rate is 94% relative to 33% in widely-used crosswalks. To illustrate the value of our crosswalks, we measure place-level population growth across the United States between 1870 and 1940 at a sub-county level, confirming predictions of Zipf’s Law and Gibrat’s Law for large cities but rejecting similar predictions for small towns. We describe how our approach can be used to accurately geocode other historical datasets.

查看原文
分享 分享
微信好友 朋友圈 QQ好友 复制链接
本刊更多论文
普查地点项目:一种对非结构化地名进行地理定位的方法
研究人员使用微观数据来研究美国的经济发展和历史政策的因果效应。这方面的研究大部分集中在县和州一级的模式和政策上,因为全面的县级以下的数据并不总是可用的。我们描述了一种新的方法,对1790-1940年十年一次的人口普查微数据中的个人和家庭居住的城镇进行地理编码和标准化。我们发布了公共人行横道,将个人和家庭与一致定义的地名、经纬度对、县和州联系起来。相对于标准的公开数据,我们的方法显著增加了分配到副县位置的个人和家庭的数量:在1790-1940年的人口普查微数据中,我们平均对83%的个人和家庭进行了地理编码,而在广泛使用的人行横道中,这一比例为23%。在个人层面的微观数据(1850-1940)中,我们的平均匹配率为94%,而在广泛使用的人行横道中,平均匹配率为33%。为了说明人行横道的价值,我们测量了1870年至1940年间美国各地次县一级的人口增长,证实了齐夫定律和直布罗特定律对大城市的预测,但拒绝了对小城镇的类似预测。我们描述了如何使用我们的方法来准确地对其他历史数据集进行地理编码。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
求助全文
约1分钟内获得全文 去求助
来源期刊
CiteScore
2.50
自引率
8.70%
发文量
27
期刊介绍: Explorations in Economic History provides broad coverage of the application of economic analysis to historical episodes. The journal has a tradition of innovative applications of theory and quantitative techniques, and it explores all aspects of economic change, all historical periods, all geographical locations, and all political and social systems. The journal includes papers by economists, economic historians, demographers, geographers, and sociologists. Explorations in Economic History is the only journal where you will find "Essays in Exploration." This unique department alerts economic historians to the potential in a new area of research, surveying the recent literature and then identifying the most promising issues to pursue.
期刊最新文献
Wealth and history: A reappraisal Family first: Defining, constructing, and applying historical patent families Reservoirs of power: The political legacy of dam construction in Franco’s Spain (In-kind) Wages and labour relations in the Middle Ages: It’s not (all) about the money Are some piece rates better than others? Cross-sectional variation in piece rates at a US cotton factory
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
现在去查看 取消
×
提示
确定
0
微信
客服QQ
Book学术公众号 扫码关注我们
反馈
×
意见反馈
请填写您的意见或建议
请填写您的手机或邮箱
已复制链接
已复制链接
快去分享给好友吧!
我知道了
×
扫码分享
扫码分享
Book学术官方微信
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术
文献互助 智能选刊 最新文献 互助须知 联系我们:info@booksci.cn
Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。
Copyright © 2023 Book学术 All rights reserved.
ghs 京公网安备 11010802042870号 京ICP备2023020795号-1