文章针对现行内容误差评估指标,利用样本未加权数据构造,从而违背抽样推断理论要求的问题,明确提出用狭义内容误差评估指标体系予以替代的研究目标.为实现这一目标,文章采用样本加权数据构造狭义内容误差指标的估计量及抽样方差估计量,解读本领域已有成果的得与失,创新内容误差研究新思路.运用这一思路,研究了人口普查内容误差的发生机制,内容误差测评指标体系的构建,测评指标的抽样估计等问题.通过上述研究,得出下列结论:在抽样调查条件下必须给出测评指标的抽样估计程序.
本文针对作为当今人口普查质量评估领域主流方法的双系统估计量因受独立性不满足困扰而有偏估计总体人数的问题,明确提出了用三系统估计量替代双系统估计量的研究目标.三系统估计量指的是依据对同一时点上同一总体的人口普查、人口普查之后的质量评估调查和人口行政记录系统这三个不同途径获得的三份人口登记数据来估计总体人数所使用的估计量.为实现上述目标,采用实地调研、数理模型、抽样估计与实证分析相结合的研究方法,以及借鉴国际相关前沿研究成果的研究思路,研究了人口普查质量评估中三系统估计量的构建、三系统估计量应用操作等问题.通过理论与实证研究,得出如下结论:三系统估计量能够有效地摆脱双系统估计量的独立性约束从而解决估计量偏误的问题;将三系统估计量应用于人口普查质量评估,必须科学合理地解决三个人口登记系统定格在统一的标准时点、人口在三个登记系统等概率分层、用人口普查质量评估调查样本对三系统估计量再估计等实践方面的问题;基于Logistic回归模型的三系统估计量可以较好地实现人口等概率分层目标,但尚有较多的理论与实践问题有待研究解决,在这一问题上有着相当大的研究空间.建议我国在未来人口普查质量评估中应用三系统估计量.
文章根据国际微观数据系列整合共享数据库中的13个亚太国家各自多次人口普查微观数据,检测了不同出生队列的小学毕业及以上受教育程度人口所占比例在历次普查之间的一致性.研究发现,中国、越南、蒙古和印度尼西亚4个国家的一致性很高,平均差异不到0.5个百分点,回归系数为0.93~1.07,R2高达0.99.然而,另外一些国家的一致性较差,有的国家与平均值的绝对差异高达16个百分点.这13个亚太国家的回归系数的变化范围为0.62~1.44,R2为0.65~0.99.总体而言,这些国家小学及以上受教育程度人口所占比例在各自历次普查中的统计一致性较高.最后,文章就如何专业化地使用数据库中统一整合过的微观数据提出了一些建议.
The IPUMS-International project, now in its fifteenth year, integrates and disseminates population microdata for twenty-two African countries (82 countries world-wide) and the number continues to increase as more National Statistical Offices cooperate with the initiative. Statistical quality is a serious concern both for the producers of the microdata as well as the researchers who use them. This paper applies the intra-cohort comparison method to pairs of integrated (harmonized) samples for fifteen African countries to assess statistical coherence using as a benchmark the proportion completing primary school by single years of birth. Samples for six countries show near perfect coherence (R2 > .9, and regression coefficients ~1.0 +/- <0.08). For a second group of five countries, coefficients are only slightly larger (R2 > 0.6 <0.9). Large deviations from 1.0 characterize samples for only four countries. On the whole, the results suggest that samples for the fifteen countries have considerable utility for socio-demographic analysis.
The Integrated Public Use Microdata Series (IPUMS)-International partnership is a project of the Minnesota Population Center and national statistical agencies, dedicated to collecting and distributing census data from around the world. IPUMS is currently disseminating data on over a half-billion persons enumerated in more than 250 census samples from 79 countries. The data series includes information on a broad range of population characteristics, including fertility, nuptiality, life-course transitions, migration, labor-force participation, occupational structure, education, ethnicity, and household composition. This paper describes sample characteristics and data structure; the data integration process including the creation of constructed family interrelationship variables; the flexible dissemination system that enables researchers to build customized extracts of pooled census samples across time and place; and some of the most significant findings that have emerged from the database.
This paper analyzes 24 African census samples from 13 countries available via the African Integrated Census MicroData website to illustrate how microdata may be used to assess development and pinpoint basic human needs at local administrative levels over time. We calculate a Human Development Index-like measure for small administrative areas, where much of the responsibility lies for executing policies related to health, education and general well-being. The methodological proposals introduced in this paper are particularly pertinent for the case of Africa. While it is true that data for much of Africa is not appropriate for economic growth rates or per-capita income estimates, the analysis in this paper demonstrates that they are good enough for many other purposes. Indeed, a major aggravating problem that contributes to the "African statistical tragedy" is the lack of accessibility to existing census microdata. This paper aims to illustrate the usefulness of census microdata-which are vastly under-utilized in Africa-and hopefully contribute to make them more transparent and freely accessible.
IPUMS-International disseminates harmonized census microdata for more than 80 countries at no cost, although access is restricted to bona-fide researchers and students who agree to the stringent conditions-of-use license. Currently over 270 samples are available, totaling more than 600 million person records. Each year, 15–20 additional samples are released, as more countries cooperate with the IPUMS initiative and the integration of 2010 round census samples is completed. With so much microdata so readily available, questions of data quality naturally arise. This article focusses on the concept of statistical coherence over time for a single concept, primary schooling completed. From an analysis of the percentage completing primary schooling by birth year for pairs of samples for 13 Asia-Pacific countries, outstanding coherence is found for four countries – China, Mongolia, Vietnam and Indonesia – with mean differences of less than 0.5 percentage points, regression coefficient ( b) ranging from 0.93 to 1.07 and R2 = 0.99. For the 13 countries as a group there is considerable variation overall with mean absolute difference as high as 16 percentage points, b ranging from 0.62–1.44 and R2 = 0.65–0.99. As a whole, statistical coherence of primary schooling is outstanding. Nonetheless, to make expert use of the harmonized microdata, researchers are cautioned to carefully study the IPUMS integrated metadata as well as the original source documentation. National Statistical Offices not currently cooperating or that have not yet entrusted 2010 round census microdata are invited to do so.
IPUMS-International disseminates more than two hundred-fifty integrated, confidentialized census microdata samples to thousands of researchers world-wide at no cost. The number of samples is increasing at the rate of several dozen per year, as quickly as the task of integrating metadata and microdata is completed. Protecting the statistical confidentiality and privacy of individuals represented in the microdata is a sine qua non of the IPUMS project. For the 2010 round of censuses, even greater protections are required, while researchers are demanding ever higher precision and utility. This paper describes a tripartite collaborative experiment using a ten percent household sample of the 2011 census of Ireland to estimate risk, mask the microdata using controlled shuffling, and assess analytical utility by comparing the masked data against the unprotected source microdata. Controlled shuffling exploits hierarchically ordered coding schemes to protect privacy and enhance utility. With controlled shuffling, the lesson seems to be the more detail means less risk and greater utility. Overall, despite substantial perturbation of the masked dataset (30
The explosive expansion of non-marital cohabitation in Latin America since the 1970s has led to the narrowing of the gap in educational homogamy between married and cohabiting couples (what we call “homogamy gap”) as shown by our analysis of 29 census samples encompassing eight countries: Argentina, Brazil, Chile, Colombia, Costa Rica, Ecuador, Mexico, and Panama ( N = 2,295,160 young couples). Most research on the homogamy gap is limited to a single decade and a small group of developed countries (the United States, Canada, and Europe). We take a historical and cross-national perspective and expand the research to a range of developing countries, where since early colonial times, traditional forms of cohabitation among the poor, uneducated sectors of society have coexisted with marriage, although to widely varying degrees from country to country. In recent decades, cohabitation is emerging in all sectors of society. We find that among married couples, educational homogamy continues to be higher than for those who cohabit, but in recent decades, the difference has narrowed substantially in all countries. We argue that assortative mating between cohabiting and married couples tends to be similar when the contexts in which they are formed are also increasingly similar.
IPUMS-International disseminates more than two hundred integrated, confidentialized census microdata samples to thousands of researchers world-wide at no cost. The number of samples is increasing at the rate of several dozen per year, as the process of integrating metadata and microdata is completed. Protecting the statistical confidentiality and privacy of individuals represented in the microdata is a sine qua non of the IPUMS project. For the 2010 round of censuses, even greater protections are required, while researchers are demanding ever higher precision and greater utility. This paper describes a tripartite collaborative experiment using a ten percent household sample of the 2011 census of Ireland to estimate risk, mask the data using controlled shuffling, and assess analytical utility by comparing the masked data against the unprotected source microdata. Controlled shuffling exploits hierarchically ordered coding schemes to protect privacy and enhance utility. With controlled shuffling, the lesson seems to be more detail means less risk and greater utility. Overall, despite substantial perturbations of the masked dataset, we find that data utility is very high and information loss is slight, almost imperceptible even for fairly complex analytical problems.. Acknowledgement. The authors greatly appreciate the cooperation of the Central Statistics Office of Ireland in providing a ten per cent household sample of the 2011 census for this experiment. The authors alone are solely responsible for the contents of this paper. The dataset described here-in is solely for experimentation and, as of this writing, the CSO has not approved its release to third parties.
Seventy years of Inter American Statistical cooperation, symbolized by the 70 anniversary of , made possible the construction of IPUMS-International, the world's largest integrated census microdata dissemination site, www.ipums.org/international. Currently, the site offers access to 238 samples totaling over 540 million person records representing 74 countries. The Americas, which account for only about one-seventh of the world's population, amount to over one-third (36%) of the person records in the IPUMS-International database. Likewise, 35% of the citations in the IPUMS-International bibliography are for studies focused on Latin America, with about half of these analyzing a single Latin American country. This article discusses salient features of the IPUMS integration methods and system. National Statistical Institutes that have not yet entrusted 2010 census microdata to the initiative are invited to do so. Researchers and teachers are invited to use the data freely in analysis and teaching. Setenta años de cooperación estadística inter-Americana, simbolizada por el 70 aniversario de la revista , han hecho posible la construcción de IPUMS-internacional, la base en línea de microdatos censales harmonizados más grande del mundo, www.ipums.org/international. Actualmente, IPUMS proporciona acceso a 238 muestras con más de 540 millones de registros individuales de 74 países. Las Américas, que albergan una séptima parte de la población mundial, representan más de un tercio (36%) de todos los registros individuales en la base de datos IPUMS-internacional. Asimismo, el 35% de todas las referencias en la bibliografía de IPUMS son de estudios realizados sobre América Latina, la mitad de éstas basadas en un sólo país de la región. Este artículo presenta las principales características del sistema de integración y difusión de datos de IPUMS. Los Institutos Nacionales de Estadísticas que todavía no ha entregado la muestra de microdatos de la ronda de 2010 son invitados a hacerlo. Los investigadores y profesores son invitados a utilizar los datos de forma gratuita para sus actividades de investigación y docencia.
Over the past decade a revolution has occurred in the dissemination and analysis of census microdata. This paper discusses the IPUMS-International initiative to liberate census data for researchers world-wide without cost. As of June 2013, academic researchers and policy makers may access, 234 anonymized samples representing 74 countries and totaling over one-half billion person records. The database expands with the addition of 20-30 samples each year. Data are downloadable as extracts from the project website: www.ipums.org/international. To facilitate good use, both metadata and microdata are integrated. The analysis of 450 citations in the project bibliography reveals patterns in publications by country and topic.
Seventy years of Inter American Statistical cooperation, symbolized by the 70th anniversary of Estadística, made possible the construction of IPUMS-International, the world's largest integrated census microdata dissemination site, www.ipums.org/international. Currently, the site offers access to 238 samples totaling over 540 million person records representing 74 countries. The Americas, which account for only about one-seventh of the world's population, amount to over one-third (36%) of the person records in the IPUMS-International database. Likewise, 35% of the citations in the IPUMS-International bibliography are for studies focused on Latin America, with about half of these analyzing a single Latin American country. This article discusses salient features of the IPUMS integration methods and system. National Statistical Institutes that have not yet entrusted 2010 census microdata to the initiative are invited to do so. Researchers and teachers are invited to use the data freely in analysis and teaching. Setenta años de cooperación estadística inter-Americana, simbolizada por el 70 aniversario de la revista Estadística, han hecho posible la construcción de IPUMS-internacional, la base en línea de microdatos censales harmonizados más grande del mundo, www.ipums.org/international. Actualmente, IPUMS proporciona acceso a 238 muestras con más de 540 millones de registros individuales de 74 países. Las Américas, que albergan una séptima parte de la población mundial, representan más de un tercio (36%) de todos los registros individuales en la base de datos IPUMS-internacional. Asimismo, el 35% de todas las referencias en la bibliografía de IPUMS son de estudios realizados sobre América Latina, la mitad de éstas basadas en un sólo país de la región. Este artículo presenta las principales características del sistema de integración y difusión de datos de IPUMS. Los Institutos Nacionales de Estadísticas que todavía no ha entregado la muestra de microdatos de la ronda de 2010 son invitados a hacerlo. Los investigadores y profesores son invitados a utilizar los datos de forma gratuita para sus actividades de investigación y docencia.
Marriage has not been, historically, a major reason for people to migrate across borders. Instead, most people migrate for economic reasons in search of land, a better job, or more opportunity. In recent decades, there has been increased international migration for family reasons to reunite with emigrant kin, to seek refuge from violence, to escape famine and natural disaster or simply to retire to a sunny paradise. Historically, if marriage was the reason to migrate, most unions would have occurred between migrants of the same nativity, strengthening family ties and reinforcing trans-national networks between countries of origin and destination. As we have seen in some migrant communities in Europe and America, often international migrants favor marriage with individuals from their country of origin. This is a common pattern of first and second generation Moroccans and Turks in Western Europe (Cottrell 1973; Cretser 1999; Lievens 1999; Glowsky 2007; Niedomysl et al. 2010), as it was a century ago with Italians, Greeks and many other ethnicities in the United States (McCaa 1993; McCaa et al. 2005)