Search results

Showing 1 – 50 of 75 results.
Curated

Census of Jail Facilities, 2006 (ICPSR 26602)

Released/updated on: 2010-01-26
Geographic coverage: United States
To reduce respondent burden and improve data quality and timeliness, the Bureau of Justice Statistics (BJS) split the jail census into two parts: The Census of Jail Inmates was conducted with a reference date of June 30, 2005. The following spring it was followed by this enumeration, the Census of Jail Facilities, which collected data as of March 31, 2006. Previous jail enumerations were conducted in 1970 (ICPSR 7641), 1972 (ICPSR 7638), 1978 (ICPSR 7737), 1983 (ICPSR 8203), 1988 (ICPSR 9256), 1993 (ICPSR 6648), and 1999 (ICPSR 3318). The United States Census Bureau collected the data for the Bureau of Justice Statistics. The 2006 Census of Jail Facilities gathered data from all jail detention facilities holding inmates beyond arraignment, a period normally exceeding 72 hours. Jail facilities were operated by cities and counties, by private entities under contract to correctional authorities, and by the Federal Bureau of Prisons (BOP). Excluded from the census were physically separate temporary holding facilities such as drunk tanks and police lockups that do not hold persons after being formally charged in court. Also excluded were state-operated facilities in Connecticut, Delaware, Hawaii, Rhode Island, Vermont, and Alaska, which have combined jail-prison systems. Fifteen independently operated jails in Alaska were included in the Census. The census collected jurisdictional level information on the number of confined inmates; average daily population; number of separate jail facilities; renovation and building plans; court orders and consent decrees; staff by occupational category and race/ethnicity; jail programs; and costs of operation. The census also collected individual jail facility information on the purpose for which the jail held offenders; gender of the inmates authorized to house; functions, such as general adult population confinement, work release, and medical treatment; whether a separate temporary holding area or lockup was operated; rated capacity; number of confined inmates by gender and adult or juvenile status; year of original construction; and whether the facility ever had a major renovation.
Curated

Demographic Characteristics of Washtenaw County, Michigan, in 1860 (ICPSR 8445)

Released/updated on: 1999-04-12
Geographic coverage: United States, Michigan
The names and occupations of inhabitants of Washtenaw County, Michigan, were transcribed from the 1860 Census enumerations to construct this data collection. In addition to demographic characteristics such as age, sex, and race, information is provided for each individual on property value, place of birth, school attendance, literacy, and physical or mental handicaps. The record for each individual includes "dwelling" and "household" codes, allowing users to analyze the data by family or by housing unit.
Curated

Puerto Rico Census Project, 1910 (ICPSR 4343)

Released/updated on: 2006-01-16
Geographic coverage: Puerto Rico, United States, Global
The data comprising the Puerto Rico Census Project, 1910 contain individual and household records drawn from the 1910 Puerto Rican Population Census. The data include variables containing basic demographic information such as age, sex, race, marital status, number of children born and surviving, family size, place of birth, immigration status, county and neighborhood of residence, urban/rural status, and citizenship. The data also describe language proficiency, literacy, school attendance, and disabilities (blind or deaf) of the individuals. Other variables provide data on occupation, industry, ownership of residence, status of mortgage, and farm ownership. There are four classifications of variables belonging to this dataset: original input variables, coded variables, constructed variables, and quality flag variables. The original input variables contain the raw data collected by the enumerators. The coded variables are variables that were recoded by the University of Wisconsin Survey Center (UWSC) as part of the Puerto Rico Census Project. Constructed variables were produced by UWSC to capture additional relevant information. For example, one constructed variable measures literacy by combining separate variables containing data on whether the individual could read and if they could write. Finally, quality flag variables were created by UWSC to indicate whether it could be logically deduced that individual records had been hand edited by the Census Office.
Curated

Puerto Rico Census Project, 1920 (ICPSR 4344)

Released/updated on: 2006-01-16
Geographic coverage: Puerto Rico, United States, Global
The data comprising the Puerto Rico Census Project, 1920 contain individual and household records drawn from the 1920 Puerto Rican Population Census. The data include variables containing basic demographic information such as age, sex, race, marital status, number of children born and surviving, family size, place of birth, immigration status, county and neighborhood of residence, urban/rural status, and citizenship. The data also describe language proficiency, literacy, school attendance, and disabilities (blind or deaf) of the individuals. Other variables provide data on occupation, industry, ownership of residence, status of mortgage, and farm ownership. There are four classifications of variables belonging to this dataset: original input variables, coded variables, constructed variables, and quality flag variables. The original input variables contain the raw data collected by the enumerators. The coded variables are variables that were recoded by the University of Wisconsin Survey Center (UWSC) as part of the Puerto Rico Census Project. Constructed variables were produced by UWSC to capture additional relevant information. For example, one constructed variable measures literacy by combining separate variables containing data on whether the individual could read and if they could write. Finally, quality flag variables were created by UWSC to indicate whether it could be logically deduced that individual records had been hand edited by the Census Office.
Curated
Simple Crosstabs

Census of Jails, 2013 (ICPSR 36128)

Released/updated on: 2018-04-25
Geographic coverage: United States

To reduce respondent burden for the 2013 collection, the Census of Jails was combined with the Deaths in Custody Reporting Program (DCRP). The census provides the sampling frame for the nationwide Survey of Inmates in Local Jails (SILJ) and the Annual Survey of Jails (ASJ). Previous jail enumerations were conducted in 1970 (ICPSR 7641), 1972 (ICPSR 7638), 1978 (ICPSR 7737), 1983 (ICPSR 8203), 1988 (ICPSR 9256), 1993 (ICPSR 6648), 1999 (ICPSR 3318), 2005 (ICPSR 20367), and 2006 (ICPSR 26602). The RTI International collected the data for the Bureau of Justice Statistics in 2013. The United States Census Bureau was the collection agent from 1970-2006.

The 2013 Census of Jails gathered data from all jail detention facilities holding inmates beyond arraignment, a period normally exceeding 72 hours. Jail facilities were operated by cities and counties, by private entities under contract to correctional authorities, and by the Federal Bureau of Prisons (BOP).

Excluded from the census were physically separate temporary holding facilities such as drunk tanks and police lockups that do not hold persons after being formally charged in court. Also excluded were state-operated facilities in Connecticut, Delaware, Hawaii, Rhode Island, Vermont, and Alaska, which have combined jail-prison systems. Fifteen independently operated jails in Alaska were included in the Census.

The 2013 census collected facility-level information on the number of confined and nonconfined inmates, number of inmates participating in weekend programs, number of confined non-U.S. citizens, number of confined inmates by sex and adult or juvenile status, number of juveniles held as adults, conviction and sentencing status, offense type, number of inmates held by race or Hispanic origin, number of inmates held for other jurisdictions or authorities, average daily population, rated capacity, number of admissions and releases, program participation for nonconfined inmates, operating expenditures, and staff by occupational category.

Curated
Partially restricted
Simple Crosstabs

Census of Jails, 2019 (ICPSR 38323)

Released/updated on: 2022-03-30
Geographic coverage: United States

To reduce respondent burden for the 2019 collection, the Census of Jails was combined with the Deaths in Custody Reporting Program (DCRP). The census provides the sampling frame for the nationwide Survey of Inmates in Local Jails (SILJ) and the Annual Survey of Jails (ASJ). Previous jail enumerations were conducted in 1970 (ICPSR 7641), 1972 (ICPSR 7638), 1978 (ICPSR 7737), 1983 (ICPSR 8203), 1988 (ICPSR 9256), 1993 (ICPSR 6648), 1999 (ICPSR 3318), 2005 (ICPSR 20367), 2006 (ICPSR 26602), and 2013 (ICPSR 36128). The RTI International collected the data for the Bureau of Justice Statistics in 2013 and 2019. The United States Census Bureau was the collection agent from 1970-2006.

The 2019 Census of Jails gathered data from all jail detention facilities holding inmates beyond arraignment, a period normally exceeding 72 hours. Jail facilities were operated by cities and counties, by private entities under contract to correctional authorities, and by the Federal Bureau of Prisons (BOP).

Excluded from the census were physically separate temporary holding facilities such as drunk tanks and police lockups that do not hold persons after being formally charged in court. Also excluded were state-operated facilities in Connecticut, Delaware, Hawaii, Rhode Island, Vermont, and Alaska, which have combined jail-prison systems. Fifteen independently operated jails in Alaska were included in the Census.

The 2019 census collected information on the number of confined inmates, number of persons supervised outside jail, number of inmates participating in weekend programs, number of confined non-U.S. citizens, number of inmates by sex and adult or juvenile status, number of juveniles held as adults, number of inmates who were parole or probation violators, number of inmates by conviction status, number of inmates by felony or misdemeanor status, number of inmates held by race or Hispanic origin, number of inmates held for other jurisdictions or authorities, average daily population, rated capacity, admissions and releases, number of staff employed by local jails, facility functions, and number of jails under court orders and consent decrees.

The 2019 census also included a module to collect data on the effects of the opioid epidemic on local jails and jail responses to the epidemic. Items included:

  • Jail practices on opioid use disorder testing, screening, and treatment.
  • Number of local jail admissions screened during June 2019.
  • Number of positive screens.
  • Number of admissions treated for opioid use disorder.
  • Number of jail inmates treated for opioid withdrawal at midyear 2019.
Curated

County-Specific Net Migration by Five-Year Age Groups, Hispanic Origin, Race, and Sex, 1990-2000: [United States] (ICPSR 4171)

Released/updated on: 2005-05-23
Geographic coverage: United States
Time period: 1990-01-01--2000-01-01
This data collection provides net migration estimates by five-year age groups, Hispanic origin, race, and sex for counties of the United States from 1990 to 2000. These estimates were derived from United States census data from 1990 to 2000, and from vital statistics collected by the National Center for Health Statistics (NCHS) for years 1990 through 1999 using the vital statistics (VS) method. The dataset contains the state and county Federal Information Processing Standards (FIPS) codes that uniquely identify counties within a state. Several data categories are presented in the collection. Vital statistics data tabulate births by sex, race, and Hispanic origin for the periods 1990-1994 and 1995-1999, and deaths by sex, race, Hispanic origin, and age groups for the period 1990-2000. The enumerated and adjusted 1990 and 2000 population categories offer population totals by sex, Hispanic origin, age groups, and race. The expected populations in 2000 are available with totals by sex, race, Hispanic origin, and age groups. Net migration estimates and net migration rates for each category also are included.
Curated

Status of Older Persons in Economic Commission for Europe (ECE) Countries, Census Microdata Samples: Turkey, 1990 (ICPSR 3292)

Released/updated on: 2013-09-27
Geographic coverage: Turkey, Global
The main objectives of this data collection effort were to assemble a set of cross-nationally comparable microdata samples based on the 1990 national population and housing censuses in countries of Europe and North America, and to use these samples to study the social and economic conditions of older persons. The samples are designed to allow research on a wide range of issues related to aging, as well as on other social phenomena. Data collected in Turkey examined the type and size of dwelling units, household composition, and employment history. Also gathered was demographic information on household members, including age, sex, ethnic background, marital status, fertility, education, employment status, income, and occupation.
Curated

Census of Population, 1910 [United States]: Oversample of Black-headed Households (ICPSR 9453)

Released/updated on: 2010-09-01
Geographic coverage: North Carolina, United States, Texas, Tennessee, Kentucky, Louisiana, Florida, Virginia, Arkansas, Maryland
Designed to facilitate analysis of the status of Blacks around the turn of the century, this oversample of Black-headed households in the United States was drawn from the 1910 manuscript census schedules. The sample complements the 1/250 Public Use Sample of the 1910 census manuscripts collected by Samuel H. Preston at the University of Pennsylvania: CENSUS OF POPULATION, 1910 [UNITED STATES]: PUBLIC USE SAMPLE (ICPSR 9166). Part 1, Household Records, contains a record for each household selected in the sample and supplies variables describing the location, type, and composition of the households. Part 2, Individual Records, contains a record for each individual residing in the sampled households and includes information on demographic characteristics, occupation, literacy, nativity, ethnicity, and fertility.
Curated

Puerto Rico's Padrones, 1779-1802 (ICPSR 30262)

Released/updated on: 2011-04-07
Geographic coverage: Puerto Rico, Global
Time period: 1779-01-01--1802-01-01
The series consists of 23 annual censuses spanning the years 1779 to 1802, a collection that for its scope and continuity is unique among serial sources of Spanish American colonial history. The padrones were born of a 1776 Royal Order requesting viceroys and executives of Capitanías Generales and Gobernaciones, such as Puerto Rico, to prepare reports on population, broken down by social status, race, and sex. The focus was the civilian population and, therefore, excludes the regular army troops. The series reports the population of Whites, Indians, free Mulattoes, free Blacks, Mulatto slaves and Black slaves for each of 30 partidos in all 23 years (producing a total of 690 observations). Each socio-racial group was subdivided by sex and an ambiguous "age" criterion, which we have interpreted as the difference between dependent (or minor) status and mayoría de edad (adulthood or full age, which in the Spanish American context was 25 years of age). For each group, there are four subdivisions: adult males, adult females, young males, and young females.
Curated

The 1915 Iowa State Census Project (ICPSR 28501)

Released/updated on: 2010-12-14
Geographic coverage: Iowa, United States
The 1915 Iowa State Census is a unique document. It was the first census in the United States to include information on education and income prior to the United States Federal Census of 1940. It contains considerable detail on other aspects of individuals and households, e.g., religion, wealth and years in the United States and Iowa. The Iowa State Census of 1915 was a complete sample of the residents of the state and the returns were written by census takers (assessors) on index cards. These cards were kept in the Iowa State Archives in Des Moines and were microfilmed in 1986 by the Genealogical Society of Salt Lake City. The census cards were sorted by county, although large cities (those having more than 25,000 residents) were grouped separately. Within each county or large city, records were alphabetized by last name and within last name by first name. This data set includes individual-level records for three of the largest Iowa cities (Des Moines, Dubuque, and Davenport; the Sioux City films were unreadable) and for ten counties that did not contain a large city. (Additional details on sample selection are available in the documentation). Variables include name, age, place of residence, earnings, education, birthplace, religion, marital status, race, occupation, military service, among others. Data on familial ties between records are also included.
Curated
Simple Crosstabs

Racial Neighborhood Inequality in the United States, 1980-2010 (ICPSR 36626)

Released/updated on: 2016-11-03
Geographic coverage: United States
Time period: 1980-01-01--2010-01-01
This project examined economic differences in the neighborhoods where whites, blacks, Hispanics, and Asians live in the U.S. Although it is commonly believed that blacks and Hispanics generally live in neighborhoods where poverty rates are higher than they are in the neighborhoods where whites and Asians live, very little research has tracked the change in racial disparities in neighborhood conditions over time. In prior research, this project's investigators found that racial differences in neighborhood economic conditions have diminished in the U.S. Since 1980 the decline in racial neighborhood inequality has been much faster than the decline in racial residential segregation. Because prior research on neighborhoods has focused on change in the residential segregation of different racial and ethnic groups, the trend in racial neighborhood inequality has been largely overlooked, and its causes are unknown. The objective of this project is to account for the decline in racial neighborhood inequality by investigating why it has declined faster in some metropolitan areas than in others.
Curated

Demographic Characteristics of the Population of Detroit, 1850-1880 (ICPSR 31)

Released/updated on: 2008-03-25
Geographic coverage: Detroit, United States, Michigan
Time period: 1850-01-01--1880-01-01
This data collection provides information for native-born Americans, Irish Americans, and German Americans living in Detroit, Michigan, between 1850 and 1880. Demographic variables provide information on age, sex, occupation, marital status, marriage patterns, ethnic background, place of birth, and spouse's and parents' place of birth. Additional information is provided on family size, number of children of adults, number of individuals in the house beyond the immediate family, total number of individuals in the nuclear family, position of individuals within the family, number of children eligible to be in school, activities of school-age children, adult male skill level, literacy level, length of time the family had been in the United States, ownership and value of real estate, constitutional and legal status, and physical condition.
Curated

Nineteenth Century Family History in Michigan: 1850-1880 (ICPSR 32)

Released/updated on: 2008-03-26
Geographic coverage: Detroit, Flint, United States, Lansing, Michigan
This data collection provides information on the characteristics of 1,194 Michigan families in rural places, towns and villages, and the urban areas of Detroit in 1850 and 1880. Data are provided on the geographic location of each household and type of locale, total number of residents in the household, and total number of children of the head of each household. Demographic variables provide information on age, race, place of birth, and occupation of the household head and their spouse, place of birth of father and mother of the household head and of their spouse, sex of the household head and their children, and age of the children. Additional variables provide information on the number of children listed as unemployed, the number of parents or parents-in-law of the household head residing in the household, the number of other related adults aged 14 and older, other related children aged 14 and younger living in the household, the number of servants or employees in the household, and the number of boarders or roomers in the household.
Curated

Census of Turin, Italy, 1705 (ICPSR 3577)

Released/updated on: 2005-12-15
Geographic coverage: Italy, Global
This study is a census enumeration of the city of Turin, Italy, in 1705. The census was ordered by Duke Victor Amadeus II as the city prepared for a siege from part of the French army, as ordered by King Louis XIV. A house-by-house survey of all inhabitants was conducted to assess how many men were able to bear arms and how many people there were to feed. All of the inhabitants, including both citizens and immigrants streaming into the city from the surrounding countryside to escape the invaders, were listed by name, along with their ages, family relationships, birthplaces, occupations, dwelling places, and any weapons kept in the household. Fifty cantonieri (ward-captains) were assigned to survey the 122 isole (city blocks). The information for 15 of the isole are absent. This study also includes a WinZipped archive of AtlasGIS files, which can be used to produce maps of Turin, Italy.
Curated

Southern Farms Study, 1860 (ICPSR 7419)

Released/updated on: 1992-02-16
Geographic coverage: United States
This study presents 1860 data on population and farm production in 5,228 farms located in 405 major cotton-producing counties in the South. The data was compiled from the agriculture, slave, and population schedules of the 1860 United States manuscript Census. For each farm, variables describing farm land, machinery, crops, and livestock are included, as well as production figures for specific crops and types of livestock on the farm. The population variables tabulate the free and slave residents of each farm by sex, race, and age in five- or ten-year categories.
Curated

Mortality in the South, 1850 (ICPSR 7424)

Released/updated on: 1992-02-16
Geographic coverage: North Carolina, United States, Texas, Tennessee, Kentucky, Louisiana, Georgia, South Carolina
This study recorded information on deaths that occurred in 1850 in seven states of the southern United States: Georgia, Kentucky, Louisiana, North Carolina, South Carolina, Tennessee, and Texas. The data were obtained from the manuscript mortality schedules of the 1850 United States Census. Variables identify the state and county in which each death occurred, and provide information on the age, sex, race, legal status (free or slave), place of birth, and occupation of the deceased. The month and cause of death as well as the number of days of illness before death are also documented.
Curated

Census of Population, 1910 [United States]: Public Use Sample (ICPSR 9166)

Released/updated on: 1992-02-16
Geographic coverage: United States
This nationally representative sample of the United States population in 1910 was drawn from manuscript census schedules. The file contains a record for each household selected in the sample, and supplies variables describing the location, type, and composition of the households. Each household record is followed by a record for each individual residing in the household. Information on individuals includes demographic characteristics, occupation, literacy, nativity, ethnicity, and fertility.
Curated
Simple Crosstabs

National ZIP Code Crosswalk, [United States], 1990-2020 (ICPSR 39431)

Released/updated on: 2025-11-10
Geographic coverage: United States
Time period: 1990-01-01--2020-12-31

ZIP Codes are administrative codes generated by the United States Postal Service (USPS) that refer to the geographic area covered by a specific set of mail delivery routes. The U.S. Census Bureau calculates and distributes aggregated social, economic, and demographic information for the population associated with "ZIP Code Tabulation Areas" (ZCTAs), which are roughly analogous to ZIP Codes and serve as identifiers for specific neighborhoods and communities. These aggregated census data, however, are unable to account for changes in ZIP Code boundaries that occur between decennial censuses, leading to measurement error and missing data problems for scholars who attempt to use the aggregated ZCTA data. The purpose of this crosswalk file is to allow researchers to overcome this limitation, enabling them to appropriately link spatial reference information (ZIP Codes) with characteristics of the populations to which they refer.

Most ZIP Codes do not change boundaries in a decade, but a large enough percentage do as to create a problem with missing or mis-specified data. Boundary changes typically involve one or more of the following three processes, although a small number of cases do not conform to these typologies: (1) two or more existing ZIP Codes are combined to create a single surviving ZIP Code, (2) an existing ZIP Code is divided into multiple resulting ZIP Codes, and (3) boundaries between two or more existing ZIP Codes are altered.

Each of these types of changes alters the geographic area that a ZIP Code refers to, and as such, the spatial unit identified by the ZIP Code includes a different population, with a different array of characteristics. By linking the spatial units associated with ZIP Codes as these boundary changes are enacted, the research team can both prevent the loss of observations due to missing data, and more accurately measure social, demographic, and economic characteristics associated with each ZIP Code.

This data set identifies changes in ZIP Code boundaries between 1990 and 2020, and provides numeric codes that cluster the ZIP Codes into the smallest geographic unit, or group of ZIP Codes, that are consistent across a decade: 1990 - 2000, 2000 - 2010, and 2010 - 2020. This "crosswalk" covers the contiguous United States, Alaska, Hawaii, and the District of Columbia. Since much administrative data is available with ZIP Code as the smallest identifiable geography, ZIP Codes are often used to embed observations from administrative data (patients, businesses, survey respondents, etc.) within their social, demographic, and economic contexts. However, ZIP Code boundaries change over time, resulting in measurement error (matching observations to the wrong contextual unit) or missing data (due to an observation reporting a ZIP Code that did not exist at the beginning of the observational period). These data were collected, and the crosswalk created, in an attempt to resolve these data quality issues.

Curated

Dynamics of Population Aging in Economic Commission for Europe (ECE) Countries, Census Microdata Samples: Estonia, 1989 (ICPSR 6780)

Released/updated on: 2013-09-27
Geographic coverage: Global, Estonia
The main objectives of this data collection effort were to assemble a set of cross-nationally comparable microdata samples for Economic Commission for Europe (ECE) countries based on the 1990 national population and housing censuses in countries of Europe and North America, and to use these samples to study the social and economic conditions of older persons. The samples are designed to allow research on a wide range of issues related to aging, as well as on other social phenomena. The Estonia microdata sample contains information on persons aged 50 and over and the persons who reside with them. Variables included in this dataset cover geographic area, type of residency, type of dwelling, and household characteristics, as well as demographic information such as age, sex, marital status, number of children, education, income, and occupation.
Curated

Dynamics of Population Aging in Economic Commission for Europe (ECE) Countries, Census Microdata Samples: Finland, 1990 (ICPSR 6797)

Released/updated on: 2013-09-27
Geographic coverage: Finland, Global
The main objectives of this data collection effort were to assemble a set of cross-nationally comparable microdata samples for Economic Commission for Europe (ECE) countries based on the 1990 national population and housing censuses in countries of Europe and North America, and to use these samples to study the social and economic conditions of older persons. The samples are designed to allow research on a wide range of issues related to aging, as well as on other social phenomena. The Finland microdata sample contains information on persons aged 50 and over and on the persons who reside with them. Variables included in this dataset provide information on geographic area, type of residency, type of dwelling, household characteristics and demographic characteristics such as age, sex, year of birth, household composition, marital status, number of children, education, income, religion, and occupation.
Curated

Dynamics of Population Aging in Economic Commission for Europe (ECE) Countries, Census Microdata Samples: Czech Republic, 1991 (ICPSR 6857)

Released/updated on: 2013-09-27
Geographic coverage: Czech Republic, Global
The main objectives of this data collection effort were to assemble a set of cross-nationally comparable microdata samples for Economic Commission for Europe (ECE) countries based on the 1990 national population and housing censuses in countries of Europe and North America, and to use these samples to study the social and economic conditions of older persons. The samples are designed to allow research on a wide range of issues related to aging, as well as on other social phenomena. Included in the Czech Republic dataset are questions on the type and characteristics of buildings/dwellings, available utility systems, and demographic information such as age, sex, marital status, number of children, education, income, religion, and occupation. Also included are questions concerning the presence of household amenities such as telephones, toilets, automobiles, baths/showers, washers, and television sets.
Curated

Dynamics of Population Aging in Economic Commission for Europe (ECE) Countries, Census Microdata Samples: Romania, 1992 (ICPSR 6900)

Released/updated on: 2013-09-27
Geographic coverage: Romania, Global
The main objectives of this data collection effort were to assemble a set of cross-nationally comparable microdata samples for Economic Commission for Europe (ECE) countries based on the 1990 national population and housing censuses in countries of Europe and North America, and to use these samples to study the social and economic conditions of older persons. The samples are designed to allow research on a wide range of issues related to aging, as well as on other social phenomena. Included in the Romania data collection are questions on type of dwelling unit and the presence of amenities, such as telephones, toilets, automobiles, baths/showers, washers, and TV sets, as well as the availability of utility systems. Also covered are the characteristics of the buildings within which these dwelling units were located. Demographic and socioeconomic information on household members includes age, sex, year of birth, household composition, marital status, number of children, education, income, religion, and occupation.
Curated

Status of Older Persons in Economic Commission for Europe (ECE) Countries, Census Microdata Samples: Lithuania, 1989 (ICPSR 3952)

Released/updated on: 2013-09-27
Geographic coverage: Lithuania, Global
The main objectives of this data collection effort were to assemble a set of cross-nationally comparable microdata samples based on the 1990 national population and housing censuses in countries of Europe and North America, and to use these samples to study the social and economic conditions of older persons. The samples are designed to allow research on a wide range of issues related to aging, as well as on other social phenomena. Data collected in the 1989 Lithuanian census examined the type and size of dwelling units, household composition, and employment history. Also gathered was demographic information on household members, including age, sex, ethnic background, marital status, fertility, education, employment status, income, and occupation.
Curated

Status of Older Persons in Economic Commission for Europe (ECE) Countries, Census Microdata Samples: Latvia, 1989 (ICPSR 2572)

Released/updated on: 2013-09-27
Geographic coverage: Latvia, Global
The main objectives of this data collection effort were to assemble a set of cross-nationally comparable microdata samples based on the 1990 national population and housing censuses in countries of Europe and North America, and to use these samples to study the social and economic conditions of older persons. The samples are designed to allow research on a wide range of issues related to aging, as well as on other social phenomena. The Latvian dataset examined the type and size of dwelling units, amenities such as flush toilets, baths/showers, and kitchens, and the type of utility systems that were available. Also covered were the characteristics of the buildings within which these dwelling units were located. Demographic and socioeconomic information on household members includes age, sex, ethnic background, household size and composition, marital status, disabilities, fertility, mortality, education, religion, employment status, income, and occupation.
Curated
Simple Crosstabs

Philadelphia Social History Project: Grid Data, 1850, 1860, 1870, 1880 (ICPSR 34982)

Released/updated on: 2014-07-30
Geographic coverage: United States, Philadelphia, Pennsylvania
This component of the Philadelphia Social History Project examines the demographic composition of city grid squares using census data from years 1850, 1860, 1870, and 1880. The collection consists of two types of data files: (1) grid tallies, and (2) grid dictionaries. The grid tally files consist of counts of individuals living in PSHP grid squares, with totals broken down by race/ethnicity, sex, and age. The grid dictionary files link lines in the census manuscripts to PSHP grid squares, allowing users to follow the movements of census-takers as they moved house-to-house on foot, adding individuals to the printed census manuscript forms. The "grid" network consists of a set of vertical and horizontal lines drawn at fixed intervals across a city map, forming the foundation for the spatial organization of the data. The grid dictionary files show when census-takers crossed from one grid square to another; each row in the grid dictionary describes a set of rows that are in a specific grid square by listing the starting page/line and the ending page/line.
Curated

Social, Demographic, and Educational Data for France, 1801-1897 (ICPSR 48)

Released/updated on: 1992-02-16
Geographic coverage: Europe, France
Time period: 1801-01-01--1897-01-01
This data collection consists of 161 selected social, demographic, and educational datasets for France in the period 1801-1897. The data were collected from published reports of three national statistical series: (1) National Censuses, (2) Vital Statistics, and (3) Primary Education. This project was supported by grants from the National Endowment for the Humanities and the National Science Foundation. The National Census data were derived from the quinquennial population censuses of France from 1801 to 1896 and were obtained from the Statistique Generale de la France. The data provide detailed social and economic information for the period 1851 to 1896. The data for 1801-1851 are less rich in subject matter coverage but do present some basic information on population characteristics. The National Census data in general describe the population, including the composition of the population by categories of age, sex, place of birth, marital status, religion, place of residence, and occupation. There is also some limited information on migration, transportation and communication, housing, and families. A large segment of the census data pertains to occupations of the population, specifying job classifications within professions, as well as information on non-employed household members that were dependent on employees in the various industries, in addition to enumerations of persons employed in various professions and trades. The Vital Statistics data files contain annual vital statistics for the French population. These data were obtained from two printed series, MOUVEMENT DE LA POPULATION (1801-1868), and STATISTIQUE ANNUELLE (1869-1897). The basic variables included in the vital statistics datasets record births, deaths, and marriages in France. Detailed cross-tabulations of these demographic indicators are presented for births, tabulated by sex, month, legitimacy status, and characteristics of the parents, and deaths, categorized by age and previous marital status of the partners. Additional cross-tabulations are provided for variables such as divorces, passports issued, medical personnel and hospitals, and a literacy indicator (signing of marriage certificates). The Primary Education data files provide information on primary schools and were obtained from the Statistique de l'enseignement Primaire. The data obtained from the series basically cover the period 1829-1897, although some recapitulative information for earlier years is also presented. The main focus of the data in this series is on primary schools, classes and buildings, enrollment, teachers, sources of funding and expenditure, and academic proficiency of the pupils. Additional information is included on literacy, teacher training (normal) schools, school age population, and libraries. A machine-readable French language codebook, describing the data items as well as the sources from which they were obtained, is provided with each dataset supplied. In addition, lists of the variables included in each dataset are included in Parts 162-164. See the related collection, DEMOGRAPHIC, SOCIAL, EDUCATIONAL AND ECONOMIC DATA FOR FRANCE, 1833-1925 (ICPSR 7529).
Curated

Big Data for Population Research (ICPSR 35978)

Released/updated on: 2015-06-17
Geographic coverage: United States
This project expands the Integrated Public Use Microdata Series (IPUMS) by adding demographic and geographic data describing the entire enumerated population of the U.S. from 1790-1930. The project provides data on the characteristics of over 600 million persons, quadrupling the quantity of U.S. census microdata available for scientific research. The data cover entire populations with full geographic detail, providing contextual information on neighborhood characteristics, including ethnic composition, demographic behavior, and population mobility.
Curated

Dominican Republic Labor Market Survey: 1980 National and 1983 Urban Sample (ICPSR 35351)

Released/updated on: 2014-10-10
Geographic coverage: Dominican Republic
The Dominican Republic Labor Market Survey includes information on housing characteristics and person-records for a 1980 national and 1983 urban sample. The original in-person survey was taken on household enumeration forms, with a sample included in the documentation. Data files contain 13,283 person-records in the order of the original questionnaire (provided in documentation). A household unique number can be matched to the person-records, allowing a variety of analyses. For more information on the data or for any questions about the data, please contact Pamela Paxton at ppaxton@prc.utexas.edu.
Curated

United States National Health Measurement Study, 2005-2006 (ICPSR 23263)

Released/updated on: 2009-06-23
Geographic coverage: United States
Time period: 2005-06-01--2006-08-01
The National Health Measurement Study (NHMS) surveyed older United States adults with a suite of health-related quality of life (HRQoL) indices to allow comparison and cross-calibration of these instruments. The design oversampled African Americans and older individuals to allow subgroup analyses. Several preference-weighted indices measuring self-reported generic HRQoL are used widely in population surveys and clinical studies in the United States and around the world. These indices are used to evaluate individual and population health. Because they have been developed using econometric methods to elicit utility weights for their scoring systems, they are generally accepted for use in cost-effectiveness analyses of health interventions. Each index uses a multidimensional representation of health, but each index covers the dimensions of health (e.g., physical function, mental function, social function, pain, other symptoms, etc.) differently, and uses questionnaires with different psychometric properties. Each index is scored so that perfect health is represented as 1.0 and dead is represented as 0.0, but they are known to have different scaling properties. Rarely have two or more of these instruments been included in a population survey, so there have been few opportunities to directly compare how they describe and measure health using multi-instrument data. In this study, respondents indicated whether they had been diagnosed with coronary heart disease, stroke, diabetes, arthritis, eye disease, sleep disorder, chronic respiratory disease, clinical depression or anxiety disorder, gastrointestinal ulcer, thyroid disorder, and/or severe chronic back pain. Census tract is not identified, however race composition, education levels, economic factors, and urbanicity of each respondent's census tract of residence are included as contextual variables. Demographic, socioeconomic, and additional health data were elicited. Respondents are characterized by census region of residence, age, gender, marital status, race, ethnicity, education, household income and assets, health insurance, weight, height, smoking status, psychological well-being scales, and everyday and lifetime discrimination items. The data were de-identified, and extensive documentation was developed. The NHMS collected data on 3,844 adults in the continental United States (1,641 males and 2,203 females, 1,086 African Americans).
Curated

Integrated Public Health Surveys, 2010-2011 (ICPSR 33822)

Released/updated on: 2024-02-14
Geographic coverage: United States
Time period: 2010-01-01--2011-01-01

This collection comprises a single data file which was produced as part of the data harmonization efforts of the Robert Wood Johnson Foundation and the United States Centers for Disease Control and Prevention. The file contains merged data from five sources:

  1. 2010 National Profile of Local Health Departments, a survey of local health departments conducted by the National Association of County and City Health Officials (NACCHO).

  2. 2011 National Profile Survey of Local Boards of Health, a survey of local boards of health conducted by the National Association of Local Boards of Health (NALBOH).

  3. 2010 State and Territorial Public Health Survey, a survey of state and United States territory health departments conducted by the Association of State and Territorial Health Officials (ASTHO).

  4. 2011 County Health Rankings, a compilation of county-level health measures and within-state county health rankings produced by the University of Wisconsin Population Health Institute.

  5. 2010 Census Demographic Profile Summary File, a series of tables with housing and population data from the 2010 Census.

Produced by matching data from the last four sources to the NACCHO data, the data file contains one case for each of the 2,107 local health departments (LHD) that responded to the NACCHO survey. Each LHD's record in the file includes the ASTHO data for its state health department and the NALBOH data for its local board of health (LBH), if it had a LBH and the LBH responded to the NALBOH survey. (If a LHD had multiple LBHs, then the first one in the NALBOH data was matched to the LHD). In addition, county (or county equivalent)-level data from the County Health Rankings and Census Demographic Profile Summary File were matched to the records of the 1,535 LHDs represented in the data file with a jurisdiction covering a single county or county equivalent.

Curated

Norwegian Ecological Data, 1949-1961 (ICPSR 40)

Released/updated on: 1992-02-16
Geographic coverage: Norway, Europe
Time period: 1949-01-01--1961-01-01
This study contains election and census data for 732 Norwegian communes in the period 1949-1961. Election returns are available for the elections of 1949, 1953, 1957, and 1961. In addition, data from the censuses of 1950 and 1960 are presented, including information on demography, education, modernization, the economy, and occupational structure, and contextual information about clusters of neighboring communes. Data are provided on the total number of registered voters and the total number of votes cast for the Norwegian Communist Party, the Norwegian Labour Party, the Liberal Party (Venstre), the Christian People's Party, the Agrarian Party (the Centre Party), the Conservative Party (Hoyre), and other political parties. Additional variables provide information on age and educational levels for males and females, the total number of economically active population employed in agriculture, forestry, fisheries, manufacturing, and construction, the total value of industrial production, and the total number of private households and occupied housing units.
Curated

Norwegian Ecological Data, 1868-1903 (ICPSR 41)

Released/updated on: 2006-01-12
Geographic coverage: Norway, Europe
Time period: 1868-01-01--1903-01-01
This data collection provides economic, social, political, and demographic information on 431 communes (or electoral parishes) of Norway in the period 1868-1903. There are four parts to this collection. Part 1 contains information from the censuses of 1875, 1891, and 1900 and the electoral censuses of 1868 and 1876 on occupation, income distribution, taxation, age, household, total population by sex, place of birth, and religious affiliation, and information about political participation, such as the number of eligible voters, registered votes, and votes cast in the Storting (unicameral parliament) elections of 1868, 1870, 1873, 1876, 1879, 1882, 1885, 1888, 1891, 1894, 1897, 1900 and 1903. Part 2 provides information from the educational censuses of 1875 and 1885 on school enrollment, the number of male and female teachers, and school expenditures. Part 3 provides information on births, miscarriages, deaths, the number of live births from unwed mothers, the number of married couples, and the number of persons emigrating overseas and to the United States in 1868, 1875, 1891 to 1895, 1896 to 1900, and 1901 to 1905. Part 4 provides information on inter-communal communication and transportation, such as railways and steamships.
Curated

Urban Growth in America: Philadelphia, 1774-1930 (ICPSR 56)

Released/updated on: 2008-03-25
Geographic coverage: United States, Philadelphia, Pennsylvania
Time period: 1774-01-01--1930-01-01
This study contains aggregate economic, political, and social data for the city of Philadelphia in the period 1774-1930. Data are provided for occupational categories in 1774 and 1860 (Parts 1 and 3), the place of birth of the city inhabitants in 1860 (File 2), and for workers aged 10 and over in 1930, tabulated by ward and industry group (Part 4).
Curated

Social Structure of Argentina: Census Data on Economic Development, 1965 (ICPSR 57)

Released/updated on: 1992-02-16
Geographic coverage: South America, Argentina, Global
This study contains data on the social structure of Argentina in 1965. Principal variables in the study cover the active population and its occupational segments, extent of commerce, industry, and rural development, production per capita, density of population, illiteracy, family size, and agricultural production. Derived measures include indices of rural occupational stability, dependency within the urban middle class, and rural landowners.
Curated

Agricultural and Demographic Records for Rural Households in the North, 1860: [Instructional Materials] (ICPSR 3463)

Released/updated on: 2002-10-17
Geographic coverage: Vermont, Indiana, United States, Minnesota, Kansas, New York (state), New Jersey, Michigan, Pennsylvania, Iowa, Illinois, Connecticut, Missouri, New Hampshire, Ohio, Maryland, Wisconsin
These instructional materials were prepared for use with AGRICULTURAL AND DEMOGRAPHIC RECORDS FOR HOUSEHOLDS IN THE NORTH, 1860 (ICPSR 7420), compiled by Fred Bateman and James D. Foust. The data file and accompanying documentation are provided to assist educators in (an SPSS portable file) instructing students about the history of agriculture and rural life in the North, just prior to the Civil War. An instructor's handout has also been included. This handout contains the following sections, among others: (1) General goals for student analysis of quantitative datasets, (2) Specific goals in studying this dataset, (3) Suggested appropriate courses for use of the dataset, (4) Tips for using the dataset, and (5) Related secondary source readings. Demographic, occupational, and economic information for over 21,000 rural households in the northern United States in 1860 are presented in the dataset. The data were obtained from the manuscript agricultural and population schedules of the 1860 United States Census and are provided for all households in a single township from each of the 102 randomly-selected counties in 16 northern states. Variables in the dataset include farm values, livestock, and crop production figures for the households that owned or operated farms (over half the households sampled), as well as value of real and personal estate, color, sex, age, literacy, school attendance, occupation, place of birth, and parents' nationality of all individuals residing in the sampled townships.
Curated

National Profile of Local Health Departments, 2010 (ICPSR 32922)

Released/updated on: 2024-02-14
Geographic coverage: United States
Conducted by the National Association of County and City Health Officials (NACCHO), the purpose of this survey of local health departments (LHDs) was to advance and support the development of a database for LHDs to describe and understand their structure, function, and capacities. A core set of questions was submitted to every LHD. In addition, some LHDs received one of two randomly assigned modules of supplemental questions. The core questions covered governance, funding, workforce (staffing levels, occupations employed, top executive education and licensure, and percentages of staff by race and Hispanic origin), LHD activities, and community health assessment and health improvement planning. The surveyed LHD activities include immunization, screening for diseases and conditions, treatment for communicable diseases, maternal and child health, epidemiology and surveillance activities, population-based primary prevention activities, and regulation, inspection and/or licensing activities. Topics covered by Module 1 included quality improvement, familiarity with a voluntary national accreditation program for state and local health departments, sharing of resources with other LHDs, emergency preparedness, and information technology. Module 2 examined human resources, policy-making and advocacy, access to health care services, practice-based research, health impact assessments, public health and law, and use of public health reports.
Curated

County-Specific Net Migration Estimates, 1980-1990 [United States] (ICPSR 26761)

Released/updated on: 2010-04-02
Geographic coverage: United States
Time period: 1980-01-01--1990-01-01

This data collection represents a set of United States county net migration estimates by age and sex for the 1980-1990 decade, and is part of a series of estimates done for each decade since 1950 (1950-1970: see NET MIGRATION OF THE POPULATION BY AGE, SEX, AND RACE, 1950-1970 [ICPSR 8493]; 1970-1980: see NET MIGRATION OF THE POPULATION OF THE UNITED STATES BY AGE, RACE, AND SEX, 1970-1980 [ICPSR 8697]; 1990-2000: see COUNTY-SPECIFIC NET MIGRATION BY FIVE-YEAR AGE GROUPS, HISPANIC ORIGIN, RACE, AND SEX, 1990-2000 [ICPSR 4171]).

Net migration, the difference between the number of people moving into an area and the number moving out over a period, is measured here, and in all the other sets of estimates in the series, by the residual method. That is, net migration is equal to the population change over the period minus the natural increase (births -- deaths). Full details on how natural increase is estimated for each county, as well as other details of the data collection, are described in the codebook.

Curated
Partially restricted
Simple Crosstabs

Historical Demographic Data of Southeastern Europe: Orasac, 1824-1975 (ICPSR 32404)

Released/updated on: 2013-05-29
Geographic coverage: Orasac, Europe, Serbia, Global
Time period: 1824-01-01--1975-01-01

The data in the Historical Demographic Data of Southeastern Europe series derive primarily from the ethnographic and archival research of Joel M. Halpern, Professor Emeritus of Anthropology at the University of Massachusetts at Amherst, in southeastern Europe from 1953 to 2006. The series is comprised of historical demographic data from several towns and villages in the countries of Bosnia, Croatia, Macedonia, Montenegro, Serbia, and Slovenia, all of which are former constituent republics of the Socialist Federal Republic of Yugoslavia. The data provide insight into the shift from agricultural to industrial production, as well as the more general processes of urbanization occurring in the last days of the Yugoslav state. With an expansive timeframe ranging from 1818 to 2006, the series also contains a wide cross-section of demographic data types. These include, but are not limited to, population censuses, tax records, agricultural and landholding data, birth records, death records, marriage and engagement records, and migration information.

This component of the series focuses exclusively on the Serbian village of Orasac and is composed of 64 datasets. These data record a variety of demographic and economic information between the years of 1824 and 1975. General population information at the individual level is available in official census records from 1863, 1884, 1948, 1953, and 1961, and from population register records for the years of 1928, 1966, and 1975. Census data at the household level is also available for the years of 1863, 1928, 1948, 1953, and 1961. These data are followed by detailed records of engagement and marriage. Many of these data were obtained through the courtesy of village and county officials. Priest book records from 1851 through 1966, as well as death records from 1863 to 1976 and tombstone records from 1975, are also available. Information regarding migrants and emigrants was obtained from the village council for the years of 1946 through 1975. Lastly, the data provide economic and financial information, including records of individual landholdings (for the years of 1863, 1952, 1966, and 1975), records of government taxation at the individual or household level (for 1813 through 1840, as well as for 1952), and livestock censuses (at both the individual and household level for the years of 1824 and 1825, and only at the individual level for the years of 1833 and 1834).

Curated

Population and Income Estimates for the United States, 1969-1973 (ICPSR 78)

Released/updated on: 1992-02-16
Geographic coverage: United States
Time period: 1969-01-01--1973-01-01
This data collection provides information on income and population estimates for the United States in the period 1969-1973. Variables include the total population in 1970, estimated population in 1973, per capita income for 1969, and estimated total money income for 1973. Data are recorded for each of the 38,529 governments (counties, townships, minor civil divisions, etc.) eligible for participation in the Federal Revenue Sharing Program. These data were prepared as part of the Bureau of the Census's Federal-State Cooperative Program for Local Population Estimates.
Curated
Partially restricted
Simple Crosstabs

China Multi-Generational Panel Dataset, Liaoning (CMGPD-LN), 1749-1909 (ICPSR 27063)

Released/updated on: 2016-09-06
Geographic coverage: Asia, China (Peoples Republic)
Time period: 1749-01-01--1909-01-01
The China Multi-Generational Panel Dataset - Liaoning (CMGPD-LN) is drawn from the population registers compiled by the Imperial Household Agency (neiwufu) in Shengjing, currently the northeast Chinese province of Liaoning, between 1749 and 1909. It provides 1.5 million triennial observations of more than 260,000 residents from 698 communities. The population mainly consists of immigrants from North China who settled in rural Liaoning during the early eighteenth century, and their descendants. The data provide socioeconomic, demographic, and other characteristics for individuals, households, and communities, and record demographic outcomes such as marriage, fertility, and mortality. The data also record specific disabilities for a subset of adult males. Additionally, the collection includes monthly and annual grain price data, custom records for the city of Yingkou, as well as information regarding natural disasters, such as floods, droughts, and earthquakes. This dataset is unique among publicly available population databases because of its time span, volume, detail, and completeness of recording, and because it provides longitudinal data not just on individuals, but on their households, descent groups, and communities.
Curated
Simple Crosstabs

European-origin and Mexican-origin Populations in Texas, 1850, 1860, 1870, 1880, 1900, 1910 (ICPSR 35032)

Released/updated on: 2016-06-20
Geographic coverage: United States, Texas
This dataset was produced in the 1990s by Myron Gutmann and others at the University of Texas to assess demographic change in European- and Mexican-origin populations in Texas from the mid-nineteenth to early-twentieth centuries. Most of the data come from manuscript records for six rural Texas counties - Angelina, DeWitt, Gillespie, Jack, Red River, and Webb - for the U.S. Censuses of 1850-1880 and 1900-1910, and tax records where available. Together, the populations of these counties reflect the cultural, ethnic, economic, and ecological diversity of rural Texas. Red River and Angelina Counties, in Eastern Texas, had largely native-born white and black populations and cotton economies. DeWitt County in Southeast Texas had the most diverse population, including European and Mexican immigrants as well as native-born white and black Americans, and its economy was divided between cotton and cattle. The population of Webb County, on the Mexican border, was almost entirely of Mexican origin, and economic activities included transportation services as well as cattle ranching. Gillespie County in Central Texas had a mostly European immigrant population and an economy devoted to cropping and livestock. Jack County in North-Central Texas was sparsely populated, mainly by native-born white cattle ranchers. These counties were selected to over-represent the European and Mexican immigrant populations. Slave schedules were not included, so there are no African Americans in the samples for 1850 or 1860. In some years and counties, the Census records were sub-sampled, using a letter-based sample with the family as the primary sampling unit (families were chosen if the surname of the head began with one of the sample letters for the county). In other counties and years, complete populations were transcribed from the Census microfilms. For details and sample sizes by county, see the County table in the Original P.I. Documentation section of the ICPSR Codebook, or see Gutmann, Myron P. and Kenneth H. Fliess, How to Study Southern Demography in the Nineteenth Century: Early Lessons of the Texas Demography Project (Austin: Texas Population Research Center Papers, no. 11.11, 1989).
Curated

Los Angeles Family and Neighborhood Survey (L.A.FANS), Los Angeles Neighborhood Services and Characteristics Data (L.A.NSC), 1980-2010 (ICPSR 37277)

Released/updated on: 2019-07-22
Geographic coverage: United States, Los Angeles, California
Time period: 1980-01-01--2010-01-01

This study, known as the Los Angeles Neighborhood Services and Characteristics data (L.A.NSC) includes data from governmental and business data sources on the characteristics of census tracts in Los Angeles County from 1980 to 2010. The unit of analysis is 1990 census tract (census tracts using 1990 census tract boundaries). The data also include crosswalks for census boundaries across census years. It is public use data.

The L.A.NSC data were assembled by the L.A.FANS projects to be used with L.A.FANS survey data restricted data versions 2, 2.5, and 3 all of which contain 1990 census tract numbers which can be linked with data in this file. However, the L.A.NSC files includes ALL census tracts in Los Angeles County, not just the 65 tracts sampled by L.A.FANS. Thus, the L.A.NSC data can be analyzed on their own without linkage to L.A.FANS survey data and provide longitudinal data on the characteristics of Los Angeles census tracts over a 30 year period.

Additional information on the data is available at the RAND website. See also:

  • The L.A.NSC Database Users' Guides (The Los Angeles Neighborhood Services and Characteristics Database: Codebook and 2010 Neighborhood Services and Characteristics Database: NSC2).
Curated
Partially restricted
Simple Crosstabs

China Multi-Generational Panel Dataset, Shuangcheng (CMGPD-SC), 1866-1913 (ICPSR 35292)

Released/updated on: 2021-10-14
Geographic coverage: Asia, China (Peoples Republic)
Time period: 1866-01-01--1913-01-01
The China Multi-Generational Panel Dataset - Shuangcheng (CMGPD-SC) provides longitudinal individual, household, and community information on the demographic and socioeconomic characteristics of a resettled population living in Shuangcheng, a county in present-day Heilongjiang Province of Northeastern China, for the period from 1866 to 1913. The dataset includes some 1.3 million annual observations of over 100,000 unique individuals descended from families who were relocated to Shuangcheng in the early 19th century. These families were divided into 3 categories based on their place of origin: metropolitan bannermen, rural bannermen, and floating bannermen. The CMGPD-SC, like its Liaoning counterpart, the CMGPD-LN (ICPSR 27063), is a valuable data source for studying longitudinal as well as multi-generational social and demographic processes. The population categories had salient differences in social origins and land entitlements, and landholding data are available at a number of time periods, thus the CMGPD-SC is especially suitable to the study of stratification processes.
Curated
Simple Crosstabs

Contraceptive Needs and Services in the United States, 1994-2016 (ICPSR 38891)

Released/updated on: 2024-01-23
Geographic coverage: United States
Time period: 1994-01-01--2016-12-31

These data come from surveillance activities conducted by the Guttmacher Institute over several decades, collecting or compiling data for the period 1994 through 2016. These activities track the numbers of women who have a potential demand for contraceptive care (because they are of reproductive age, sexually active and not seeking to become pregnant), the subset of these women who likely need public support for care (because of their family income level or their age), the numbers of women who receive contraceptive services from publicly funded clinics, and the numbers of clinics providing publicly supported contraceptive services. These efforts have been conducted periodically, typically about every five years, but at times the intervals between efforts were shorter or longer than five years. The most recent data were collected or compiled for 2015 (women served) and 2016 (women with potential demand for services).

This release includes two separate datasets. Dataset 1, "Need for contraceptive services," provides county-level aggregate data for 6 different years (1995, 2000, 2002, 2006, 2010, and 2016). For each county, the data represent estimates of the number of women who have a potential demand for contraceptive services and the number who likely need public support for care, both in total, and according to key socio-demographic characteristics. Dataset 2, "Clinics providing contraceptive services and women served," provides county-level aggregate data for six different years (1994, 1997, 2001, 2006, 2010, and 2015). For each county, the data represent the number of publicly funded clinics according to clinic type and funding status and the number of female contraceptive patients served at those clinics.

Curated
Partially restricted

Pathways to Adulthood: A Three-Generation Urban Study, 1960-1994: [Baltimore, Maryland] (ICPSR 2420)

Released/updated on: 2019-11-26
Geographic coverage: Baltimore, United States, Maryland
Time period: 1960-01-01--1994-01-01
This collection incorporates both prospective and retrospective data on three generations of families initially living in inner-city Baltimore, Maryland. The prospective data were selected from data collected as part of the Johns Hopkins Collaborative Perinatal Study (JHCPS), a survey of pregnant women seeking prenatal care and delivery at Johns Hopkins Hospital during 1960-1964. JHCPS studied these women (the first-generation mothers, abbreviated as G1) and the children born to them during 1960-1965 (the second-generation children, abbreviated as G2) until the children were 8 years old. The retrospective data come from a follow-up study, conducted in 1992-1994, of G1, G2, and the children born to G2 (the third-generation children, abbreviated as G3). Data from JHCPS on G1 include obstetrical and reproductive history at registration for prenatal care, sociological/family history variables at or around delivery of G2, observations of mother with child when G2 was 4 months old and 8 months old, and family history, demographic, and sociological variables when G2 was age 7. For G2, the data from JHCPS include delivery room observations at birth, pediatric examination data at age 4 months, developmental evaluation data at age 8 months, pediatric-neurological examination data at age 12 months, language, hearing, and speech evaluation summary data at age 36 months, psychological, behavior profile, physical growth, and other tests at age 48 months, psychological, motor, behavior, neurological, vision, physical, and other tests at age 7-1/2 years, and language, hearing, and speech evaluations, physical growth, interval medical history, and other tests at age 8 years. Retrospective data from the follow-up study on G1 include variables on education, employment, family composition, health and health care usage, housing conditions, income and income sources, marital status, partnerships and changes, neighborhood characteristics at registration to JHCPS and current, and reproductive history. For G2, data from the follow-up include information on aspirations, education, schooling, employment, family composition, health and health care usage, housing conditions, income and income sources, legal problems, living arrangements, marriage, partnership and changes, neighborhood characteristics at birth, at ages 11/12 and 16/17, and current, reproductive history, social relationships, smoking, and substance abuse. Data for the assessed third-generation children, i.e., G3s who were 7-8 years old during the follow-up period, include information on cognitive development, academic achievement and behavior, prenatal care, health, day care, and parental aspirations.
Curated

Mexican-American Families in Los Angeles, 1844-1880 (ICPSR 7582)

Released/updated on: 2010-06-29
Geographic coverage: United States, Los Angeles, California
Time period: 1844-01-01--1880-01-01
This data collection contains two data files created from manuscript census returns. Part 1 is an aggregation of social characteristics of Spanish-surnamed and Mexican-born families in the city of Los Angeles from 1844-1880. The data were used to study family composition and socioeconomic mobility. Data items include real property held by head of household (1844, 1850, and 1880 missing), number of children in household, number of adults who were literate in household (no data for 1844), last name of head of household, place of birth of head of household, and occupational category (i.e., rancher or farmer, professional, mercantile, clerk, skilled, and unskilled). Part 2 is composed of data used to study the socioeconomic development of the Mexican-American community in Los Angeles. The main emphasis was on an analysis of literacy, occupational mobility, schooling, family structure, demographic changes, and property mobility. Data items include last name, first name, age, sex, occupational code, real property, personal property, place of birth, literacy, race, head of household, wife of head, child of head, parent of head, sibling of head, and common law spouse. Definitions of family types and discussion of the methodology and rationale used to generate the data in both files can be found in Appendix A of del Castillo, Richard Griswold. "La Raza Hispano Americana: The Emergence of an Urban Culture Among the Spanish Speaking of Los Angeles, 1850-1880." Ph.D. dissertation, University of California, Los Angeles, CA, 1974.
Curated

United States Southern Cities in 1870 and 1880: A Study of Individuals and Families (ICPSR 7568)

Released/updated on: 2006-01-18
Geographic coverage: Charleston (South Carolina), Savannah, United States, Atlanta, Louisiana, New Orleans, Georgia, Alabama, Virginia, Mobile, South Carolina, Norfolk
This data collection contains individual-level and family-level information collected from the 1870 and 1880 manuscript schedules of the United States Population Census for seven Southern cities: Charleston, South Carolina, Richmond, Virginia, Atlanta, Georgia, Savannah, Georgia, Mobile, Alabama, Norfolk, Virginia, and New Orleans, Louisiana. Approximately 5,000 individuals and 1,500 families are represented for each of the two census years studied. Part 1 contains data for 1870, and Part 2 contains data for 1880. The data gathered for sampled individuals include age, sex, race, marital status, presence of health defect, school attendance, ability to read, ability to write, occupational classification (female and male), nationality, and real and personal wealth (for 1870 only). Both datasets include a variable that uniquely identifies each family in the sample to facilitate the aggregation of the data for the creation of family-level data for each member, e.g., sex, race, age, marital status, school attendance, member status in the family, occupation, health, unemployment, city of residence, nationality and parents' nationality, and real and personal wealth.
Curated

Tsogolo La Thanzi 3 (TLT-3): Household Listing Data, Malawi, 2019 [Healthy Futures] (ICPSR 39855)

Released/updated on: 2026-07-13
Geographic coverage: Balaka, Malawi, Africa

Tsogolo la Thanzi (TLT) is a longitudinal study in Balaka, Malwai designed to examine how young people navigate reproduction in an AIDS epidemic. Tsogolo la Thanzi means "Healthy Futures" in Chichewa, which is Malawi's most widely spoken language.

The 2019 Household Listing Data are the result of a new census of the original catchment area, 10 years after the initial study began. These data are supplemental to the main TLT series, providing insights into how the Balaka area developed over the past 10 years. It includes data from all persons living within 7 kilometers of the TLT research center.

Curated
Simple Crosstabs

Land Use, Agropastoral Production, Family Composition, and Household Economy in Santarem, Para, Brazil, June-August 2003 (ICPSR 34347)

Released/updated on: 2013-04-01
Geographic coverage: Brazil, Global, Santarem
The 2003 Santarem dataset consists of 8 interconnected datasets and 1 linking file. The primary unit of analysis is the rural property or lot. Each lot in the sample contains a minimum of 1 household with a mean of 1.33 households per lot in the final sample. Within households, data were collected on subsets of individuals as well as additional properties used by the households in the study. These 2003 Santarem data come from interviews with farm families in an agricultural zone south of the city of Santarem in the Brazilian state of Para. Santarem is a relatively old settlement within the Brazilian Amazon that has experienced waves of regional settlement in the 1930s, mid-century, and the 1970s. The study region is adjacent to the confluence of the Amazon and Tapajos Rivers and the northern terminus of the BR-163 (the Cuiaba-Santarem Highway). BR-163 links intensive agropastoral production (particularly mechanized soybean farming) in the state of Mato Grosso to Santarem, where the multinational corporation Cargill runs a deepwater port (opened in 2003) for loading soybeans onto oceangoing ships. The opening of this port has accelerated the process of urbanization and led to a transformation from a landscape of small family farming to a landscape of mechanized agriculture (description adapted from VanWey, Leah K., and Kara B. Cebulko, 2007, Journal of Marriage and the Family 69: 1257-1270). The discourse on deforestation has focused on the alarming rates of deforestation in the Amazon Basin to the neglect of the dynamic and reciprocal influences between the human population and the environment. Deforestation is a process mediated by human intervention, from the act of clearing to how such a clearing is used and managed over time. It would be helpful to know whether observable rates of forest removal represent a stage in the developmental cycle of households or represents the simple and direct impact of increasing population in these environments. From the point of view of theory and method, it is necessary to develop new approaches that effectively link demographic process to the interactive relationship of population to specific aspects of an environmental matrix. This project addressed multiple scales, from household dynamics to landscape dynamics and has developed methods by which to scale between them. We hypothesize that as households occupy frontier areas past the first generation, they move from a strategy of managing their land under the constraints of available household labor to a strategy that gives greater recognition of the constraints posed by land quality and of the risks to their farm operation coming from external socioeconomic forces and biophysical constraints. In the first generation, the labor available to a household is determined by the size of the household making the initial trip to the frontier (primarily young couples is common in frontier regions) and later by the fertility of these initial migrants. As these initial migrants age and their children enter adulthood (thereby becoming the second generation), labor supply is determined by the reproductive and land use choices of these children. Given the precipitous decline in female fertility, other factors gain salience in the second generation: the suitability of the land for various uses, the availability of off-farm employment and educational opportunities (both locally and those requiring migration), and macroeconomic factors affecting the economic viability of farming. These decisions then directly determine the entries into and exits from the household. This study investigated five basic questions: (1) Does the changing availability of household labor over the household life cycle affect the trajectory of deforestation and land use change in the same way for later generations of Amazonian farmers as for first generation in-migrants? (2) What are the determinants of changing household labor supply? Specifically, what are the biophysical and socioeconomic determinants of entries into and exits from the household through fertility, migration, and marriage? (3) How are the decisions of households regarding land use and labor allocation constrained by soil quality, access to water supplies, interannual drought events (e.g. El Nino type events), and other resource scarcities? (4) Are there notable differences in land use choices made by landholders who live in an urban area (away from the piece of land owned in the rural area) in contrast to the decisions made by those who live on their rural properties? (5) What are the bases for the precipitous decline in female fertility in these frontier regions, especially the use of sterilization after two pregnancies? Households will be surveyed in the Santarem region, in the Lower Tapajos Basin, Brazilian Amazon to collect detailed demographic, land-use histories, and economic data. The sampling of households for inclusion in the study will be based on a stratified random sample by period of occupation in Santarem, to capture intergenerational processes that preceded the availability of satellite images. Based on the particular combination of methodologies used in this investigation (traditional household surveys, satellite image analysis, and GIS, and the scaling up and down from households to landscape), future environmental changes were projected for the regional landscape under various scenarios of continued settlement, household life cycles, combinations of credit, and changing environmental conditions.
Back to top