Adaptation Process of Cuban and Mexican Immigrants in the United States, 1972-1979 (ICPSR 9672)
Agricultural and Demographic Records for Rural Households in the North, 1860: [Instructional Materials] (ICPSR 3463)
Chinese Household Income Project, 1988 (ICPSR 9836)
The purpose of this project was to measure and estimate the distribution of income in both rural and urban areas of the People's Republic of China. The principal investigators based their definition of income on cash payments and on a broad range of additional components: payments in kind valued at market prices, agricultural output produced for self-consumption valued at market prices, the value of ration coupons and other direct subsidies, and the imputed value of housing. The rural component of this collection consists of two data files, one in which the individual is the unit of analysis and a second in which the household is the unit of analysis. Individual rural respondents reported on their employment status, level of education, Communist Party membership, type of employer (e.g., public, private, or foreign), type of economic sector in which employed, occupation, whether they held a second job, retirement status, monthly pension, monthly wage, and other sources of income. Demographic variables include relationship to householder, gender, age, and student status. Rural households reported extensively on the character of the household and residence. Information was elicited on type of terrain surrounding the house, geographic position, type of house, and availability of electricity. Also reported were sources of household income (e.g., farming, industry, government, rents, and interest), taxes paid, value of farm, total amount and type of cultivated land, financial assets and debts, quantity and value of various crops (e.g., grains, cotton, flax, sugar, tobacco, fruits and vegetables, tea, seeds, nuts, lumber, livestock and poultry, eggs, fish and shrimp, wool, honey, and silkworm cocoons), amount of grain purchased or provided by a collective, use of chemical fertilizers, gasoline, and oil, quantity and value of agricultural machinery, and all household expenditures (e.g., food, fuel, medicine, education, transportation, and electricity). The urban component of this collection also consists of two data files, one in which the individual is the unit of analysis and a second in which the household is the unit of analysis. Individual urban respondents reported on their economic status within the household, Communist Party membership, sex, age, nature of employment, and relationship to the household head. Information was collected on all types and sources of income from each member of the household whether working, nonworking, or retired, all revenue received by owners of private or individual enterprises, and all in-kind payments (e.g., food and durable and non-durable goods). Urban households reported total income (including salaries, interest on savings and bonds, dividends, rent, leases, alimony, gifts, and boarding fees), all types and values of food rations received, and total debt. Information was also gathered on household accommodations and living conditions, including number of rooms, total living area in square meters, availability and cost of running water, sanitary facilities, heating and air-conditioning equipment, kitchen availability, location of residence, ownership of home, and availability of electricity and telephone. Households reported on all of their expenditures including amounts spent on food items such as wheat, rice, edible oils, pork, beef and mutton, poultry, fish and seafood, sugar, and vegetables by means of both coupons in state-owned stores and at free market prices. Information was also collected on rents paid by the households, fuel available, type of transportation used, and availability and use of medical and child care.
The Chinese Household Income Project collected data in 1988, 1995, 2002, and 2007. ICPSR holds data from the first three collections, and information about these can be found on the series description page. Data collected in 2007 are available through the China Institute for Income Distribution.
Chinese Household Income Project, 1995 (ICPSR 3012)
The purpose of this project was to measure and estimate the distribution of personal income in both rural and urban areas of the People's Republic of China. The principal investigators based their definition of income on cash payments and on a broad range of additional components: payments in kind valued at market prices, agricultural output produced for self-consumption valued at market prices, the value of food and other direct subsidies, and the imputed value of housing services. The rural component of this collection consists of two data files, one in which the individual is the unit of analysis (Part 1) and a second in which the household is the unit of analysis (Part 2). Individual rural respondents reported on their employment status, level of education, Communist Party membership, type of employer (e.g., public, private, or foreign), type of economic sector in which they were employed, occupation, whether they held a second job, retirement status, monthly pension, monthly wage, and other sources of income. Demographic variables include relationship to householder, gender, age, and student status. Rural households reported extensively on the character of the household and residence. Information was elicited on type of terrain surrounding the house, geographic position, type of house, and availability of electricity. Also reported were sources of household income (e.g., farming, industry, government, rents, and interest), taxes paid, value of farm, total amount and type of cultivated land, financial assets and debts, quantity and value of various crops, amount of grain purchased or provided by a collective, use of chemical fertilizers, gasoline, and oil, quantity and value of agricultural machinery, and all household expenditures (e.g., food, fuel, medicine, education, transportation, and electricity). The urban component of this collection also consists of two data files, one in which the individual is the unit of analysis (Part 3) and a second in which the household is the unit of analysis (Part 4). Individual urban respondents reported on their economic status within the household, Communist Party membership, sex, age, nature of employment, and relationship to the household head. Information was collected on all types and sources of income from each member of the household whether working, nonworking, or retired, all revenue received by owners of private or individual enterprises, and all in-kind payments (e.g., food, durable goods, and nondurable goods). Urban households reported total income (including salaries, interest on savings and bonds, dividends, rent, leases, alimony, gifts, and boarding fees), all types and values of food subsidies received, and total debt. Information was also gathered on household accommodations and living conditions, including number of rooms, total living area in square meters, availability and cost of running water, sanitary facilities, heating and air-conditioning equipment, kitchen availability, location of residence, ownership of home, and availability of electricity and telephone. Households reported on all their expenditures including amounts spent on food items such as wheat, rice, edible oils, pork, beef and mutton, poultry, fish and seafood, sugar, and vegetables by means of coupons in state-owned stores and at free market prices. Information was also collected on rents paid by the households, fuel available, type of transportation used, and availability and use of medical and child care.
The Chinese Household Income Project collected data in 1988, 1995, 2002, and 2007. ICPSR holds data from the first three collections, and information about these can be found on the series description page. Data collected in 2007 are available through the China Institute for Income Distribution.
Consumer Pyramids Survey, 2014 [India] (ICPSR 36782)
The Consumer Pyramids is the largest survey of households in India. The survey contains record-level data that are delivered in the form of population estimates. The survey contains multiple databases that contain population estimates on household demographics, household income and expenses, borrowing by household, and household assets. The data also contain individual-level health status, financial inclusion, education level, and caste and literacy estimates. Demographic information collected include gender, age, religion, education, and occupation.
Database Composition: The Consumer Pyramids Survey is conducted over the course of four-month periods or waves throughout the year totaling three rounds a year. This collection includes the following six databases: People of India; Household Income and Expenses; Household Amenities, Assets, and Liabilities; Household Expenses; Composition of Incomes at the member level; Composition of Incomes at the household level.
Demographic Characteristics of the Population of Detroit, 1850-1880 (ICPSR 31)
Demographic Characteristics of Washtenaw County, Michigan, in 1860 (ICPSR 8445)
Demographic, Social, Educational and Economic Data for France, 1833-1925 (ICPSR 7529)
Demography in Frontier Indiana, 1820 (ICPSR 7504)
Historical Demographic Data of Southeastern Europe: Orasac, 1824-1975 (ICPSR 32404)
The data in the Historical Demographic Data of Southeastern Europe series derive primarily from the ethnographic and archival research of Joel M. Halpern, Professor Emeritus of Anthropology at the University of Massachusetts at Amherst, in southeastern Europe from 1953 to 2006. The series is comprised of historical demographic data from several towns and villages in the countries of Bosnia, Croatia, Macedonia, Montenegro, Serbia, and Slovenia, all of which are former constituent republics of the Socialist Federal Republic of Yugoslavia. The data provide insight into the shift from agricultural to industrial production, as well as the more general processes of urbanization occurring in the last days of the Yugoslav state. With an expansive timeframe ranging from 1818 to 2006, the series also contains a wide cross-section of demographic data types. These include, but are not limited to, population censuses, tax records, agricultural and landholding data, birth records, death records, marriage and engagement records, and migration information.
This component of the series focuses exclusively on the Serbian village of Orasac and is composed of 64 datasets. These data record a variety of demographic and economic information between the years of 1824 and 1975. General population information at the individual level is available in official census records from 1863, 1884, 1948, 1953, and 1961, and from population register records for the years of 1928, 1966, and 1975. Census data at the household level is also available for the years of 1863, 1928, 1948, 1953, and 1961. These data are followed by detailed records of engagement and marriage. Many of these data were obtained through the courtesy of village and county officials. Priest book records from 1851 through 1966, as well as death records from 1863 to 1976 and tombstone records from 1975, are also available. Information regarding migrants and emigrants was obtained from the village council for the years of 1946 through 1975. Lastly, the data provide economic and financial information, including records of individual landholdings (for the years of 1863, 1952, 1966, and 1975), records of government taxation at the individual or household level (for 1813 through 1840, as well as for 1952), and livestock censuses (at both the individual and household level for the years of 1824 and 1825, and only at the individual level for the years of 1833 and 1834).
Historical, Demographic, Economic, and Social Data: The United States, 1790-1970 (ICPSR 3)
Historical, Demographic, Economic, and Social Data: The United States, 1790-2002 (ICPSR 2896)
International Data Base, February 1990 (ICPSR 8490)
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 1, Public Data, 2000-2001 (ICPSR 37279)
This study includes public user data files of two waves of interviews with L.A.FANS respondents. There often are multiple respondents in L.A.FANS households and Wave 2 includes both panel respondents and a new sample. Users' Guides which explain the design and how to use the sample are available for Wave 1 and Wave 2 at the RAND website.
The Los Angeles Family and Neighborhood Survey (L.A.FANS) is a two-wave study of adults and children in Los Angeles County and of the neighborhoods in which they live. The first wave (L.A.FANS-1), which was fielded between April 2000 and January 2002, interviewed adults and children living in 3,085 households in a stratified probability sample of 65 neighborhoods throughout Los Angeles County. The samples of neighborhoods and individuals were representative of neighborhoods and residents of Los Angeles County. Poorer neighborhoods and households with children were oversampled. In Wave 2 of L.A.FANS (L.A.FANS-2), Wave 1 respondents living in Los Angeles County were reinterviewed and updated information was collected on Wave 1 respondents who had moved away from Los Angeles County. A sample of individuals who moved into each sampled neighborhood between Waves 1 and 2 was also interviewed, for a total of 2,319 adults and 1,382 children (ages less than 18 years). Additional information on the project is available at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 1, Restricted Data Version 1, 2000-2001 (ICPSR 37242)
This study includes a restricted data file for Wave 1 of the L.A.FANS data. To compare L.A.FANS restricted data, version 1 with other restricted data versions, see the table on the series page for the L.A.FANS data here. Data in this study are designed for use with the public use data files for L.A.FANS, Wave 1 (study 1). This file adds only a few variables to the L.A.FANS, Wave 1 public use files. Specifically, it adds a "pseudo-tract ID" which is a number from 1 to 65, randomly assigned to each census tract (neighborhood) in the study. It is not possible to link pseudo-tract IDs in any way to real tract IDs or other neighborhood characteristics. However, pseudo-tract IDs permit users to conduct analyses which take into account the clustered sample design in which neighborhoods (tracts) were selected first and then individuals were sampled within neighborhoods. Pseudo-tract IDs do so because they identify which respondents live in the same neighborhood. It also includes certain variables, thought to be sensitive, which are not available in the public use data. These variables are identified in the L.A.FANS Wave 1 Users Guide and Codebook. Finally, some distance variables and individual characteristics which are treated in the public use data to make it harder to identify individuals are provided in an untreated form in the Version 1 restricted data file. Please note that L.A. FANS restricted data may only be accessed within the ICPSR Virtual Data Enclave (VDE) and must be merged with the L.A. FANS public data prior to beginning any analysis.
A Users' Guide which explains the design and how to use the samples are available for Wave 1 at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 1, Restricted Data Version 2, 2000-2001 (ICPSR 37269)
This study includes restricted data file, version 2, for Wave 1 of the L.A.FANS data. To compare L.A.FANS restricted data, version 2 with other restricted data versions, see the table on the series page for the L.A.FANS data here. Data in this study are designed for use with the public use data files for L.A.FANS, Wave 1 (study 1). This file adds only a few variables to the L.A.FANS, Wave 1 public use files. Specifically, it adds the census tract number for the tract each respondent lives in. It also includes certain variables, thought to be sensitive, which are not available in the public use data. These variables are identified in the L.A.FANS Wave 1 Users Guide and Codebook. Finally, some distance variables and individual characteristics which are treated in the public use data to make it harder to identify individuals are provided in an untreated form in the Version 2 restricted data file. Please note that L.A. FANS restricted data may only be accessed within the ICPSR Virtual Data Enclave (VDE) and must be merged with the L.A. FANS public data prior to beginning any analysis.
A Users' Guide which explains the design and how to use the samples are available for Wave 1 at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 1, Restricted Data Version 2.5, 2000-2001 (ICPSR 37270)
This study includes restricted data version 2.5, for Wave 1 of the L.A.FANS data. To compare L.A.FANS restricted data, version 2.5 with other restricted data versions, see the table on the series page for the L.A.FANS data here. Data in this study are designed for use with the public use data files for L.A.FANS, Wave 1 (study 1). This file adds only a few variables to the L.A.FANS, Wave 1 public use files. Specifically, it adds the census tract and block number for the tract each respondent lives in. It also includes certain variables, thought to be sensitive, which are not available in the public use data. These variables are identified in the L.A.FANS Wave 1 Users Guide and Codebook. Finally, some distance variables and individual characteristics which are treated in the public use data to make it harder to identify individuals are provided in an untreated form in the Version 2.5 restricted data file. Please note that L.A. FANS restricted data may only be accessed within the ICPSR Virtual Data Enclave (VDE) and must be merged with the L.A. FANS public data prior to beginning any analysis.
A Users' Guide which explains the design and how to use the samples are available for Wave 1 at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 1, Restricted Data Version 3, 2000-2001 (ICPSR 37271)
This study includes restricted data version 3, for Wave 1 of the L.A.FANS data. To compare L.A.FANS restricted data, version 3 with other restricted data versions, see the table on the series page for the L.A.FANS data here. Data in this study are designed for use with the public use data files for L.A.FANS, Wave 1 (study 1). This file adds only a few variables to the L.A.FANS, Wave 1 public use files. Specifically, it adds the census tract and block number for the tract each respondent lives in and geographic coordinates data for a number of locations reported by the respondent (including home, grocery store, place of work, place of worship, schools, etc.). It also includes certain variables, thought to be sensitive, which are not available in the public use data. These variables are identified in the L.A.FANS Wave 1 Users Guide and Codebook. Finally, some distance variables and individual characteristics which are treated in the public use data to make it harder to identify individuals are provided in an untreated form in the Version 3 restricted data file. Please note that L.A. FANS restricted data may only be accessed within the ICPSR Virtual Data Enclave (VDE) and must be merged with the L.A. FANS public data prior to beginning any analysis.
A Users' Guide which explains the design and how to use the samples are available for Wave 1 at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 2, Public Data, 2006-2008 (ICPSR 37278)
This study includes public user data files of two waves of interviews with L.A.FANS respondents. There often are multiple respondents in L.A.FANS households and Wave 2 includes both panel respondents and a new sample. Users' Guides which explain the design and how to use the sample are available for Wave 1 and Wave 2 at the RAND website.
The Los Angeles Family and Neighborhood Survey (L.A.FANS) is a two-wave study of adults and children in Los Angeles County and of the neighborhoods in which they live. The first wave (L.A.FANS-1), which was fielded between April 2000 and January 2002, interviewed adults and children living in 3,085 households in a stratified probability sample of 65 neighborhoods throughout Los Angeles County. The samples of neighborhoods and individuals were representative of neighborhoods and residents of Los Angeles County. Poorer neighborhoods and households with children were oversampled. In Wave 2 of L.A.FANS (L.A.FANS-2), Wave 1 respondents living in Los Angeles County were reinterviewed and updated information was collected on Wave 1 respondents who had moved away from Los Angeles County. A sample of individuals who moved into each sampled neighborhood between Waves 1 and 2 was also interviewed, for a total of 2,319 adults and 1,382 children (ages less than 18 years). Additional information on the project is available at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 2, Restricted Data Version 1, 2006-2008 (ICPSR 37259)
This study includes a restricted data file for Wave 2 of the L.A.FANS data. To compare L.A.FANS restricted data, version 1 with other restricted data versions, see the table on the series page for the L.A.FANS data here. Data in this study are designed for use with the public use data files for L.A.FANS, Wave 2 (study 2). This file adds only a few variables to the L.A.FANS, Wave 2 public use files. Specifically, it adds a "pseudo-tract ID" which is a number from 1 to 65, randomly assigned to each census tract (neighborhood) in the study. It is not possible to link pseudo-tract IDs in any way to real tract IDs or other neighborhood characteristics. However, pseudo-tract IDs permit users to conduct analyses which take into account the clustered sample design in which neighborhoods (tracts) were selected first and then individuals were sampled within neighborhoods. Pseudo-tract IDs do so because they identify which respondents live in the same neighborhood. It also includes certain variables, thought to be sensitive, which are not available in the public use data. These variables are identified in the L.A.FANS Wave 2 Users Guide and Codebook. Finally, some distance variables and individual characteristics which are treated in the public use data to make it harder to identify individuals are provided in an untreated form in the Version 1 restricted data file. Please note that L.A. FANS restricted data may only be accessed within the ICPSR Virtual Data Enclave (VDE) and must be merged with the L.A. FANS public data prior to beginning any analysis.
A Users' Guide which explains the design and how to use the samples are available for Wave 2 at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 2, Restricted Data Version 2, 2006-2008 (ICPSR 37265)
This study includes a restricted data file, version 2, for Wave 2 of the L.A.FANS data. To compare L.A.FANS restricted data, version 2 with other restricted data versions, see the table on the series page for the L.A.FANS data here. Data in this study are designed for use with the public use data files for L.A.FANS, Wave 2 (study 2). This file adds only a few variables to the L.A.FANS, Wave 2 public use files. Specifically, it adds the census tract number for the tract each respondent lives in. It also includes certain variables, thought to be sensitive, which are not available in the public use data. These variables are identified in the L.A.FANS Wave 2 Users Guide and Codebook. Finally, some distance variables and individual characteristics which are treated in the public use data to make it harder to identify individuals are provided in an untreated form in the Version 2 restricted data file. Please note that L.A. FANS restricted data may only be accessed within the ICPSR Virtual Data Enclave (VDE) and must be merged with the L.A. FANS public data prior to beginning any analysis.
A Users' Guide which explains the design and how to use the samples are available for Wave 2 at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 2, Restricted Data Version 2.5, 2006-2008 (ICPSR 37266)
This study includes a restricted data file, version 2.5, for Wave 2 of the L.A.FANS data. To compare L.A.FANS restricted data, version 2.5 with other restricted data versions, see the table on the series page for the L.A.FANS data here. Data in this study are designed for use with the public use data files for L.A.FANS, Wave 2 (study 2). This file adds only a few variables to the L.A.FANS, Wave 2 public use files. Specifically, it adds the census tract and block number for the tract each respondent lives in. It also includes certain variables, thought to be sensitive, which are not available in the public use data. These variables are identified in the L.A.FANS Wave 2 Users Guide and Codebook. Finally, some distance variables and individual characteristics which are treated in the public use data to make it harder to identify individuals are provided in an untreated form in the Version 2.5 restricted data file. Please note that L.A. FANS restricted data may only be accessed within the ICPSR Virtual Data Enclave (VDE) and must be merged with the L.A. FANS public data prior to beginning any analysis.
A Users' Guide which explains the design and how to use the samples are available for Wave 2 at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 2, Restricted Data Version 3, 2006-2008 (ICPSR 37267)
This study includes a restricted data file, version 3, for Wave 2 of the L.A.FANS data. To compare L.A.FANS restricted data, version 3 with other restricted data versions, see the table on the series page for the L.A.FANS data here. Data in this study are designed for use with the public use data files for L.A.FANS, Wave 2 (study 2). This file adds only a few variables to the L.A.FANS, Wave 2 public use files. Specifically, it adds the census tract and block number for the tract each respondent lives in and geographic coordinates data for a number of locations reported by the respondent (including home, grocery store, place of work, place of worship, schools, etc.). It also includes certain variables, thought to be sensitive, which are not available in the public use data. These variables are identified in the L.A.FANS Wave 2 Users Guide and Codebook. Finally, some distance variables and individual characteristics which are treated in the public use data to make it harder to identify individuals are provided in an untreated form in the Version 3 restricted data file. Please note that L.A. FANS restricted data may only be accessed within the ICPSR Virtual Data Enclave (VDE) and must be merged with the L.A. FANS public data prior to beginning any analysis.
A Users' Guide which explains the design and how to use the samples are available for Wave 2 at the RAND website.
Additional information on the project, survey design, sample, and variables are available from:
- Sastry, Narayan, Bonnie Ghosh-Dastidar, John Adams, and Anne R. Pebley (2006). The Design of a Multilevel Survey of Children, Families, and Communities: The Los Angeles Family and Neighborhood Survey, Social Science Research, Volume 35, Number 4, Pages 1000-1024
- The Users' Guides (Wave 1 and Wave 2)
- RAND Documentation Reports page
Los Angeles Family and Neighborhood Survey (L.A.FANS), Wave 3, Public Data | Mixed Income Project (MIP), 2011-2013 (ICPSR 37845)
This study includes one public use data file of follow-up interviews, conducted between 2011 and 2013, with respondents to Wave 2 of L.A.FANS (Los Angeles Family and Neighborhood Survey). This follow-up data collection effort (hereafter called L.A.FANS-3 or Wave 3) was part of the broader Mixed Income Project (MIP), which was designed to allow for detailed examination of neighborhood context, residential mobility, and mixed-income housing in Los Angeles and Chicago. The two anchor studies for the MIP are L.A.FANS and the Project on Human Development in Chicago Neighborhoods (PHDCN).
Wave 3 targeted a random probability sample of approximately 1,000 randomly selected adults and children from the prior wave of L.A.FANS, which was fielded between 2006 and 2008, who still resided within Los Angeles County. The Los Angeles field operation first assigned selected respondents to a telephone survey center for interviews. Cases that were not interviewed by telephone were transferred to experienced field interviewers in the Los Angeles area. The final response rate was 75 percent of eligible participants (i.e., residents who still resided in Los Angeles County and who were not institutionalized, incapacitated, or deceased) for a combined sample of 1,032. Two-hundred and two (202) of these respondents were reached during a preliminary Field Test in 2011, after which point the survey was slightly revised. After making these revisions, 830 respondents were reached during the Main Study. For more details on sampling procedures for the Field Test and Main Study, see Methodology section below.
For context, the L.A.FANS is a study of adults and children in Los Angeles County, and of the neighborhoods in which they live. The first wave (L.A.FANS-1 or Wave 1), which was fielded between April 2000 and January 2002, interviewed adults and children living in 3,085 households in a stratified probability sample of 65 neighborhoods throughout Los Angeles County. The samples of neighborhoods and individuals were representative of neighborhoods and residents of Los Angeles County. Poorer neighborhoods and households with children were oversampled. In Wave 2 of L.A.FANS (L.A.FANS-2), Wave 1 respondents still living in Los Angeles County were re-interviewed, while updated information was collected on Wave 1 respondents who had moved away from Los Angeles County. A sample of individuals who moved into each sampled neighborhood between Waves 1 and 2 was also interviewed, for a total of 2,319 adults and 1,382 children (ages less than 18 years). Additional information on the project is available at the RAND website.
The Mexican American Study Project II (MASP II), 1998-2000 (ICPSR 28481)
Midlife in the United States (MIDUS 3), 2013-2014 (ICPSR 36346)
In 1995-1996, the MacArthur Midlife Research Network carried out a national survey of over 7,000 Americans aged 25 to 74 [ICPSR 2760]. The purpose of the study was to investigate the role of behavioral, psychological, and social factors in understanding age-related differences in physical and mental health. The study was innovative for its broad scientific scope, its diverse samples (which included siblings of the main sample respondents and a national sample of twin pairs), and its creative use of in-depth assessments in key areas (e.g. daily diary of stressful experiences [ICPSR 3725] and cognitive functioning [ICPSR 3596]) on a subset of participants. A detailed description of the study and findings generated by it are available at: http://www.midus.wisc.edu
With support from the National Institute on Aging, a follow-up of the original Midlife Development in the United States (MIDUS) sample was conducted in 2004 (MIDUS 2 [ICPSR 4652]). The daily stress and cognitive functioning projects were repeated and expanded at MIDUS 2; in addition the protocol was expanded to include biomarkers and neuroscience.
In 2013 a third wave (MIDUS 3) of survey data was collected on longitudinal participants. Data collection for this follow-up wave largely repeated baseline assessments (e.g., phone interview and extensive self-administered questionnaire), with additional questions in selected areas such as economic recession experiences. Cognitive functioning data were also collected at the same time, while data collection for the daily diary, biomarker, and neuroscience projects commenced in 2017.
MIDUS also maintains a Colectica portal, which allows users to interact with variables across waves and create customized subsets. Registration is required.
National Health and Nutrition Examination Survey (NHANES), 1999-2000 (ICPSR 25501)
National Health and Nutrition Examination Survey (NHANES), 2001-2002 (ICPSR 25502)
National Health and Nutrition Examination Survey (NHANES), 2003-2004 (ICPSR 25503)
The National Health and Nutrition Examination Surveys (NHANES) is a program of studies designed to assess the health and nutritional status of adults and children in the United States. The NHANES combines personal interviews and physical examinations, which focus on different population groups or health topics. These surveys have been conducted by the National Center for Health Statistics (NCHS) on a periodic basis from 1971 to 1994. In 1999 the NHANES became a continuous program with a changing focus on a variety of health and nutrition measurements which were designed to meet current and emerging concerns. The surveys examine a nationally representative sample of approximately 5,000 persons each year. These persons are located in counties across the United States, 15 of which are visited each year.
For NHANES 2003-2004, there were 12,761 persons selected for the sample, 10,122 of those were interviewed (79.3 percent) and 9,643 (75.6 percent) were examined in the mobile examination centers (MEC). Many of the NHANES 2003-2004 questions were also asked in NHANES II 1976-1980, Hispanic HANES 1982-1984, NHANES III 1988-1994, and NHANES 1999-2002. New questions were added to the survey based on recommendations from survey collaborators, NCHS staff, and other interagency work groups. As in past health examination surveys, data were collected on the prevalence of chronic conditions in the population. Estimates for previously undiagnosed conditions, as well as those known to and reported by survey respondents, are produced through the survey. Risk factors, those aspects of a person's lifestyle, constitution, heredity, or environment that may increase the chances of developing a certain disease or condition, were examined. Data on smoking, alcohol consumption, sexual practices, drug use, physical fitness and activity, weight, and dietary intake were collected. Information on certain aspects of reproductive health, such as use of oral contraceptives and breastfeeding practices, were also collected. The diseases, medical conditions, and health indicators that were studied include: anemia, cardiovascular disease, diabetes and lower extremity disease, environmental exposures, equilibrium, hearing loss, infectious diseases and immunization, kidney disease, mental health and cognitive functioning, nutrition, obesity, oral health, osteoporosis, physical fitness and physical functioning, reproductive history and sexual behavior, respiratory disease (asthma, chronic bronchitis, emphysema), sexually transmitted diseases, skin diseases, and vision. The sample for the survey was selected to represent the United States population of all ages. Special emphasis in the 2003-2004 NHANES was on adolescent health and the health of older Americans. To produce reliable statistics for these groups, adolescents aged 15-19 years and persons aged 60 years and older were over-sampled for the survey. African Americans and Mexican Americans were also over-sampled to enable accurate estimates for these groups. Several important areas in adolescent health, including nutrition and fitness and other aspects of growth and development, were addressed. Since the United States has experienced dramatic growth in the number of older people during the twentieth century, the aging population has major implications for health care needs, public policy, and research priorities. NCHS is working with public health agencies to increase the knowledge of the health status of older Americans. NHANES has a primary role in this endeavor. In the examination, all participants visit the physician who takes their pulse or blood pressure. Dietary interviews and body measurements are included for everyone. All but the very young have a blood sample taken and see the dentist. Depending upon the age of the participant, the rest of the examination includes tests and procedures to assess the various aspects of health listed above. Usually, the older the individual, the more extensive the examination. Some persons who are unable or unwilling to come to the examination center may be given a less extensive examination in their homes.
Demographic data file variables are grouped into three broad categories: (1) Status Variables: provide core information on the survey participant. Examples of the core variables include interview status, examination status, and sequence number. (Sequence number is a unique ID assigned to each sample person and is required to match the information on this demographic file to the rest of the NHANES 2003-2004 data). (2) Recoded Demographic Variables: these variables include age (age in months for persons through age 19 years, 11 months; age in years for 1- to 84-year-olds, and a top-coded age group of 85 years of age and older), gender, a race/ethnicity variable, current or highest grade of education completed, (less than high school, high school, and more than high school education), country of birth (United States, Mexico, or other foreign born), Poverty Income Ratio (PIR), income, and a pregnancy status variable (adjudicated from various pregnancy related variables). Some of the groupings were made due to limited sample sizes for the two-year data set. (3) Interview and Examination Sample Weight Variables: sample weights are available for analyzing NHANES 2003-2004 data. For a complete listing of survey contents for all years of the NHANES see the document -- Survey Content -- NHANES 1999-2010.
National Health and Nutrition Examination Survey (NHANES), 2005-2006 (ICPSR 25504)
National Health and Nutrition Examination Survey (NHANES), 2007-2008 (ICPSR 25505)
National Longitudinal Survey (NLS) of College Graduates, 1967-1985 (ICPSR 9390)
National Longitudinal Survey of Older and Young Men (ICPSR 34937)
The New Immigrant Survey Round 1 (NIS-2003-1), United States, 2003-2004 [Public and Restricted-Use Version 1] (ICPSR 38031)
The New Immigrant Survey (NIS) was a nationally representative, longitudinal study of new legal immigrants to the United States and their children. The sampling frame was based on the electronic administrative records compiled for new legal permanent residents (LPRs) by the U.S. government (via, formerly, the U.S. Immigration and Naturalization Service (INS) and now its successor agencies, the U.S. Citizenship and Immigration Services (USCIS) and the Office of Immigration Statistics (OIS)). The sample was drawn from new legal immigrants during May through November of 2003. The geographic sampling design took advantage of the natural clustering of immigrants. It included all top 85 Metropolitan Statistical Areas (MSAs) and all top 38 counties, plus a random sample of MSAs and counties. The baseline survey was conducted from June 2003 to June 2004 and yielded data on:
- 8,573 Adult Sample respondents,
- 810 sponsor-parents of the Sampled Child,
- 4,915 spouses,
- and 1,072 children aged 8-12.
Interviews were conducted in the respondents' language of choice. The Round 1 questionnaire items that were used in social-demographic-migration surveys around the world as well as the major U.S. longitudinal surveys were reviewed in order to achieve comparability. The NIS content includes the following information: demographic, health and insurance, migration history, living conditions, transfers, employment history, income, assets, social networks, religion, housing environment, and child assessment tests.
The New Immigrant Survey Round 1 (NIS-2003-1), United States, 2003-2004 [Restricted-Use Version 2] (ICPSR 38063)
The New Immigrant Survey (NIS) is a nationally representative, multi-cohort, longitudinal study of new legal immigrants to the United States and their children. The sampling frame is based on the electronic administrative records compiled for new legal permanent residents (LPRs) by the U.S. government (via, formerly, the U.S. Immigration and Naturalization Service (INS) and now its successor agencies, the U.S. Citizenship and Immigration Services (USCIS) and the Office of Immigration Statistics (OIS)). The geographic sampling design takes advantage of the natural clustering of immigrants. It includes all top 85 Metropolitan Statistical Areas (MSAs) and all top 38 counties, plus a random sample of MSAs and counties. The baseline survey (ICPSR 38031) was conducted from June 2003 to June 2004 and yielded data on:
- 8,573 Adult Sample respondents
- 810 sponsor-parent of the Sampled Child
- 4,915 spouses
- and 1,072 children aged 8-12.
Interviews were conducted in the respondents' language of choice. Round 2 instruments were designed to track changes from the baseline and also included new questions. As with the Round 1 questionnaire, questions that were used in social-demographic-migration surveys around the world as well as the major U.S. longitudinal surveys were reviewed in order to achieve comparability.
The New Immigrant Survey Round 2 (NIS-2003-2), United States, 2007-2009 [Public and Restricted-Use Version 1] (ICPSR 38061)
The New Immigrant Survey (NIS) was a nationally representative, longitudinal study of new legal immigrants to the United States and their children. The sampling frame was based on the electronic administrative records compiled for new legal permanent residents (LPRs) by the U.S. government (via, formerly, the U.S. Immigration and Naturalization Service (INS) and now its successor agencies, the U.S. Citizenship and Immigration Services (USCIS) and the Office of Immigration Statistics (OIS)). The sample was drawn from new legal immigrants during May through November of 2003. The geographic sampling design took advantage of the natural clustering of immigrants. It included all top 85 Metropolitan Statistical Areas (MSAs) and all top 38 counties, plus a random sample of MSAs and counties. The baseline survey (ICPSR 38031) was conducted from June 2003 to June 2004 and yielded data on:
- 8,573 Adult Sample respondents,
- 810 sponsor-parents of the Sampled Child,
- 4,915 spouses,
- and 1,072 children aged 8-12.
This study contains the follow-up interview, conducted from June 2007 to October 2009, and yielded data on:
- 3,902 Adult Sample respondents,
- 351 sponsor-parents of the Sampled Child,
- 1,771 spouses,
- and 41 now-adult main children.
Interviews were conducted in the respondents' language of choice. Round 2 instruments were designed to track changes from the baseline and also included new questions. As with the Round 1 questionnaire, questions that were used in social-demographic-migration surveys around the world as well as the major U.S. longitudinal surveys were reviewed in order to achieve comparability. The NIS content includes the following information: demographics, health and insurance, migration history, living conditions, transfers, employment history, income, assets, social networks, religion, housing environment, and child assessment tests.
The New Immigrant Survey Round 2 (NIS-2003-2), United States, 2007-2009 [Restricted-Use Version 2] (ICPSR 38064)
The New Immigrant Survey (NIS) was a nationally representative, longitudinal study of new legal immigrants to the United States and their children. The sampling frame was based on the electronic administrative records compiled for new legal permanent residents (LPRs) by the U.S. government (via, formerly, the U.S. Immigration and Naturalization Service (INS) and now its successor agencies, the U.S. Citizenship and Immigration Services (USCIS) and the Office of Immigration Statistics (OIS)). The sample was drawn from new legal immigrants during May through November of 2003. The geographic sampling design took advantage of the natural clustering of immigrants. It included all top 85 Metropolitan Statistical Areas (MSAs) and all top 38 counties, plus a random sample of MSAs and counties. The baseline survey (ICPSR 38031) was conducted from June 2003 to June 2004 and yielded data on:
- 8,573 Adult Sample respondents,
- 810 sponsor-parents of the Sampled Child,
- 4,915 spouses,
- and 1,072 children aged 8-12.
This study contains the follow-up interview, conducted from June 2007 to October 2009, and yielded data on:
Interviews were conducted in the respondents' language of choice. Round 2 instruments were designed to track changes from the baseline and also included new questions. As with the Round 1 questionnaire, questions that were used in social-demographic-migration surveys around the world as well as the major U.S. longitudinal surveys were reviewed in order to achieve comparability. The NIS content includes the following information: demographic, health and insurance, migration history, living conditions, transfers, employment history, income, assets, social networks, religion, housing environment, and child assessment tests.