Population Assessment of Tobacco and Health (PATH) Study [United States] Public-Use Files (ICPSR 36498)
The Population Assessment of Tobacco and Health (PATH) Study began originally surveying 45,971 adult and youth respondents. The PATH Study was launched in 2011 to inform Food and Drug Administration's regulatory activities under the Family Smoking Prevention and Tobacco Control Act (TCA). The PATH Study is a collaboration between the National Institute on Drug Abuse (NIDA), National Institutes of Health (NIH), and the Center for Tobacco Products (CTP), Food and Drug Administration (FDA). The study sampled over 150,000 mailing addresses across the United States to create a national sample of people who use or do not use tobacco.
45,971 adults and youth constitute the first (baseline) wave of data collected by this longitudinal cohort study. These 45,971 adults and youth along with 7,207 "shadow youth" (youth ages 9 to 11 sampled at Wave 1) make up the 53,178 participants that constitute the Wave 1 Cohort. Respondents are asked to complete an interview at each follow-up wave. Youth who turn 18 by the current wave of data collection are considered "aged-up adults" and are invited to complete the Adult Interview. Additionally, "shadow youth" are considered "aged-up youth" upon turning 12 years old, when they are asked to complete an interview after parental consent.
At Wave 4, a probability sample of 14,098 adults, youth, and shadow youth ages 10 to 11 was selected from the civilian, noninstitutionalized population at the time of Wave 4. This sample was recruited from residential addresses not selected for Wave 1 in the same sampled Primary Sampling Units (PSUs) and segments using similar within-household sampling procedures. This "replenishment sample" was combined for estimation and analysis purposes with Wave 4 adult and youth respondents from the Wave 1 Cohort who were in the civilian, noninstitutionalized population at the time of Wave 4. This combined set of Wave 4 participants, 52,731 participants in total, forms the Wave 4 Cohort.
Dataset 0001 (DS0001) contains the data from the Master Linkage file. This file contains 14 variables and 67,276 cases. The file provides a master list of every person's unique identification number and what type of respondent they were for each wave.
At Wave 7, a probability sample of 14,863 adults, youth, and shadow youth ages 9 to 11 was selected from the civilian, noninstitutionalized population at the time of Wave 7. This sample was recruited from residential addresses not selected for Wave 1 or Wave 4 in the same sampled PSUs and segments using similar within-household sampling procedures. This second replenishment sample was combined for estimation and analysis purposes with Wave 7 adult and youth respondents from the Wave 4 Cohort who were at least age 15 and in the civilian, noninstitutionalized population at the time of Wave 7. This combined set of Wave 7 participants, 46,169 participants in total, forms the Wave 7 Cohort.
Please refer to the Public-Use Files User Guide that provides further details about children designated as "shadow youth" and the formation of the Wave 1, Wave 4, and Wave 7 Cohorts.
Dataset 1001 (DS1001) contains the data from the Wave 1 Adult Questionnaire. This data file contains 1,732 variables and 32,320 cases. Each of the cases represents a single, completed interview.
Dataset 1002 (DS1002) contains the data from the Youth and Parent Questionnaire. This file contains 1,228 variables and 13,651 cases.
Dataset 2001 (DS2001) contains the data from the Wave 2 Adult Questionnaire. This data file contains 2,197 variables and 28,362 cases. Of these cases, 26,447 also completed a Wave 1 Adult Questionnaire. The other 1,915 cases are "aged-up adults" having previously completed a Wave 1 Youth Questionnaire.
Dataset 2002 (DS2002) contains the data from the Wave 2 Youth and Parent Questionnaire. This data file contains 1,389 variables and 12,172 cases. Of these cases, 10,081 also completed a Wave 1 Youth Questionnaire. The other 2,091 cases are "aged-up youth" having previously been sampled as "shadow youth."
Dataset 3001 (DS3001) contains the data from the Wave 3 Adult Questionnaire. This data file contains 2,139 variables and 28,148 cases. Of these cases, 26,241 are continuing adults having completed a prior Adult Questionnaire. The other 1,907 cases are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 3002 (DS3002) contains the data from the Wave 3 Youth and Parent Questionnaire. This data file contains 1,309 variables and 11,814 cases. Of these cases, 9,769 are continuing youth having completed a prior Youth Interview. The other 2,045 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 3101, 3102, 3201 and 3202 (DS3101, DS3102, DS3201, and DS3202) are data files comprising the weight variables for Wave 3. The weight variables for Wave 1 and Wave 2 are included in the main data files. However, in Wave 3, the weight variables have been separated into individual data files for Adult and Youth Questionnaires. The "all-waves" weight files contain weights for those respondents who have completed an interview during all three waves of data collection. The "single-wave" weight files contain weights for all respondents in Wave 3 regardless of their participation in previous waves.
Dataset 3503 (DS3503) contains data derived from responses to Wave 1-3 questionnaires indicating if participants had ever/never used various tobacco products as of the Wave 3 study period. This data file contains 25 variables for all 53,178 study participants as of Wave 3. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 4001 (DS4001) contains the data from the Wave 4 Adult Questionnaire. This data file contains 2,182 variables and 33,822 cases. Of these cases, 25,857 are continuing adults having completed a prior Adult Questionnaire, 1,900 are "aged-up adults" having previously completed a Youth Questionnaire, and 6,065 are "replenishment sample adults" (also known as "new cohort adults" in the annotated instrument).
Dataset 4002 (DS4002) contains the data from the Wave 4 Youth and Parent Questionnaire. This data file contains 1,389 variables and 14,798 cases. Of these cases, 9,365 are continuing youth having completed a prior Youth Interview, 1,694 cases are "aged-up youth" having previously been sampled as "shadow youth," and 3,739 are "replenishment sample youth" (also known as "new cohort youth" in the annotated instrument).
Datasets 4111, 4112, 4211, 4212, 4321, and 4322 (DS4111, DS4112, DS4211, DS4212, DS4321, and DS4322) are data files comprising the weight variables for Wave 4. In Wave 4, the weight variables have been separated into individual data files corresponding to the Wave 1 and Wave 4 Cohorts and different weight types. The "all-waves" weight files contain weights for those Wave 1 Cohort respondents who completed an interview for all waves in which they were old enough or verified their information for waves in which they were not old enough to be interviewed. The "single-wave" weight files contain weights for Wave 1 Cohort respondents at Wave 4 who completed an interview at Wave 1, regardless of their participation in previous waves. The "cross-sectional" weight files contain weights for all respondents in the Wave 4 Cohort.
Dataset 4503 (DS4503) contains data derived from responses to Wave 1-4 questionnaires indicating if participants had ever/never used various tobacco products as of the Wave 4 data collection period. This data file contains 27 variables for all 67,276 study participants as of the Wave 4 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 5001 (DS5001) contains the data from the Wave 5 Adult Questionnaire. This data file contains 2,315 variables and 34,309 cases. Of these cases, 29,876 are continuing adults having completed a prior Adult Questionnaire, 4,433 are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 5002 (DS5002) contains the data from the Wave 5 Youth and Parent Questionnaire. This data file contains 1,530 variables and 12,098 cases. Of these cases, 10,446 are continuing youth having completed a prior Youth Interview, 1,652 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 5111, 5112, 5211, 5212, 5221, 5222, 5711, 5712, 5721, and 5722 (DS5111, DS5112, DS5211, DS5212, DS5221, DS5222, DS5711, DS5712, DS5721, and DS5722) are data files comprising the weight variables for Wave 5. In Wave 5, the weight variables are in individual data files corresponding to the Wave 1 and Wave 4 Cohorts and different weight types. The "all-waves" weight files contain weights for those Wave 1 Cohort participants who completed a Wave 5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, and 4.
Dataset 5503 (DS5503) contains data derived from responses to Wave 1-5 (including Wave 4.5) questionnaires indicating if participants had ever/never used various tobacco products as of the Wave 5 data collection period. This data file contains 26 variables for all 67,276 study participants as of the Wave 5 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
There are two separate sets of files with "single wave" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "single-wave" weight files for the Wave 1 Cohort contain weights for participants who completed an interview in Wave 1 and in Wave 5, regardless of their participation in the intervening waves. The "single-wave" weight files for the Wave 4 Cohort contain weights for all Wave 5 interview respondents in the Wave 4 Cohort.
There are also two separate sets of files with "special collection all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "special collection all-waves" weight files for the Wave 1 Cohort contains weights for participants who completed a Wave 5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, and the special collection in Wave 4.5. The "special collection all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Wave 4 and the special collection in Wave 4.5.
Dataset 6001 (DS6001) contains the data from the Wave 6 Adult Questionnaire. This data file contains 2,589 variables and 30,516 cases. Of these cases, 28,852 are continuing adults having completed a prior Adult Questionnaire and 1,664 are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 6002 (DS6002) contains the data from the Wave 6 Youth and Parent Questionnaire. This data file contains 1,822 variables and 5,652 cases. Of these cases, 5,622 are continuing youth having completed a prior Youth interview and 30 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 6111, 6112, 6121, 6122, 6211, 6212, 6221, 6222, 6711, 6712, 6721, and 6722 (DS6111, DS6112, DS6121, DS6122, DS6211, DS6212, DS6221, DS6222, DS6711, DS6712, DS6721, and DS6722) are data files comprising the weight variables for Wave 6. In Wave 6, the weight variables are in individual data files corresponding to the Wave 1 and Wave 4 Cohorts and different weight types. There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 6 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, and 5. The "all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 6 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4 and 5.
There are two separate sets of files with "single-wave" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "single-wave" weight files for the Wave 1 Cohort contain weights for participants who completed an interview in Wave 1 and in Wave 6, regardless of their participation in the intervening waves. The "single-wave" weight files for the Wave 4 Cohort contain weights for participants who completed an interview in Wave 4 and in Wave 6, regardless of their participation in the intervening waves.
There are also two separate sets of files with "special collection all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "special collection all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 6 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, and the special collections in Wave 4.5, and Wave 5.5 or PATH-ATS. The "special collection all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 6 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4 and 5, and the special collections in Wave 4.5, and Wave 5.5 or PATH-ATS.
Dataset 6503 (DS6503) contains data derived from responses to Wave 1-6 (including Wave 4.5, Wave 5.5, and PATH-ATS) questionnaires indicating if participants had ever/never used various tobacco products as of the Wave 6 data collection period. This data file contains 24 variables for all 67,276 study participants as of the Wave 6 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 7001 (DS7001) contains the data from the Wave 7 Adult Questionnaire. This data file contains 2,813 variables and 30,801 cases. Of these cases, 27,258 are continuing adults having completed a prior Adult Questionnaire, 1,740 are "aged-up adults" having previously completed a Youth Questionnaire, and 1,803 are "replenishment sample adults" (also known as "new cohort adults" in the annotated instrument).
Dataset 7002 (DS7002) contains the data from the Wave 7 Youth and Parent Questionnaire. This data file contains 1,897 variables and 10,834 cases. Of these cases, 3,512 are continuing youth having completed a prior Youth Interview, 1 case is an "aged-up youth" having previously been sampled as "shadow youth," and 7,321 are "replenishment sample youth" (also known as "new cohort youth" in the annotated instrument).
Datasets 7111, 7112, 7121, 7122, 7211, 7212, 7221, 7222, 7331, 7332, 7711, 7712, 7721, and 7722 (DS DS7111, DS7112, DS7121, DS7122, DS7211, DS7212, DS7221, DS7222, DS7331, DS7332, DS7711, DS7712, DS7721, and DS7722) are data files comprising the weight variables for Wave 7. In Wave 7, the weight variables are in individual data files corresponding to the Wave 1, Wave 4, and Wave 7 Cohorts and different weight types.
There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 7 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, and 6. The "all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 7 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 5, and 6.
There are two separate sets of files with "single-wave" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "single-wave" weight files for the Wave 1 Cohort contain weights for participants who completed an interview in Wave 1 and in Wave 7, regardless of their participation in the intervening waves. The "single-wave" weight files for the Wave 4 Cohort contain weights for participants who completed an interview in Wave 4 and in Wave 7, regardless of their participation in the intervening waves.
There are also two separate sets of files with "special collection all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "special collection all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 7 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, 6, and the special collections in Wave 4.5, and Wave 5.5 or PATH-ATS. The "special collection all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 7 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 5, 6, and the special collections in Wave 4.5, and Wave 5.5 or PATH-ATS.
The "cross-sectional" weight files contain weights for all respondents in the Wave 7 Cohort.
Dataset 8001 (DS8001) contains data from the Wave 8 Adult Questionnaire. This data file contains 3,467 variables and 31,477 cases. Of these cases, 30,021 are continuing adults having completed a prior Adult Questionnaire and 1,456 are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 8002 (DS8002) contains data from the Wave 8 Youth and Parent Questionnaire. This data file contains 2,393 variables and 8,002 cases. Of these cases, 7,046 are continuing youth having completed a prior Youth Interview and 956 are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 8111, 8121, 8122, 8211, 8221, 8231, 8232, 8711, 8721, 8722, 8731, and 8732 (DS8111, DS8121, DS8122, DS8211, DS8221, DS8231, DS8232, DS8711, 8DS721, DS8722, DS8731, and DS8732) are data files comprising the weight variables for Wave 8. In Wave 8, the weight variables are in individual data files corresponding to the Wave 1, Wave 4, and Wave 7 Cohorts and different weight types.
There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, 6, and 7. Note that only adults have "all-waves" weights for the Wave 1 Cohort; youth from the Wave 1 Cohort aged-up to adults by the time of Wave 8. The "all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 5, 6, and 7.
There are three separate sets of files with "single-wave" weights: one for the Wave 1 Cohort, one for the Wave 4 Cohort, and one for the Wave 7 Cohort. The "single-wave" weight files for the Wave 1 Cohort contain weights for participants who completed an interview in Wave 1 and in Wave 8, regardless of their participation in the intervening waves. The "single-wave" weight files for the Wave 4 Cohort contain weights for participants who completed an interview in Wave 4 and in Wave 8, regardless of their participation in the intervening waves. Note that only adults have "single-wave" weights for the Wave 1 and Wave 4 Cohorts; youth from the Wave 1 Cohort aged-up to adults by the time of Wave 8 and youth from the Wave 4 Cohort were selected as shadow youth so they do not have any interview data from Wave 4. The "single wave" weights files for the Wave 7 Cohort contain weights for participants who completed an interview in Wave 7 and in Wave 8.
There are also three separate sets of files with "special collection all-waves" weights: one for the Wave 1 Cohort, one for the Wave 4 Cohort, and one for the Wave 7 Cohort. The "special collection all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, 6, 7 and the special collections in Wave 4.5, Wave 5.5, and Wave 7.5. Note that only adults have "special collection all-waves" weights for the Wave 1 Cohort; youth from the Wave 1 Cohort aged-up to adults by the time of Wave 8. The "special collection all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 5, 6, 7, and the special collections in Wave 4.5, Wave 5.5, and Wave 7.5. The "special collection all-waves" weight files for the Wave 7 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Wave 7 and the special collection in Wave 7.5.
Each case in an Adult data file represents a single, completed interview. Each case in a Youth data file represents one youth and his or her parent's responses about that youth. Parents who provided permission for their child to participate in a Youth Interview were asked to complete a brief interview about their child. In all waves of data collection, less than 0.5 percent of the parents did not complete an interview. Most questions are asked about the child.
When multiple youth from the same household were selected to be in the study, the parent(s) completed separate interviews about each youth. If one parent completed two or more interviews, that parent only answered questions about himself/herself once. Those questions were then skipped in the subsequent interview(s) for the other child(ren) and the responses duplicated in that child(ren)'s data file(s).
Population Assessment of Tobacco and Health (PATH) Study [United States] Restricted-Use Files (ICPSR 36231)
The PATH Study was launched in 2011 to inform the Food and Drug Administration's regulatory activities under the Family Smoking Prevention and Tobacco Control Act (TCA). The PATH Study is a collaboration between the National Institute on Drug Abuse (NIDA), National Institutes of Health (NIH), and the Center for Tobacco Products (CTP), Food and Drug Administration (FDA). The study sampled over 150,000 mailing addresses across the United States to create a national sample of people who use or do not use tobacco.
45,971 adults and youth constitute the first (baseline) wave, Wave 1, of data collected by this longitudinal cohort study. These 45,971 adults and youth along with 7,207 "shadow youth" (youth ages 9 to 11 sampled at Wave 1) make up the 53,178 participants that constitute the Wave 1 Cohort. Respondents are asked to complete an interview at each follow-up wave. Youth who turn 18 by the current wave of data collection are considered "aged-up adults" and are invited to complete the Adult Interview. Additionally, "shadow youth" are considered "aged-up youth" upon turning 12 years old, when they are asked to complete an interview after parental consent.
At Wave 4, a probability sample of 14,098 adults, youth, and shadow youth ages 10 to 11 was selected from the civilian, noninstitutionalized population (CNP) at the time of Wave 4. This sample was recruited from residential addresses not selected for Wave 1 in the same sampled Primary Sampling Units (PSUs) and segments using similar within-household sampling procedures. This "replenishment sample" was combined for estimation and analysis purposes with Wave 4 adult and youth respondents from the Wave 1 Cohort who were in the CNP at the time of Wave 4. This combined set of Wave 4 participants, 52,731 participants in total, forms the Wave 4 Cohort.
At Wave 7, a probability sample of 14,863 adults, youth, and shadow youth ages 9 to 11 was selected from the CNP at the time of Wave 7. This sample was recruited from residential addresses not selected for Wave 1 or Wave 4 in the same sampled PSUs and segments using similar within-household sampling procedures. This "second replenishment sample" was combined for estimation and analysis purposes with the Wave 7 adult and youth respondents from the Wave 4 Cohort who were at least age 15 and in the CNP at the time of Wave 7. This combined set of Wave 7 participants, 46,169 participants in total, forms the Wave 7 Cohort.
Please refer to the Restricted-Use Files User Guide that provides further details about children designated as "shadow youth" and the formation of the Wave 1, Wave 4, and Wave 7 Cohorts.
Dataset 0002 (DS0002) contains the data from the State Design Data. This file contains 7 variables and 82,139 cases. The state identifier in the State Design file reflects the participant's state of residence at the time of selection and recruitment for the PATH Study.
Dataset 1011 (DS1011) contains the data from the Wave 1 Adult Questionnaire. This data file contains 2,021 variables and 32,320 cases. Each of the cases represents a single, completed interview.
Dataset 1012 (DS1012) contains the data from the Wave 1 Youth and Parent Questionnaire. This file contains 1,431 variables and 13,651 cases.
Dataset 1411 (DS1411) contains the Wave 1 State Identifier data for Adults and has 5 variables and 32,320 cases. Dataset 1412 (DS1412) contains the Wave 1 State Identifier data for Youth (and Parents) and has 5 variables and 13,651 cases. The same 5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state Federal Information Processing System (FIPS), state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 1, which is also their state of residence at the time of recruitment.
Dataset 1611 (DS1611) contains the Tobacco Universal Product Code (UPC) data from Wave 1. This data file contains 32 variables and 8,601 cases. This file contains UPC values on the packages of tobacco products used or in the possession of adult respondents at the time of Wave 1. The UPC values can be used to identify and validate the specific products used by respondents and augment the analyses of the characteristics of tobacco products used by these respondents at the time of Wave 1.
Dataset 1801 (DS1801) contains Location Characteristics for Wave 1 Adults. This data file contains 4 variables and 32,320 cases.
Dataset 1802 (DS1802) contains Location Characteristics for Wave 1 Youth. This data file contains 4 variables and 13,651 cases.
Dataset 1901 (DS1901) contains Study Research Derived Variables for Wave 1 Adults created by PATH Study analysts. This data file contains 104 variables and 32,320 cases.
Dataset 1902 (DS1902) contains Study Research Derived Variables for Wave 1 Youth created by PATH Study analysts. This data file contains 89 variables and 13,651 cases.
Dataset 2011 (DS2011) contains the data from the Wave 2 Adult Questionnaire. This data file contains 2,421 variables and 28,362 cases. Of these cases, 26,447 also completed a Wave 1 Adult Questionnaire. The other 1,915 cases are "aged-up adults" having previously completed a Wave 1 Youth Questionnaire.
Dataset 2012 (DS2012) contains the data from the Wave 2 Youth and Parent Questionnaire. This data file contains 1,596 variables and 12,172 cases. Of these cases, 10,081 also completed a Wave 1 Youth Questionnaire. The other 2,091 cases are "aged-up youth" having previously been sampled as "shadow youth."
Dataset 2411 (DS2411) contains the Wave 2 State Identifier data for Adults and has 5 variables and 28,362 cases. Dataset 2412 (DS2412) contains the Wave 2 State Identifier data for Youth and Parents and has 5 variables and 12,172 cases. The same 5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 2.
Dataset 2611 (DS2611) contains the Tobacco Universal Product Code (UPC) data from Wave 2. This data file contains 32 variables and 7,295 cases. This file contains UPC values on the packages of tobacco products used or in the possession of adult respondents at the time of Wave 2. The UPC values can be used to identify and validate the specific products used by respondents and augment the analyses of the characteristics of tobacco products used by these respondents at the time of Wave 2.
Dataset 2801 (DS2801) contains Location Characteristics for Wave 2 Adults. This data file contains 4 variables and 28,362 cases.
Dataset 2802 (DS2802) contains Location Characteristics for Wave 2 Youth. This data file contains 4 variables and 12,172 cases.
Dataset 2901 (DS2901) contains Study Research Derived Variables for Wave 2 Adults created by PATH Study analysts. This data file contains 178 variables and 28,362 cases.
Dataset 2902 (DS2902) contains Study Research Derived Variables for Wave 2 Youth created by PATH Study analysts. This data file contains 123 variables and 12,172 cases.
Dataset 3011 (DS3011) contains the data from the Wave 3 Adult Questionnaire. This data file contains 2,359 variables and 28,148 cases. Of these cases, 26,241 are continuing adults having completed a prior Adult Questionnaire. The other 1,907 cases are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 3012 (DS3012) contains the data from the Wave 3 Youth and Parent Questionnaire. This data file contains 1,492 variables and 11,814 cases. Of these cases, 9,769 are continuing youth having completed a prior Youth Interview. The other 2,045 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 3111, 3211, 3112, and 3212 (DS3111, DS3211, DS3112, and DS3212) are data files comprising the weight variables for Wave 3. The weight variables for Wave 1 and Wave 2 are included in the main data files. However, starting with Wave 3, the weight variables have been separated into individual data files. The "all-waves" weight files contain weights for respondents who completed an interview for all waves in which they were old enough to do so or verified their information with the study for waves in which they were not old enough to be interviewed. The "single-wave" weight files contain weights for all respondents in Wave 3 regardless of their participation in previous waves.
Dataset 3503 (DS3503) contains data derived from responses to Wave 1-3 questionnaires indicating if participants had ever/never used various tobacco products as of the Wave 3 study period. This data file contains 25 variables for all 53,178 study participants as of Wave 3. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 3411 (DS3411) contains the Wave 3 State Identifier data for Adults and has 5 variables and 28,148 cases. Dataset 3412 (DS3412) contains the Wave 3 State Identifier data for Youth and Parents and has 5 variables and 11,814 cases. The same 5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 3.
Dataset 3611 (DS3611) contains the Tobacco Universal Product Code (UPC) data from Wave 3. This data file contains 32 variables and 6,768 cases. This file contains UPC values on the packages of tobacco products used or in the possession of adult respondents at the time of Wave 3. The UPC values can be used to identify and validate the specific products used by respondents and augment the analyses of the characteristics of tobacco products used by these respondents at the time of Wave 3.
Dataset 3801 (DS3801) contains Location Characteristics for Wave 3 Adults. This data file contains 4 variables and 28,148 cases.
Dataset 3802 (DS3802) contains Location Characteristics for Wave 3 Youth. This data file contains 4 variables and 11,814 cases.
Dataset 3901 (DS3901) contains Study Research Derived Variables for Wave 3 Adults created by PATH Study analysts. This data file contains 107 variables and 28,148 cases.
Dataset 3902 (DS3902) contains Study Research Derived Variables for Wave 3 Youth created by PATH Study analysts. This data file contains 88 variables and 11,814 cases.
Dataset 4001 (DS4001) contains the data from the Wave 4 Adult Questionnaire. This data file contains 2,504 variables and 33,822 cases. Of these cases, 25,857 are continuing adults having completed a prior Adult Questionnaire, 1,900 are "aged-up adults" having previously completed a Youth Questionnaire, and 6,065 are "replenishment sample adults" (also known as "new cohort adults" in the annotated instrument).
Dataset 4002 (DS4002) contains the data from the Wave 4 Youth and Parent Questionnaire. This data file contains 1,600 variables and 14,798 cases. Of these cases, 9,365 are continuing youth having completed a prior Youth Interview, 1,694 cases are "aged-up youth" having previously been sampled as "shadow youth," and 3,739 are "replenishment sample youth" (also known as "new cohort youth" in the annotated instrument).
Datasets 4111, 4211, 4321, 4112, 4212, and 4322 (DS4111, DS4211, DS4321, DS4112, DS4212, and DS4322) are data files comprising the weight variables for Wave 4. In Wave 4, the weight variables have been separated into individual data files corresponding to the Wave 1 and Wave 4 Cohorts and different weight types. The "all-waves" weight files contain weights for those Wave 1 Cohort respondents who completed an interview for all waves in which they were old enough or verified their information for waves in which they were not old enough to be interviewed. The "single-wave" weight files contain weights for Wave 1 Cohort respondents at Wave 4 who completed an interview at Wave 1, regardless of their participation in previous waves. The "cross-sectional" weight files contain weights for all respondents in the Wave 4 Cohort.
Dataset 4401 (DS4401) contains the Wave 4 State Identifier data for Adults and has 5 variables and 33,822 cases. Dataset 4402 (DS4402) contains the Wave 4 State Identifier data for Youth and Parents and has 5 variables and 14,798 cases. The same 5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 4. For adults and youth from the replenishment sample, the values also represent state of residence at the time of recruitment.
Dataset 4503 (DS4503) contains data derived from responses to Wave 1-4 questionnaires, indicating if participants had ever/never used various tobacco products as of the Wave 4 data collection period. This data file contains 27 variables for all 67,276 study participants as of the Wave 4 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 4601 (DS4601) contains the Tobacco Universal Product Code (UPC) data from Wave 4. This data file contains 32 variables and 7,684 cases. This file contains UPC values on the packages of tobacco products used or in the possession of adult respondents at the time of Wave 4. The UPC values can be used to identify and validate the specific products used by respondents and augment the analyses of the characteristics of tobacco products used by these respondents at the time of Wave 4.
Dataset 4801 (DS4801) contains Location Characteristics for Wave 4 Adults. This data file contains 4 variables and 33,822 cases.
Dataset 4802 (DS4802) contains Location Characteristics for Wave 4 Youth. This data file contains 4 variables and 14,798 cases.
Dataset 5001 (DS5001) contains the data from the Wave 5 Adult Questionnaire. This data file contains 2,606 variables and 34,309 cases. Of these cases, 29,876 are continuing adults having completed a prior Adult Questionnaire and 4,433 are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 5002 (DS5002) contains the data from the Wave 5 Youth and Parent Questionnaire. This data file contains 1,776 variables and 12,098 cases. Of these cases, 10,446 are continuing youth having completed a prior Youth Interview and 1,652 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 5111, 5112, 5211, 5212, 5221, 5222, 5711, 5712, 5721, and 5722 (DS5111, DS5112, DS5211, DS5212, DS5221, DS5222, DS5711, DS5712, DS5721, and DS5722) are data files comprising the weight variables for Wave 5. In Wave 5, the weight variables are in individual data files corresponding to the Wave 1 and Wave 4 Cohorts and different weight types. The "all-waves" weight files contain weights for those Wave 1 Cohort participants who completed a Wave 5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, and 4.
There are two separate sets of files with "single wave" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "single-wave" weight files for the Wave 1 Cohort contain weights for participants who completed an interview in Wave 1 and in Wave 5, regardless of their participation in the intervening waves. The "single-wave" weight files for the Wave 4 Cohort contain weights for all Wave 5 interview respondents in the Wave 4 Cohort.
There are also two separate sets of files with "special collection all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "special collection all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, and the special collection in Wave 4.5. The "special collection all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Wave 4 and the special collection in Wave 4.5.
Dataset 5401 (DS5401) contains the Wave 5 State Identifier data for Adults and has 5 variables and 34,309 cases. Dataset 5402 (DS5402) contains the Wave 5 State Identifier data for Youth and Parents and has 5 variables and 12,098 cases. The same 5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 5.
Dataset 5503 (DS5503) contains data derived from responses to Wave 1-5 (including Wave 4.5) questionnaires indicating if participants had ever/never used various tobacco products as of the Wave 5 data collection period. This data file contains 26 variables for all 67,276 study participants as of the Wave 5 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 5601 (DS5601) contains the Tobacco Universal Product Code (UPC) data from Wave 5. This data file contains 33 variables and 6,678 cases. This file contains UPC values on the packages of tobacco products used or in the possession of adult respondents at the time of Wave 5. The UPC values can be used to identify and validate the specific products used by respondents and augment the analyses of the characteristics of tobacco products used by these respondents at the time of Wave 5.
Dataset 5801 (DS5801) contains Location Characteristics for Wave 5 Adults. This data file contains 4 variables and 34,309 cases.
Dataset 5802 (DS5802) contains Location Characteristics for Wave 5 Youth. This data file contains 4 variables and 12,098 cases.
Dataset 6001 (DS6001) contains the data from the Wave 6 Adult Questionnaire. This data file contains 2,935 variables and 30,516 cases
Of these cases, 28,852 are continuing adults having completed a prior Adult Questionnaire and 1,664 are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 6002 (DS6002) contains the data from the Wave 6 Youth and Parent Questionnaire. This data file contains 2,080 variables and 5,652 cases. Of these cases, 5,622 are continuing youth having completed a prior Youth Interview and 60 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 6111, 6112, 6121, 6122, 6211, 6212, 6221, 6222, 6711, 6712, 6721, and 6722 (DS6111, DS6112, DS6121, DS6122, DS6211, DS6212, DS62221, DS6222, DS6711, DS6712, DS6721, and DS6722) are data files comprising the weight variables for Wave 6. In Wave 6, the weight variables are in individual data files corresponding to the Wave 1 and Wave 4 Cohorts and different weight types. There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 6 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, and 5. The "all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 6 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4 and 5.
There are two separate sets of files with "single-wave" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "single-wave" weight files for the Wave 1 Cohort contain weights for participants who completed an interview in Wave 1 and in Wave 6, regardless of their participation in the intervening waves. The "single-wave" weight files for the Wave 4 Cohort contain weights for participants who completed an interview in Wave 4 and in Wave 6, regardless of their participation in the intervening waves.
There are also two separate sets of files with "special collection all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "special collection all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 6 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, and the special collections in Wave 4.5, and Wave 5.5 or PATH-ATS. The "special collection all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 6 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4 and 5, and the special collections in Wave 4.5, and Wave 5.5 or PATH-ATS.
Dataset 6401 (DS6401) contains the Wave 6 State Identifier data for Adults and has 5 variables and 30,516 cases. Dataset 6402 (DS6402) contains the Wave 6 State Identifier data for Youth and Parents and has 5 variables and 5,652 cases. The same 5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 6.
Dataset 6503 (DS6503) contains data derived from responses to questionnaires in Waves 1-6 (including the special collections in Wave 4.5, Wave 5.5, and PATH-ATS) indicating if participants had ever/never used various tobacco products as of the Wave 6 data collection period. This data file contains 24 variables for all 67,276 study participants as of the Wave 6 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 6601 (DS6601) contains the Tobacco Universal Product Code (UPC) data from Wave 6. This data file contains 53 variables and 5,408 cases. This file contains UPC values on the packages of tobacco products used or in the possession of adult respondents at the time of Wave 6. The UPC values can be used to identify and validate the specific products used by respondents and augment the analyses of the characteristics of tobacco products used by these respondents at the time of Wave 6.
Dataset 6801 (DS6801) contains Location Characteristics for Wave 6 Adults. This data file contains 4 variables and 30,516 cases.
Dataset 6802 (DS6802) contains Location Characteristics for Wave 6 Youth. This data file contains 4 variables and 5,652 cases.
Dataset 7001 (DS7001) contains the data from the Wave 7 Adult Questionnaire. This data file contains 3,221 variables and 30,801 cases. Of these cases, 27,258 are continuing adults having completed a prior Adult Questionnaire, 1,740 are "aged-up adults" having previously completed a Youth Questionnaire, and 1,803 are "replenishment sample adults" (also known as "new cohort adults" in the annotated instrument).
Dataset 7002 (DS7002) contains the data from the Wave 7 Youth and Parent Questionnaire. This data file contains 2,171 variables and 10,834 cases. Of these cases, 3,512 are continuing youth having completed a prior Youth Interview, 1 case is an "aged-up youth" having previously been sampled as "shadow youth," and 7,321 are "replenishment sample youth" (also known as "new cohort youth" in the annotated instrument).
Datasets 7111, 7112, 7121, 7122, 7211, 7212, 7221, 7222, 7331, 7332, 7711, 7712, 7721, and 7722 (DS DS7111, DS7112, DS7121, DS7122, DS7211, DS7212, DS7221, DS7222, DS7331, DS7332, DS7711, DS7712, DS7721, and DS7722) are data files comprising the weight variables for Wave 7. In Wave 7, the weight variables are in individual data files corresponding to the Wave 1, Wave 4, and Wave 7 Cohorts and different weight types.
There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 7 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, and 6. The "all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 7 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 5, and 6.
There are two separate sets of files with "single-wave" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "single-wave" weight files for the Wave 1 Cohort contain weights for participants who completed an interview in Wave 1 and in Wave 7, regardless of their participation in the intervening waves. The "single-wave" weight files for the Wave 4 Cohort contain weights for participants who completed an interview in Wave 4 and in Wave 7, regardless of their participation in the intervening waves.
There are also two separate sets of files with "special collection all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "special collection all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 7 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, 6, and the special collections in Wave 4.5, and Wave 5.5 or PATH-ATS. The "special collection all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 7 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 5, 6, and the special collections in Wave 4.5, and Wave 5.5 or PATH-ATS.
The "cross-sectional" weight files contain weights for all respondents in the Wave 7 Cohort.
Dataset 7401 (DS7401) contains the Wave 7 State Identifier data for Adults and has 5 variables and 30,801 cases. Dataset 7402 (DS7402) contains the Wave 7 State Identifier data for Youth and Parents and has 5 variables and 10,834 cases. The same 5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 7.
Dataset 7503 (DS7503) contains data derived from responses to questionnaires in Waves 1-7 (including the special collections in Wave 4.5, Wave 5.5, and PATH-ATS) indicating if participants had ever/never used various tobacco products as of the Wave 7 data collection period. This data file contains 26 variables for all 82,139 study participants as of the Wave 7 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 7601 (DS7601) contains the Tobacco Universal Product Code (UPC) data from Wave 7. This data file contains 53 variables and 4,533 cases. This file contains UPC values on the packages of tobacco products used or in the possession of adult respondents at the time of Wave 7. The UPC values can be used to identify and validate the specific products used by respondents and augment the analyses of the characteristics of tobacco products used by these respondents at the time of Wave 7.
Dataset 7801 (DS7801) contains Location Characteristics for Wave 7 Adults. This data file contains 4 variables and 30,801 cases.
Dataset 7802 (DS7802) contains Location Characteristics for Wave 7 Youth. This data file contains 4 variables and 10,834 cases.
Dataset 8001 (DS8001) contains the data from the Wave 8 Adult Questionnaire. This data file contains 3,467 variables and 31,477 cases. Of these cases, 30,021 are continuing adults having completed a prior Adult Questionnaire and 1,456 are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 8002 (DS8002) contains the data from the Wave 8 Youth and Parent Questionnaire. This data file contains 2,393 variables and 8,002 cases. Of these cases, 7,046 are continuing youth having completed a prior Youth Interview and 956 are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 8111, 8121, 8122, 8211, 8221, 8231, 8232, 8711, 8721, 8722, 8731, and 8732 (DS8111, DS8121, DS8122, DS8211, DS8221, DS8231, DS8232, DS8711, 8DS721, DS8722, DS8731, and DS8732) are data files comprising the weight variables for Wave 8. In Wave 8, the weight variables are in individual data files corresponding to the Wave 1, Wave 4, and Wave 7 Cohorts and different weight types.
There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, 6, and 7. Note that only adults have "all-waves" weights for the Wave 1 Cohort; youth from the Wave 1 Cohort aged-up to adults by the time of Wave 8. The "all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 5, 6, and 7.
There are three separate sets of files with "single-wave" weights: one for the Wave 1 Cohort, one for the Wave 4 Cohort, and one for the Wave 7 Cohort. The "single-wave" weight files for the Wave 1 Cohort contain weights for participants who completed an interview in Wave 1 and in Wave 8, regardless of their participation in the intervening waves. The "single-wave" weight files for the Wave 4 Cohort contain weights for participants who completed an interview in Wave 4 and in Wave 8, regardless of their participation in the intervening waves. Note that only adults have "single-wave" weights for the Wave 1 and Wave 4 Cohorts; youth from the Wave 1 Cohort aged-up to adults by the time of Wave 8 and youth from the Wave 4 Cohort were selected as shadow youth so they do not have any interview data from Wave 4. The "single wave" weights files for the Wave 7 Cohort contain weights for participants who completed an interview in Wave 7 and in Wave 8.
There are also three separate sets of files with "special collection all-waves" weights: one for the Wave 1 Cohort, one for the Wave 4 Cohort, and one for the Wave 7 Cohort. The "special collection all-waves" weight files for the Wave 1 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 5, 6, 7 and the special collections in Wave 4.5, Wave 5.5, and Wave 7.5. Note that only adults have "special collection all-waves" weights for the Wave 1 Cohort; youth from the Wave 1 Cohort aged-up to adults by the time of Wave 8. The "special collection all-waves" weight files for the Wave 4 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 5, 6, 7, and the special collections in Wave 4.5, Wave 5.5, and Wave 7.5. The "special collection all-waves" weight files for the Wave 7 Cohort contain weights for participants who completed a Wave 8 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Wave 7 and the special collection in Wave 7.5.
Dataset 8401 (DS8401) contains the Wave 8 State Identifier data for Adults and has 5 variables and 31,477 cases. Dataset 8402 (DS8402) contains the Wave 8 State Identifier data for Youth and Parents and has 5 variables and 8,002 cases. The same 5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 8.
Dataset 8801 (DS8801) contains Location Characteristics for Wave 8 Adults. This data file contains 4 variables and 31,477 cases.
Dataset 8802 (DS8802) contains Location Characteristics for Wave 8 Youth. This data file contains 4 variables and 8,002 cases.
Each case in an Adult data file represents a single, completed interview. Each case in a Youth data file represents one youth and his or her parent's responses about that youth. Parents who provided permission for their child to participate in a Youth Interview were asked to complete a brief interview about their child. In all waves of data collection, less than 0.5 percent of the parents did not complete an interview. Most questions are asked about the child.
When multiple youth from the same household were selected to be in the study, the parent(s) completed separate interviews about each youth. If one parent completed two or more interviews, that parent only answered questions about himself/herself once. Those questions were then skipped in the subsequent interview(s) for the other child(ren) and the responses duplicated in that child(ren)'s data file(s).
Population Assessment of Tobacco and Health (PATH) Study [United States] Biomarker Restricted-Use Files (ICPSR 36840)
The Population Assessment of Tobacco and Health (PATH) Study is a collaboration between the National Institute on Drug Abuse (NIDA), National Institutes of Health (NIH), and the Center for Tobacco Products (CTP), Food and Drug Administration (FDA). The study was launched in 2011 to inform the FDA's tobacco regulatory activities under the Family Smoking Prevention and Tobacco Control Act (TCA). For Wave 1 (baseline), the PATH Study sampled over 150,000 mailing addresses across the United States to create a national sample of people who use or do not use tobacco, yielding interviews with 45,971 adult and youth respondents.
45,971 adults and youth constitute the first (baseline) wave, Wave 1, of data collected by this longitudinal cohort study. These 45,971 adults and youth along with 7,207 "shadow youth" (youth ages 9 to 11 sampled at Wave 1) make up the 53,178 participants that constitute the Wave 1 Cohort. Respondents are asked to complete an interview at each follow-up wave. Youth who turn 18 by the current wave of data collection are considered "aged-up adults" and are invited to complete the Adult Interview. Additionally, "shadow youth" are considered "aged-up youth" upon turning 12 years old, when they are asked to complete an interview after parental consent.
At Wave 4, a probability sample of 14,098 adults, youth, and shadow youth ages 10 to 11 was selected from the civilian, noninstitutionalized population at the time of Wave 4. This sample was recruited from residential addresses not selected for Wave 1 in the same sampled PSUs and segments using similar within-household sampling procedures. This "replenishment sample" was combined for estimation and analysis purposes with Wave 4 adult and youth respondents from the Wave 1 Cohort who were in the civilian, noninstitutionalized population at the time of Wave 4. This combined set of Wave 4 participants, 52,731 participants in total, forms the Wave 4 Cohort.
At Wave 7, a probability sample of 14,863 adults, youth, and shadow youth ages 9 to 11 was selected from the civilian, noninstitutionalized population at the time of Wave 7. This sample was recruited from residential addresses not selected for Wave 1 or Wave 4 in the same sampled PSUs and segments using similar within-household sampling procedures. This second replenishment sample was combined for estimation and analysis purposes with Wave 7 adult and youth respondents from the Wave 4 Cohort who were at least age 15 and in the civilian, noninstitutionalized population at the time of Wave 7. This combined set of Wave 7 participants, 46,169 participants in total, forms the Wave 7 Cohort
Please refer to the Restricted-Use Files User Guide that provides further details about children designated as "shadow youth" and the formation of the Wave 1, Wave 4, and Wave 7 Cohorts.
Biospecimen Collection
Each adult respondent, who completed the interview at Wave 1, was asked to provide at least two biospecimens. Providing biospecimens was voluntary and was not a condition of participation. Respondents were asked to report their use of all nicotine-containing products during the 3-day period prior to the time of any biospecimen collection (Nicotine Exposure Questions (NEQs)) to facilitate interpretation of biomarker results.
Of the 32,320 respondents who completed the Adult Interview at Wave 1, 21,801 (67.4 percent) provided a urine specimen and 14,520 (44.9 percent) provided a blood specimen. For the purposes of subsampling adults into the Wave 1 Biomarker Core, adult participants were grouped by tobacco product use at Wave 1 into nine mutually exclusive groups.
A sample of 11,522 adults who provided sufficient urine for the planned analyses were selected from the first six tobacco product use groups (see section 3.1 of the Biomarker Restricted-Use Files User Guide) representing people who never used tobacco, currently use tobacco, and formerly used tobacco (within the last 12 months). This group constitutes the original Wave 1 Biomarker Core. Of the 11,522 adults, 7,159 also provided a blood specimen. All urine and blood specimens provided by the Wave 1 Biomarker Core were sent for laboratory analysis.
Subsequent to this selection, an additional stratified probability sample of adults who completed the Wave 1 Adult Interview and provided a sufficient amount of urine for the planned analyses at Wave 1 (independent of whether they provided a blood specimen) was selected from the remaining three product use groups (see section 3.1 of the Biomarker Restricted-Use Files User Guide). Wave 1 blood and urine specimens from this expansion sample were also sent for laboratory analysis. The original and expansion samples together form the expanded Wave 1 Biomarker Core. The expansion sample did not provide urine specimens for laboratory analysis again until Wave 7.
Each youth who completed the Wave 4 interview was asked to provide a urine specimen. Each Wave 4 shadow youth (ages 10 and 11 at Wave 4) who completed the Wave 5 youth interview was also asked to provide a urine specimen. Providing this urine biospecimen was voluntary and was not a condition of participation.
Of the 14,798 respondents who completed the Youth Interview at Wave 4, 13,097 (88.5 percent) provided a urine specimen. A sample of 3,509 Wave 4 Cohort youth ages 12 to 17 who completed the Wave 4 Youth Interview and provided a sufficient amount of urine for the planned laboratory analyses was selected from a diverse mix of five tobacco product use and non-use groups. In addition, a sample of 528 Wave 4 shadow youth who completed a Wave 5 interview and provided a sufficient amount of urine for the planned laboratory analyses at Wave 5 was also selected. These 4,037 sampled youth and shadow youth constitute the Wave 4 Biomarker Core. All urine specimens provided by the Wave 4 Biomarker Core were sent for laboratory analysis.
As members of the Wave 1 and Wave 4 Biomarker Cores age over time, a new Wave 7 Biomarker Core was designed to provide nationally representative estimates for the U.S. civilian noninstitutionalized adult (ages 18 and older) population (CNP) at the time of Wave 7 (2022-2023). To that end, Aat the conclusion of Wave 7, a new biomarker core was selected from Wave 7 Cohort adults who completed an interview and provided a urine specimen at Wave 7. The Wave 7 Biomarker Core sample selection was a two-stage process. Prior to the start of data collection, a subsample of continuing participants expected to be adults at the time of their Wave 7 interview, including some participants who were part of the Wave 1 or Wave 4 Biomarker Cores, was selected and flagged for urine collection; additionally, a subsample of replenishment sample address was selected and flagged so that any Wave 7 Adult Interview respondents living at the selected addresses would be asked to provide a urine specimen. Of the 10,698 Adult Interview respondents from these subsamples, 9,187 (85.9 percent) provided a urine specimen. A sample of 7,750 Wave 7 Cohort adults who completed the Wave 7 Adult Interview and provided a sufficient amount of urine for the planned laboratory analyses was selected from six mutually exclusive and exhaustive tobacco use groups (see section 3.3 of the Biomarker Restricted-Use Files User Guide). All urine specimens provided by the Wave 7 Biomarker Core were sent for laboratory analysis.
Biomarker Restricted Use Files
Wave 1 Restricted-Use Biomarker Data Files (Biomarker RUF) consists of three different types of files for the Wave 1 Biomarker Core:
- 2 Collection and NEQ files for Urine (DS1001) and Blood (DS1101)
- 2 Biomarker Weight files including variables for use in variance estimation for Urine (DS1021) and Blood (DS1121). Both files are updated to include records for the expanded Wave 1 Biomarker Core.
- 8 Urine Panels (DS1031 to DS1038), 4 Serum Panels (DS1131 to DS1134) and 1 Plasma Panel (DS1231) containing biomarker assay results. 6 Urine Panels (DS1032, DS1033, DS1035, DS1036, DS1037, and DS1038) and 2 Serum Panels (DS1131 and DS1132) are updated to include records for the expanded Wave 1 Biomarker Core.
All files updated to include records for the expanded Wave 1 Biomarker Core contain an indicator R01_A_W1BC_TYPE (1 = Original, 2 = Expansion) to identify respondents in the Wave 1 Biomarker Core original and expansion subsamples.
For Wave 2, urine biospecimens were requested from the original Wave 1 Biomarker Core. Respondents were also asked to complete the NEQs prior to biospecimen collection.
The Wave 2 Biomarker RUF consists of three different types of files:
- 1 Collection and NEQ file for Urine (DS2001)
- 2 Biomarker Weight files including variables for use in variance estimation for Urine (DS2021) and F2PG2a (DS2022)
- 8 Urine Panels (DS2031 to DS2038) containing biomarker assay results.
For Wave 3, urine biospecimens were requested from the original Wave 1 Biomarker Core. Respondents were also asked to complete the NEQs prior to biospecimen collection.
The Wave 3 Biomarker RUF consists of three different types of files:
- 1 Collection and NEQ file for Urine (DS3001)
- 4 Biomarker Weight files including variables for use in variance estimation for Urine (DS3021 and DS3022) and F2PG2a (DS3023 and DS3024).
- 7 Urine Panels (DS3032 to DS3038) containing biomarker assay results.
For Wave 4, urine biospecimens were requested from the original Wave 1 Biomarker Core and all youth who completed the Wave 4 interview. Respondents were also asked to complete the NEQs prior to biospecimen collection.
The Wave 4 Biomarker RUF consists of the following files for each Biomarker Core:
Wave 1 Biomarker Core:
- 1 Collection and NEQ file for Urine (DS4001)
- 4 Biomarker Weight files including variables for use in variance estimation for Urine (DS4021 and DS4022) and F2PG2a (DS4023 and DS4024).
- 7 Urine Panels (DS4032, DS4033, DS4034, DS4035, DS4036, DS4037 and DS4038) containing biomarker assay results.
Wave 4 Biomarker Core:
- 1 Collection and NEQ file for Youth Urine (DS4011)
- 1 Biomarker Weight files including variables for use in variance estimation for Urine (DS4043)
- 7 Urine Panels (DS4051, DS4053, DS4054, DS4055, DS4056, DS4057 and DS4058) containing biomarker assay results.
For Wave 5, urine biospecimens were requested from the original Wave 1 Biomarker Core and the Wave 4 Biomarker Core. Respondents were also asked to complete the NEQs prior to biospecimen collection.
The Wave 5 Biomarker RUF consists of the following files for each Biomarker Core:
Wave 1 Biomarker Core:
- 1 Collection and NEQ file for Urine (DS5001)
- 4 Biomarker Weight files including variables for use in variance estimation for Urine (DS5021 and DS5022) and F2PG2a (DS5023 and DS5024)
- 6 Urine Panels (DS5032, DS5033, DS5035, DS5036, DS5037, and DS5038) containing biomarker assay results.
Wave 4 Biomarker Core:
- 1 Collection and NEQ file for Youth Urine (DS5011)
- 1 Collection and NEQ file for Adult Urine (DS5001)
- 1 Biomarker Weight file including variables for use in variance estimation for Urine (DS5042)
- 7 Urine Panels (DS5051, DS5053, DS5054, DS5055, DS5056, DS5057, and DS5058) containing biomarker assay results.
Note that the initial release of 3 Urine Panels and Biomarker weights for the Wave 4 Biomarker Core only included records for those among the 3,509 members who responded in Wave 5 and provided urine specimens in sufficient quantities for laboratory analyses. As of version 20, the Wave 5 biomarker data files and weights include data for all Wave 4 Biomarker Core members who provided urine specimens at Wave 5 in sufficient quantities for laboratory analyses, including the Wave 4 shadow youth who completed their first interviews at Wave 5. This means that records were added to previously released urine panel data files (DS5051, DS5053, and DS5056) and biomarker weights (DS5042) to include data for the Wave 4 shadow youth (N=528) who completed their first interviews at Wave 5. All panels released in version 20 and beyond will include records for the complete Wave 4 Biomarker Core.
Also note that the Collection and NEQ file for Adult Urine (DS5001) includes data for both the Wave 1 Biomarker Core and Wave 4 Biomarker Core.
For Wave 7, urine biospecimens were requested from the Wave 1 Biomarker Core, the Wave 4 Biomarker Core, and those in the subsample eligible for the Wave 7 biomarker Core. Respondents were also asked to complete the NEQs prior to biospecimen collection.
The Wave 7 Biomarker RUF consists of the following files for each Biomarker Core:
Wave 1 Biomarker Core:
- 1 Collection and NEQ file for Urine (DS7001)
- 4 Biomarker Weight files including variables for use in variance estimation for Urine (DS7021 and DS7022) and F2PG2a (DS7023 and DS7024)
- 6 Urine Panels (DS7032, DS7033, DS7035, DS7036, DS7037, and DS7038) containing biomarker assay results.
Wave 4 Biomarker Core:
- 1 Collection and NEQ file for Youth Urine (DS7011)
- 1 Collection and NEQ file for Adult Urine (DS7001)
- 2 Biomarker Weight files including variables for use in variance estimation for Urine (DS7041 and DS7042)
- 6 Urine Panels (DS7051, DS7053, DS7055, DS7056, DS7057, and DS7058) containing biomarker assay results.
Wave 7 Biomarker Core:
- 1 Collection and NEQ file for Urine (DS7001)
- 1 Biomarker Weight file including variables for use in variance estimation for Urine (DS7061)
- 6 Urine Panels (DS7072, DS7073, DS7075, DS7076, DS7077, and DS7078) containing biomarker assay results.
The Collection and NEQ file for Adult Urine (DS7001) includes data for the Wave 1 Biomarker Core, Wave 4 Biomarker Core, and Wave 7 Biomarker Core.
Please refer to the Biomarker Restricted-Use Files User Guide for additional information about the Biomarker Cores.
References to the collection of biospecimens will be specified by the collected specimen, i.e., urine and (whole) blood. However, references to biomarker analyses and analytes will be specified by the type of matrix (serum, plasma, or urine) used for the analysis.
Population Assessment of Tobacco and Health (PATH) Study [United States] Special Collection Public-Use Files (ICPSR 37786)
The PATH Study was launched in 2011 to inform the Food and Drug Administration's regulatory activities under the Family Smoking Prevention and Tobacco Control Act (TCA). The PATH Study is a collaboration between the National Institute on Drug Abuse (NIDA), National Institutes of Health (NIH), and the Center for Tobacco Products (CTP), Food and Drug Administration (FDA). The study sampled over 150,000 mailing addresses across the United States to create a national sample of people who do and do not use tobacco.
45,971 adults and youth constitute the first (baseline) wave, Wave 1, of data collected by this longitudinal cohort study. These 45,971 adults and youth along with 7,207 "shadow youth" (youth ages 9 to 11 sampled at Wave 1) make up the 53,178 participants that constitute the Wave 1 Cohort. Respondents are asked to complete an interview at each follow-up wave. Youth who turn 18 by the current wave of data collection are considered "aged-up adults" and are invited to complete the Adult Interview. Additionally, "shadow youth" are considered "aged-up youth" upon turning 12 years old, when they are asked to complete an interview after parental consent.
At Wave 4, a probability sample of 14,098 adults, youth, and shadow youth ages 10 to 11 was selected from the civilian, noninstitutionalized population (CNP) at the time of Wave 4. This sample was recruited from residential addresses not selected for Wave 1 in the same sampled Primary Sampling Units (PSUs) and segments using similar within-household sampling procedures. This "replenishment sample" was combined for estimation and analysis purposes with Wave 4 adult and youth respondents from the Wave 1 Cohort who were in the CNP at the time of Wave 4. This combined set of Wave 4 participants, 52,731 participants in total, forms the Wave 4 Cohort.
At Wave 7, a probability sample of 14,863 adults, youth, and shadow youth ages 9 to 11 was selected from the CNP at the time of Wave 7. This sample was recruited from residential addresses not selected for Wave 1 or Wave 4 in the same sampled PSUs and segments using similar within-household sampling procedures. This "second replenishment sample" was combined for estimation and analysis purposes with the Wave 7 adult and youth respondents from the Wave 4 Cohorts who were at least age 15 and in the CNP at the time of Wave 7. This combined set of Wave 7 participants, 46,169 participants in total, forms the Wave 7 Cohort.
Please refer to the Public-Use Files User Guide that provides further details about children designated as "shadow youth" and the formation of the Wave 1, Wave 4, and Wave 7 Cohorts.
Wave 4.5 was a special data collection for youth only who were aged 12 to 17 at the time of the Wave 4.5 interview. Wave 4.5 was the fourth annual follow-up wave for those who were members of the Wave 1 Cohort. For those who were sampled at Wave 4, Wave 4.5 was the first annual follow-up wave.
Wave 5.5, conducted in 2020, was a special data collection for Wave 4 Cohort youth and young adults ages 13 to 19 at the time of the Wave 5.5 interview. Also in 2020, a subsample of Wave 4 Cohort adults ages 20 and older were interviewed via the PATH Study Adult Telephone Survey (PATH-ATS).
Wave 7.5 was a special collection for Wave 4 and Wave 7 Cohort youth and young adults ages 12 to 22 at the time of the Wave 7.5 interview. For those who were sampled at Wave 7, Wave 7.5 was the first annual follow-up wave.
Dataset 1002 (DS1002) contains the data from the Wave 4.5 Youth and Parent Questionnaire. This file contains 1,395 variables and 13,131 cases. Of these cases, 11,378 are continuing youth having completed a prior Youth Interview. The other 1,753 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 1112, 1212, and 1222, (DS1112, DS1212, and DS1222) are data files comprising the weight variables for Wave 4.5. The "all-waves" weight file contains weights for participants in the Wave 1 Cohort who completed a Wave 4.5 Youth Interview and completed interviews (if old enough to do so) or verified their information with the study (if not old enough to be interviewed) in Waves 1, 2, 3, and 4.
There are two separate files with "single wave" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "single-wave" weight file for the Wave 1 Cohort contains weights for youth who completed an interview in Wave 1 and in Wave 4.5, regardless of their participation in the intervening waves. The "single-wave" weight file for the Wave 4 Cohort contains weights for all Wave 4.5 Youth Interview respondents in the Wave 4 Cohort.
Dataset 1503 (DS1503) contains data derived from responses to questionnaires in Wave 1, Wave 2, Wave 3, Wave 4, and Wave 4.5 indicating if participants had ever/never used various tobacco products as of the Wave 4.5 data collection period. This data file contains 26 variables for all 67,276 study participants as of the Wave 4.5 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 2001 (DS2001) contains the data from the Wave 5.5 Adult Questionnaire. This file contains 2,323 variables and 3,628 cases. Of these cases, 1,014 are continuing adults having completed a prior Adult Questionnaire. The other 2,614 cases are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 2002 (DS2002) contains the data from the Wave 5.5 Youth and Parent Questionnaire. This file contains 1,625 variables and 7,129 cases. Of these cases, 7,076 are continuing youth having completed a prior Youth Interview. The other 53 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 2111, 2112, 2121, 2122, 2221, and 2222 (DS2111, DS2112, DS2121, DS2122, DS2221, and DS2222) are data files comprising the weight variables for Wave 5.5. In Wave 5.5, the weight variables are in individual data files corresponding to the Wave 1 and Wave 4 Cohorts and different weight types.
There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight file for the Wave 1 Cohort contains weights for participants who completed a Wave 5.5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 4.5, and 5. The "all-waves" weight file for the Wave 4 Cohort contains weights for participants who completed a Wave 5.5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 4.5, and 5.
The "single-wave" weight file for the Wave 4 Cohort contains weights for all Wave 5.5 interview respondents.
Dataset 3001 (DS3001) contains the data from PATH-ATS. This file contains 908 variables and 8,874 cases, all of which are continuing adults having completed a prior Adult Questionnaire, with their most recent interview in Wave 5.
Datasets 3111 and 3121 (DS3111 and DS3121) are data files comprising weights for PATH-ATS. In PATH-ATS, weight variables are in individual files corresponding to the Wave 1 and Wave 4 Cohorts.
The "all-waves" weight file for the Wave 1 Cohort contains weights for participants who completed an interview in PATH-ATS and completed interviews in Waves 1, 2, 3, 4, and 5. The "all-waves" weight file for the Wave 4 Cohort contains weights for participants who completed an interview in PATH-ATS; all PATH-ATS respondents completed interviews in Wave 4 and Wave 5.
Dataset 2503 (DS2503) contains data derived from responses to questionnaires in Wave 1, Wave 2, Wave 3, Wave 4, Wave 4.5, Wave 5, Wave 5.5, and PATH-ATS, indicating if participants had ever/never used various tobacco products as of the Wave 5.5/PATH-ATS data collection period. This data file contains 26 variables for all 67,276 study participants as of the Wave 5.5/PATH-ATS data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 4001 (DS4001) contains the data from the Wave 7.5 Adult Questionnaire. This file contains 2,760 variables and 7,961 cases. Of these cases, 5,952 are continuing adults having completed a prior Adult Questionnaire. The other 2,009 cases are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 4002 (DS4002) contains the data from the Wave 7.5 Youth and Parent Questionnaire. This file contains 1,889 variables and 8,949 cases. Of these cases, 7,064 are continuing youth having completed a prior Youth Interview. The other 1,885 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 4111, 4112, 4121, 4122, 4221, 4222, 4231, and 4232 (DS4111, DS4112, DS4121, DS4122, DS4221, DS4222, DS4231, and DS4232) are data files comprising the weight variables for Wave 7.5. In Wave 7.5, the weight variables are in individual data files corresponding to the Wave 1, Wave 4, and Wave 7 Cohorts and different weight types.
There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight file for the Wave 1 Cohort contains weights for participants who completed a Wave 7.5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 4.5, 5, 5.5, 6, and 7. The "all-waves" weight file for the Wave 4 Cohort contains weights for participants who completed a Wave 7.5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 4.5, 5, 5.5, 6, and 7.
There are two separate sets of files with "single-waves" weights: one for the Wave 4 Cohort and one for the Wave 7 Cohort. The "single-wave" weight file for the Wave 4 Cohort contains weights for Wave 7.5 interview respondents in the Wave 4 Cohort, regardless of their response status at Waves 4.5, 5, 5.5, 6, or 7. The "single-wave" weight file for the Wave 7 Cohort contains weights for all Wave 7.5 interview respondents in the Wave 7 Cohort.
Dataset 4503 (DS4503) contains data derived from responses to questionnaires in Wave 1, Wave 2, Wave 3, Wave 4, Wave 4.5, Wave 5, Wave 5.5, PATH-ATS, Wave 6, Wave 7, and Wave 7.5, indicating if participants had ever/never used various tobacco products as of the Wave 7.5 data collection period. This data file contains 25 variables for all 82,139 study participants as of the Wave 7.5 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Monitoring the Future: Age 60 Panel Data, United States, 2018-2021 [Restricted-Use] (ICPSR 39779)
The longitudinal Monitoring the Future (MTF) Panel study extends the work of the cross-sectional MTF Main study by following a subsample of graduating seniors through the entire adult life course. The selected respondents are surveyed every two years from ages 19-30. Starting at age 35, respondents are surveyed every five years, at ages 35, 40, 45, 50, 55, and 60 (FZ surveys). The FZ surveys cover many of the same topics as the 12th grade and follow-up surveys and include additional questions on life events and health.
This study contains only the survey data for age 60 for the MTF longitudinal panel study participants that have reached age 60 (FZ6) through the 2021 data collection.
NOTE: Users must also request the core panel data file: MTF: Base Year and Follow-Up Core Panel Data, Ages 18-30, 1976-2021 (ICPSR 39223) because demographic information (e.g. sex, race/ethnicity) for the participants of the age 60 survey is included in the core panel data file.
Researchers can merge the Age 60 study data file with other MTF follow-up data in this series. This includes:
- MTF: Base Year and Follow-Up Core Data, Ages 18-30, 1976-2021 (ICPSR 39223)
- MTF: Base Year and Follow-Up Form 1 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39282)
- MTF: Base Year and Follow-Up Form 2 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39325)
- MTF: Base Year and Follow-Up Form 3 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39389)
- MTF: Base Year and Follow-Up Form 4 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39326)
- MTF: Base Year and Follow-Up Form 5 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39283)
- MTF: Base Year and Follow-Up Form 6 Panel Data, Ages 18-30, 1989-2021 (ICPSR 39388)
- MTF: Age 35 Panel Data, 1993-2021 [Restricted-Use] (ICPSR 39749)
- MTF: Age 40 and 45 Panel Data, 1998-2021 [Restricted-Use] (ICPSR 39767)
- MTF: Ages 50 and 55 Panel Data, 2008-2021 [Restricted-Use] (ICPSR 39804)
In addition to questions about lifetime, annual, and 30-day substance use, the Age 60 (FZ6) survey also includes questions covering:
- Substance use and its consequences (alcohol, marijuana/cannabis, other illicit drugs, substance use disorder symptoms)
- Methods of marijuana/cannabis use
- Own attitudes and perceptions about substance use
- Living arrangements and household characteristics
- Dating, marriage, and significant relationships
- Family roles, obligations, burdens, emotional support
- Employment/retirement
- Income, financial security, satisfaction
- Community involvement, social issues
- Local and global concerns
- Political interest and preferences
- Happiness; satisfaction with life domains and self
- Psychosocial constructs: self-esteem, locus of control, loneliness, risk-taking, boredom
- Health symptoms and illnesses, healthy behaviors, COVID-19, medical treatments
Please see the study documentation available on the MTF Panel series page for question-specific details, including content areas included in all survey forms.
HIGHLIGHTS of this update:
- Missing data coding has been changed/simplified in this release. Please see the User Guide for details.
- Panel analysis weights are now included in the data file instead of a stand-alone file. Please see the updated documentation for information.
Please be alert for variable coding differences between paper and web survey versions, especially for questions skipped based on answers to other questions. Note the following:
- The web-based version of the survey was introduced in 2020.
- Paper vs. Web coding differences will be most noticeable for the questions related to substance use, relationship/marital status, employment, and family composition.
- Users will need to explore their data using V60035 (89940:FZ PAPER OR WEB - RESPONSE) to look for and understand any coding differences.
Extensive work has been done to document the history and use of the MTF substance use disorder questions and criteria. Please see Substance use disorder criteria sums in the Monitoring the Future Panel Study (Occasional Paper No. 101).
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Monitoring the Future: Ages 50 and 55 Panel Data, United States, 2008-2021 [Restricted-Use] (ICPSR 39804)
The longitudinal MTF Panel study extends the work of the cross-sectional MTF Main study by following a subsample of graduating seniors through the entire adult life course. The selected respondents are surveyed every two years from ages 19-30. Starting at age 35, respondents are surveyed every five years, at ages 35, 40, 45, 50, 55, and 60 (FZ surveys). The FZ surveys cover many of the same topics as the 12th grade and follow-up surveys and include additional questions on life events and health.
This study contains only the survey data for ages 50 and 55 for the MTF longitudinal panel study participants that have reached age 50 (FZ4) and/or age 55 (FZ5) through the 2021 data collection.
NOTE: Users must also request the core panel data file: MTF: Base Year and Follow-Up Core Panel Data, Ages 18-30, 1976-2021 (ICPSR 39223) because demographic information (e.g. sex, race/ethnicity) for the participants of the age 50 and 55 surveys is included in the core panel data file.
Researchers can merge the Age 50-55 study data file with other MTF follow-up data in this series. This includes:
- MTF: Base Year and Follow-Up Core Data, Ages 18-30, 1976-2021 (ICPSR 39223)
- MTF: Base Year and Follow-Up Form 1 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39282)
- MTF: Base Year and Follow-Up Form 2 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39325)
- MTF: Base Year and Follow-Up Form 3 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39389)
- MTF: Base Year and Follow-Up Form 4 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39326)
- MTF: Base Year and Follow-Up Form 5 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39283)
- MTF: Base Year and Follow-Up Form 6 Panel Data, Ages 18-30, 1989-2021 (ICPSR 39388)
- MTF: Age 35 Panel Data, 1993-2021 [Restricted-Use] (ICPSR 39749)
- MTF: Ages 40 and 45 Panel Data, 1998-2021 [Restricted-Use] (ICPSR 39767)
- MTF: Age 60 Panel Data, 2018-2021 [Restricted-Use] (ICPSR 39779)
In addition to questions about lifetime, annual, and 30-day substance use, the Age 50 (FZ4) and Age 55 (FZ5) surveys also includes questions covering:
- Substance use and its consequences (alcohol, marijuana/cannabis, other illicit drugs, substance use disorder symptoms)
- Methods of marijuana/cannabis use
- Own attitudes and perceptions about substance use
- Living arrangements and household characteristics
- Dating, marriage, and significant relationships
- Family roles, obligations, burdens, emotional support
- Employment: income, financial security, satisfaction
- Community involvement, social issues
- Local and global concerns
- Political interest and preferences
- Happiness; satisfaction with life domains and self
- Psychosocial constructs: self-esteem, locus of control, loneliness, risk-taking, boredom
- Health symptoms and illnesses, healthy behaviors, COVID-19, medical treatments
Please see the study documentation available on the MTF Panel series page for question-specific details, including content areas included in all survey forms.
HIGHLIGHTS of this update:
- Missing data coding has been changed/simplified in this release. Please see the User Guide for details.
- Panel analysis weights are now included in the data file instead of a stand-alone file. Please see the updated documentation for information.
Please be alert for variable coding differences between paper and web survey versions, especially for questions skipped based on answers to other questions. Note the following:
- The web-based version of the survey was introduced in 2020.
- Paper vs. Web coding differences will be most noticeable for the questions related to substance use, relationship/marital status, employment, and family composition.
- Users will need to explore their data using V50035/V55035 (89940:FZ PAPER OR WEB - RESPONSE) to look for and understand any coding differences
Extensive work has been done to document the history and use of the MTF substance use disorder questions and criteria. Please see Substance use disorder criteria sums in the Monitoring the Future Panel Study (Occasional Paper No. 101).
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Population Assessment of Tobacco and Health (PATH) Study [United States] Special Collection Restricted-Use Files (ICPSR 37519)
The PATH Study was launched in 2011 to inform the Food and Drug Administration's regulatory activities under the Family Smoking Prevention and Tobacco Control Act (TCA). The PATH Study is a collaboration between the National Institute on Drug Abuse (NIDA), National Institutes of Health (NIH), and the Center for Tobacco Products (CTP), Food and Drug Administration (FDA). The study sampled over 150,000 mailing addresses across the United States to create a national sample of people who use or do not use tobacco.
45,971 adults and youth constitute the first (baseline) wave, Wave 1, of data collected by this longitudinal cohort study. These 45,971 adults and youth along with 7,207 "shadow youth" (youth ages 9 to 11 sampled at Wave 1) make up the 53,178 participants that constitute the Wave 1 Cohort. Respondents are asked to complete an interview at each follow-up wave. Youth who turn 18 by the current wave of data collection are considered "aged-up adults" and are invited to complete the Adult Interview. Additionally, "shadow youth" are considered "aged-up youth" upon turning 12 years old, when they are asked to complete an interview after parental consent.
At Wave 4, a probability sample of 14,098 adults, youth, and shadow youth ages 10 to 11 was selected from the civilian, noninstitutionalized population (CNP) at the time of Wave 4. This sample was recruited from residential addresses not selected for Wave 1 in the same sampled Primary Sampling Units (PSUs) and segments using similar within-household sampling procedures. This "replenishment sample" was combined for estimation and analysis purposes with Wave 4 adult and youth respondents from the Wave 1 Cohort who were in the CNP at the time of Wave 4. This combined set of Wave 4 participants, 52,731 participants in total, forms the Wave 4 Cohort.
At Wave 7, a probability sample of 14,863 adults, youth, and shadow youth ages 9 to 11 was selected from the CNP at the time of Wave 7. This sample was recruited from residential addresses not selected for Wave 1 or Wave 4 in the same sampled PSUs and segments using similar within-household sampling procedures. This "second replenishment sample" was combined for estimation and analysis purposes with the Wave 7 adult and youth respondents from the Wave 4 Cohorts who were at least age 15 and in the CNP at the time of Wave 7. This combined set of Wave 7 participants, 46,169 participants in total, forms the Wave 7 Cohort.
Please refer to the Restricted-Use Files User Guide that provides further details about children designated as "shadow youth" and the formation of the Wave 1, Wave 4, and Wave 7 Cohorts.
Wave 4.5 was a special data collection for youth only who were aged 12 to 17 at the time of the Wave 4.5 interview. Wave 4.5 was the fourth annual follow-up wave for those who were members of the Wave 1 Cohort. For those who were sampled at Wave 4, Wave 4.5 was the first annual follow-up wave.
Wave 5.5, conducted in 2020, was a special data collection for Wave 4 Cohort youth and young adults ages 13 to 19 at the time of the Wave 5.5 interview. Also in 2020, a subsample of Wave 4 Cohort adults ages 20 and older were interviewed via the PATH Study Adult Telephone Survey (PATH-ATS).
Wave 7.5 was a special collection for Wave 4 and Wave 7 Cohort youth and young adults ages 12 to 22 at the time of the Wave 7.5 interview. For those who were sampled at Wave 7, Wave 7.5 was the first annual follow-up wave.
Dataset 1002 (DS1002) contains the data from the Wave 4.5 Youth and Parent Questionnaire. This file contains 1,617 variables and 13,131 cases. Of these cases, 11,378 are continuing youth having completed a prior Youth Interview. The other 1,753 cases are "aged-up youth" having previously been sampled as "shadow youth"
Datasets 1112, 1212, and 1222, (DS1112, DS1212, and DS1222) are data files comprising the weight variables for Wave 4.5. The "all-waves" weight file contains weights for participants in the Wave 1 Cohort who completed a Wave 4.5 Youth Interview and completed interviews (if old enough to do so) or verified their information with the study (if not old enough to be interviewed) in Waves 1, 2, 3, and 4.
There are two separate files with "single wave" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "single-wave" weight file for the Wave 1 Cohort contains weights for youth who completed an interview in Wave 1 and in Wave 4.5, regardless of their participation in the intervening waves. The "single-wave" weight file for the Wave 4 Cohort contains weights for all Wave 4.5 Youth Interview respondents in the Wave 4 Cohort.
Dataset 1402 (DS1402) contains the Wave 4.5 State Identifier data for Youth and Parents and has 5 variables and 13,131 cases. The State Identifier dataset includes PERSONID for linking the State Identifier to the questionnaire data and 3 variables designating the state (state Federal Information Processing System (FIPS), state abbreviation, and full name of the state). The State Identifier values in this dataset represent participants' state of residence at the time of Wave 4.5.
Dataset 1503 (DS1503) contains data derived from responses to questionnaires in Wave 1, Wave 2, Wave 3, Wave 4, and Wave 4.5 indicating if participants had ever/never used various tobacco products as of the Wave 4.5 data collection period. This data file contains 26 variables for all 67,276 study participants as of the Wave 4.5 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 2001 (DS2001) contains the data from the Wave 5.5 Adult Questionnaire. This file contains 2,619 variables and 3,628 cases. Of these cases, 1,014 are continuing adults having completed a prior Adult Questionnaire. The other 2,614 cases are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 2002 (DS2002) contains the data from the Wave 5.5 Youth and Parent Questionnaire. This file contains 1,871 variables and 7,129 cases. Of these cases, 7,076 are continuing youth having completed a prior Youth Interview. The other 53 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 2111, 2112, 2121, 2122, 2221, and 2222 (DS2111, DS2112, DS2121, DS2122, DS2221, and DS2222) are data files comprising the weight variables for Wave 5.5. In Wave 5.5, the weight variables are in individual data files corresponding to the Wave 1 and Wave 4 Cohorts and different weight types.
There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight file for the Wave 1 Cohort contains weights for participants who completed a Wave 5.5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 4.5, and 5. The "all-waves" weight file for the Wave 4 Cohort contains weights for participants who completed a Wave 5.5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 4.5 and 5.
The "single-wave" weight file for the Wave 4 Cohort contains weights for all Wave 5.5 interview respondents.
Dataset 2401 (DS2401) contains the Wave 5.5 State Identifier data for Adults and has 5 variables and 3,628 cases. Dataset 2402 (DS2402) contains the Wave 5.5 State Identifier data for Youth and Parents and has 5 variables and 7,129 cases. The same 5.5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 5.5.
Dataset 2503 (DS2503) contains data derived from responses to questionnaires in Wave 1, Wave 2, Wave 3, Wave 4, Wave 4.5, Wave 5, and Wave 5.5 indicating if participants had ever/never used various tobacco products as of the Wave 5.5 data collection period. This data file contains 26 variables for all 67,276 study participants as of the Wave 5.5 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 3001 (DS3001) contains the data from PATH-ATS. This file contains 977 variables and 8,874 cases, all of which are continuing adults having completed a prior Adult Questionnaire, with their most recent interview in Wave 5.
Datasets 3111 and 3121 (DS3111 and DS3121) are data files comprising weights for PATH-ATS. In PATH-ATS, weight variables are in individual files corresponding to the Wave 1 and Wave 4 Cohorts.
The "all-waves" weight file for the Wave 1 Cohort contains weights for participants who completed an interview in PATH_-ATS and completed interviews in Waves 1, 2, 3, 4, and 5. The "all-waves" weight file for the Wave 4 Cohort contains weights for participants who completed an interview in PATH-ATS; all PATH-ATS respondents completed interviews in Wave 4 and Wave 5.
Dataset 3401 (DS3401) contains the PATH-ATS State Identifier data and has 5 variables and 8,874 cases. The State Identifier dataset includes PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in this dataset represents participants' state of residence at the time of PATH-ATS.
Dataset 4001 (DS4001) contains the data from the Wave 7.5 Adult Questionnaire. This file contains 3,142 variables and 7,961 cases. Of these cases, 5,952 are continuing adults having completed a prior Adult Questionnaire. The other 2,009 cases are "aged-up adults" having previously completed a Youth Questionnaire.
Dataset 4002 (DS4002) contains the data from the Wave 7.5 Youth and Parent Questionnaire. This file contains 2,169 variables and 8,949 cases. Of these cases, 7,064 are continuing youth having completed a prior Youth Interview. The other 1,885 cases are "aged-up youth" having previously been sampled as "shadow youth."
Datasets 4111, 4112, 4121, 4122, 4221, 4222, 4231, and 4232 (DS4111, DS4112, DS4121, DS4122, DS4221, DS4222, DS4231, and DS4232) are data files comprising the weight variables for Wave 7.5. In Wave 7.5, the weight variables are in individual data files corresponding to the Wave 1, Wave 4, and Wave 7 Cohorts and different weight types.
There are two separate sets of files with "all-waves" weights: one for the Wave 1 Cohort and one for the Wave 4 Cohort. The "all-waves" weight file for the Wave 1 Cohort contains weights for participants who completed a Wave 7.5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 1, 2, 3, 4, 4.5, 5, 5.5, 6, and 7. The "all-waves" weight file for the Wave 4 Cohort contains weights for participants who completed a Wave 7.5 interview and completed interviews (if old enough to do so) or verified their information (if not old enough to be interviewed) in Waves 4, 4.5, 5, 5.5, 6, and 7.
There are two separate sets of files with "single-waves" weights: one for the Wave 4 Cohort and one for the Wave 7 Cohort. The "single-wave" weight file for the Wave 4 Cohort contains weights for Wave 7.5 interview respondents in the Wave 4 Cohort, regardless of their response status at Waves 4.5, 5, 5.5, 6, or 7. The "single-wave" weight file for the Wave 7 Cohort contains weights for all Wave 7.5 interview respondents in the Wave 7 Cohort.
Dataset 4401 (DS4401) contains the Wave 7.5 State Identifier data for Adults and has 5 variables and 7,961 cases. Dataset 4402 (DS4402) contains the Wave 7.5 State Identifier data for Youth and Parents and has 5 variables and 8,949 cases. The same 7.5 variables are in each State Identifier dataset, including PERSONID for linking the State Identifier to the questionnaire and biomarker data and 3 variables designating the state (state FIPS, state abbreviation, and full name of the state). The State Identifier values in these datasets represent participants' state of residence at the time of Wave 7.5.
Dataset 4503 (DS4503) contains data derived from responses to questionnaires in Wave 1, Wave 2, Wave 3, Wave 4, Wave 4.5, Wave 5, Wave 5.5, PATH-ATS, Wave 6, Wave 7, and Wave 7.5 indicating if participants had ever/never used various tobacco products as of the Wave 7.5 data collection period. This data file contains 25 variables for all 82,139 study participants as of the Wave 7.5 data collection. This file is provided for reference only to simplify the definitions of tobacco use variables in the Adult and Youth data files for subsequent waves.
Dataset 4601 (DS4601) contains the Tobacco Universal Product Code (UPC) data from Wave 7.5. This data file contains 53 variables and 157 cases. This file contains UPC values on the packages of tobacco products used or in the possession of adult respondents at the time of Wave 7.5. The UPC values can be used to identify and validate the specific products used by respondents and augment the analyses of the characteristics of tobacco products used by these respondents at the time of Wave 7.5.
National Neighborhood Data Archive (NaNDA): Liquor, Tobacco, Cannabis, Vape, and Convenience Stores by Census Tract and ZCTA, United States, 1990-2022 (ICPSR 208907)
This dataset provides annual measures of the number and density of liquor, tobacco, cannabis, vape, and convenience stores per census tract and ZIP Code Tabulation Area (ZCTA) across the United States from 1990 through 2022. Data are derived from the National Establishment Time Series (NETS) database and are available for four geographies: Census Tract 2010, Census Tract 2020, ZCTA 2010, and ZCTA 2020.
Monitoring the Future: Age 35 Panel Data, United States, 1993-2021 [Restricted-Use] (ICPSR 39749)
The longitudinal Monitoring the Future (MTF) Panel study extends the work of the cross-sectional MTF Main study by following a subsample of graduating seniors through the entire adult life course. The selected respondents are surveyed every two years from ages 19-30. Starting at age 35, respondents are surveyed every five years, at ages 35, 40, 45, 50, 55, and 60 (FZ surveys). The FZ surveys cover many of the same topics as the 12th grade and follow-up surveys and include additional questions on life events and health.
This study contains only the age 35 survey data for the MTF longitudinal panel study participants that have reached age 35 (FZ1) through the 2021 data collection.
NOTE: Users must also request the core panel data file: MTF: Base Year and Follow-Up Core Panel Data, Ages 18-30, 1976-2021 (ICPSR 39223) because demographic information (e.g. sex, race/ethnicity) for the participants of the age 35 survey is included in the core panel data file.
Researchers can merge the Age 35 (FZ1) study data file with other MTF follow-up data in this series. This includes:
- MTF: Base Year and Follow-Up Core Data, Ages 18-30, 1976-2021 (ICPSR 39223)
- MTF: Base Year and Follow-Up Form 1 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39282)
- MTF: Base Year and Follow-Up Form 2 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39325)
- MTF: Base Year and Follow-Up Form 3 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39389)
- MTF: Base Year and Follow-Up Form 4 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39326)
- MTF: Base Year and Follow-Up Form 5 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39283)
- MTF: Base Year and Follow-Up Form 6 Panel Data, Ages 18-30, 1989-2021 (ICPSR 39388)
- MTF: Age 40-45 Panel Data, 1998-2021 [Restricted-Use] (ICPSR 39767)
- Forthcoming: MTF Panel Data for Ages 50-55, and 60
In addition to questions about lifetime, annual, and 30-day substance use, the Age 35 (FZ1) survey also includes questions covering:
- Substance use and its consequences (alcohol, marijuana/cannabis, other illicit drugs, substance use disorder symptoms)
- Methods of marijuana/cannabis use
- Own attitudes and perceptions about substance use
- Living arrangements and household characteristics
- Dating, marriage, and significant relationships
- Parenthood and family
- Employment: experiences, income, financial security, satisfaction
- Leisure time
- Local and global concerns
- Political interest and preferences
- Happiness; satisfaction with life domains and self
- Psychosocial constructs: self-esteem, locus of control, loneliness, risk-taking, boredom
- Health symptoms, healthy behaviors, COVID-19
Please see the study documentation available on the MTF Panel series page for question-specific details, including content areas included in all survey forms.
HIGHLIGHTS of this update:
- Missing data coding has been changed/simplified in this release. Please see the User Guide for details.
- Panel analysis weights are now included in the data file instead of a stand-alone file. Please see the updated documentation for information.
Please be alert for variable coding differences between paper and web survey versions, especially for questions skipped based on answers to other questions. Note the following:
- The web-based version of the survey was introduced in 2020.
- Paper vs. Web coding differences will be most noticeable for the questions related to substance use, relationship/marital status, employment, and family composition.
- Users will need to explore their data using V35035 (89940:FZ PAPER OR WEB - RESPONSE) to look for and understand any coding differences.
Extensive work has been done to document the history and use of the MTF substance use disorder questions and criteria. Please see Substance use disorder criteria sums in the Monitoring the Future Panel Study (Occasional Paper No. 101)
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Monitoring the Future: Ages 40 and 45 Panel Data, United States, 1998-2021 [Restricted-Use] (ICPSR 39767)
The longitudinal Monitoring the Future (MTF) Panel study extends the work of the cross-sectional MTF Main study by following a subsample of graduating seniors through the entire adult life course. The selected respondents are surveyed every two years from ages 19-30. Starting at age 35, respondents are surveyed every five years, at ages 35, 40, 45, 50, 55, and 60 (FZ surveys). The FZ surveys cover many of the same topics as the 12th grade and follow-up surveys and include additional questions on life events and health.
This study contains only the survey data for ages 40 and 45 for the MTF longitudinal panel study participants that have reached age 40 (FZ2) and/or age 45 (FZ3) through the 2021 data collection.
NOTE: Users must also request the core panel data file: MTF: Base Year and Follow-Up Core Panel Data, Ages 18-30, 1976-2021 (ICPSR 39223) because demographic information (e.g. sex, race/ethnicity) for the participants of the age 40 and 45 surveys is included in the core panel data file.
Researchers can merge the Age 40-45 study data file with other MTF follow-up data in this series. This includes:
- MTF: Base Year and Follow-Up Core Data, Ages 18-30, 1976-2021 (ICPSR 39223)
- MTF: Base Year and Follow-Up Form 1 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39282)
- MTF: Base Year and Follow-Up Form 2 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39325)
- MTF: Base Year and Follow-Up Form 3 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39389)
- MTF: Base Year and Follow-Up Form 4 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39326)
- MTF: Base Year and Follow-Up Form 5 Panel Data, Ages 18-30, 1976-2021 (ICPSR 39283)
- MTF: Base Year and Follow-Up Form 6 Panel Data, Ages 18-30, 1989-2021 (ICPSR 39388)
- MTF: Age 35 Panel Data, 1993-2021 [Restricted-Use] (ICPSR 39749)
- Forthcoming: MTF panel data for ages, 50-55, and 60
In addition to questions about lifetime, annual, and 30-day substance use, the Age 40 (FZ2) and Age 45 (FZ3) surveys also includes questions covering:
- Substance use and its consequences (alcohol, marijuana/cannabis, other illicit drugs, substance use disorder symptoms)
- Methods of marijuana/cannabis use
- Own attitudes and perceptions about substance use
- Living arrangements and household characteristics
- Dating, marriage, and significant relationships
- Family roles, obligations, burdens
- Employment: experiences, income, financial security, satisfaction
- Leisure time
- Local and global concerns
- Political interest and preferences
- Happiness; satisfaction with life domains and self
- Psychosocial constructs: self-esteem, locus of control, loneliness, risk-taking, boredom
- Health symptoms and illnesses, healthy behaviors, COVID-19
Please see the study documentation available on the MTF Panel series page for question-specific details, including content areas included in all survey forms.
HIGHLIGHTS of this update:
- Missing data coding has been changed/simplified in this release. Please see the User Guide for details.
- Panel analysis weights are now included in the data file instead of a stand-alone file. Please see the updated documentation for information.
Please be alert for variable coding differences between paper and web survey versions, especially for questions skipped based on answers to other questions. Note the following:
- The web-based version of the survey was introduced in 2020.
- Paper vs. Web coding differences will be most noticeable for the questions related to substance use, relationship/marital status, employment, and family composition.
- Users will need to explore their data using V40035/V45035 (89940:FZ PAPER OR WEB - RESPONSE) to look for and understand any coding differences.
Extensive work has been done to document the history and use of the MTF substance use disorder questions and criteria. Please see Substance use disorder criteria sums in the Monitoring the Future Panel Study (Occasional Paper No. 101)
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Boys Town Study of Youth Development, United States, mid-1970s (ICPSR 34595)
Monitoring the Future: Base Year & Follow-Up Core Panel Data, Ages 18-30, United States, 1976-2021 [Restricted-Use] (ICPSR 39223)
The Monitoring the Future (MTF) project is a long-term epidemiologic and etiologic study of substance use among youth and adults in the United States. It is conducted at the University of Michigan's Institute for Social Research and is funded by a series of investigator-initiated research grants from the National Institute on Drug Abuse.
The MTF panel study consists of six different survey forms (five forms from 1976-1988), and each survey contains a "core" set of questions about demographics and substance use. This study contains the "core" data for these questions compiled across all survey forms and years in which they are included for the longitudinal panel participants. Each record in the core panel dataset includes the respondent's data for their base year (BY) 12th grade survey (modal age 18) and their young adult follow-up FU surveys (modal ages 19-30).
The core panel dataset should be selected by all researchers. Use the linking variable available on all datasets, MTFID, to link the core dataset with all other MTF panel datasets.
Here is a list of subjects included in the core dataset:
Administrative variables
- Year of administration
- Survey form
- Survey date
- BY survey weight, sampling stratum and cluster
- FU panel analysis weights
Demographics
BY only
- #Parents in household
- Parent education levels
- Respondent's age in months
- Sex
- Race/Ethnicity
- Region of the country (school location)
- Population density/Urbanicity (school location)
- High school Zip Code, State and County FIPS codes (can be linked to user-provided data; results can be reported at no unit smaller than US geographical region)
- Absenteeism (illness, cutting, skipping class)
- High school program, Grades, post-high school plans
FU only
- Pregnancy status
- Household type
- Urbanicity
- Absenteeism (missing work due to illness, other)
- Vocational/Technical education, Armed forces, College attendance
- College grades, attendance, Greek life
BY and FU
- Marital status
- Household composition
- Political preference
- Religious attendance, importance, preference
- Evenings out, Dating
- Employment
- Salary/earned Income and Other Income
- Driving, tickets, and accidents related to alcohol and other substance use
Substance use
- Cigarette use
- Alcohol use (including binge drinking (e.g. 5+ drinks in a row/2 weeks), drunkenness)
- Marijuana/cannabis, hashish use
- LSD use
- Hallucinogen use, other than LSD
- Cocaine use (including cocaine, crack, other forms)
- Amphetamine use
- Sedatives/Barbiturate use
- Tranquilizer use
- Heroin use (with and without needles)
- Narcotics use (other than Heroin)
- Inhalant use
- Steroid use
- Ice use
- Methamphetamine use
- MDMA use
- Vaping: nicotine, marijuana, flavoring
Please see the study documentation available on the MTF Panel series page for question-specific details.
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Multilevel Influences on HIV and Substance Use in a YMSM Cohort (RADAR), Chicago Metropolitan Area, 2015-2020 (ICPSR 37603)
The National Institute on Drug Abuse (NIDA) funded RADAR in 2014 to collect multilevel, longitudinal data and biospecimens from an ethnically and racially diverse cohort of young, sexual and gender minorities (SGM; e.g., men who have sex with men (MSM), transgender women, gender non-conforming individuals) who were assigned male at birth (AMAB) (current core cohort n=1,113). The primary objective of this study is to apply a multilevel perspective to a syndemic of health issues associated with human immunodeficiency virus (HIV) in this population. The multilevel design focuses on individual, dyadic (i.e., sexual and romantic relationships), network (i.e., social, drug, and sexual connections) and biologic factors that may be associated with HIV. The cohort contains both HIV-negative and HIV-positive individuals, which allows for the development of a repository of biospecimens and HIV sequence data from both pre-infection and post-infection visits that will help facilitate future projects evaluating substance use, HIV risk, and pathogenesis.
A multiple cohort, accelerated longitudinal design was utilized by initially enrolling two existing SGM cohorts and then expanded through the use of convenience and snowball sampling methods. Enrollment criteria varied slightly based on the recruitment method, but overall inclusion criteria required participants to be AMAB, between 16 and 29 years of age, report having had sex with a man in the prior year or identify as a SGM, live in the Chicago metropolitan area, and be an English speaker. Study recruitment opened in February 2015. Participants are followed through the developmental period of late adolescence to early adulthood, which is a critical period of initiation and acceleration of sexual behavior and substance use. Study visits occur every six months.
Monitoring the Future: Base Year & Follow-Up Form 4 Panel Data, Ages 18-30, United States, 1976-2021 [Restricted-Use] (ICPSR 39326)
The MTF study consists of six different survey forms (five forms from 1976-1988). This study contains the data for Form 4 longitudinal panel participants. The MTF Form 4 restricted panel dataset includes data for the base year (BY) 12th grade surveys (modal age 18) and their young adult follow-up FU surveys (modal ages 19-30).
In addition to demographic-related questions and questions about lifetime, annual, and 30-day substance use that are included on all survey forms, Form 4 also includes questions covering:
- Beer, Wine, hard liquor, wine coolers use
- Vaping sources
- Flavored, small, and large cigars
- Hookah, dissolvable tobacco, snus, and smokeless tobacco use
- Own attitudes and perceptions about substance use
- Perceived risk of substance use
- Perceived friends' substance use
- Perceived addictiveness of substances
- Legal Issues Regarding Drugs
- Delinquency, victimization, and feeling safe at school
- Vocational plans, aspirations, expectations
- Preferences regarding job characteristics
- Desirability of different working arrangements and settings
- Work ethic/success orientation
- Dating and marriage: status, attitudes, expectations
- Parenthood: attitudes, expectations
- Values surrounding marriage and family
- Personal materialism
- Ecological concerns, conservation of resources
- Attitudes toward governmental policies and practices
- Local and global concerns
- Voting behavior
- Attitudes toward the military as an institution and occupation
- Happiness; satisfaction with life domains and self
Please see the study documentation available on the MTF Panel series page for question-specific details.
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Monitoring the Future: Base Year & Follow-up Form 3 Panel Data, Ages 18-30, United States, 1976-2021 [Restricted-Use] (ICPSR 39389)
The MTF study consists of six different survey forms (forms 1-5 began in 1976; form 6 was added in 1989). This study contains the data for Form 3 longitudinal panel participants. The MTF Form 3 Panel dataset includes data for the base year (BY) 12th grade surveys (modal age 18) and their young adult follow-up FU surveys (modal ages 19-30).
In addition to demographic-related questions and questions about lifetime, annual, and 30-day substance use that are included on all survey forms, Form 3 also includes questions covering:
- Attitudes toward governmental policies and practices
- Dating and marriage: status, attitudes, expectations
- Ecological concerns, conservation of resources
- Happiness; satisfaction with life domains and self
- Health symptoms, healthy behaviors, COVID-19
- Leisure time, including computer, cell phone, and social media use
- Local and global concerns
- Methods of marijuana use
- Own attitudes and perceptions about substance use
- Parenthood: attitudes, expectations
- Perceived friends' substance use
- Perceived risk of substance use
- Race relations
- Substance use consequences (alcohol, marijuana/cannabis, other illicit drugs)
NOTE: In 2020, school-based data collection was halted due to COVID-19. BY sample sizes were affected, and data for some questions on forms 2 and 3 were suppressed. The list of variables affected is found in the 2020 12th grade Codebook available through NAHDAP.
Please see the study documentation available on the MTF Panel series page for question-specific details, including content areas included in all survey forms.
NOTE: Researchers are encouraged to begin their work with the "core" data file, NAHDAP study 39223. Please see the User's Guide, IV. Working with the MTF Restricted Panel Data, for details.
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Monitoring the Future: Base Year & Follow-Up Form 2 Panel Data, Ages 18-30, United States, 1976-2021 [Restricted-Use] (ICPSR 39325)
The MTF study consists of six different survey forms (five forms from 1976-1988). This study contains the data for Form 2 longitudinal panel participants. The MTF Form 2 restricted panel dataset includes data for the base year (BY) 12th grade surveys (modal age 18) and their young adult follow-up FU surveys (modal ages 19-30).
In addition to demographic-related questions and questions about lifetime, annual, and 30-day substance use that are included on all survey forms, Form 2 also includes questions covering:
- Availability of drugs
- Confidence/trust in government
- Dating, marriage, and family
- Delinquency and victimization
- Expected future substance use
- Exposure to substance use
- Healthy behaviors, illness, COVID-19
- Leisure time activities, high school and post-high school
- Methods of marijuana use
- Military: plans for service, draft opinion
- Own attitudes and perceptions about substance use
- Perceived friends' substance use
- Perceived risk of substance use
- Psychosocial domains: boredom, loneliness, self-esteem, depressive affect, social support, self-efficacy, risk taking
- Satisfaction with life domains
- Sources of help and treatment for substance use
- Sources of marijuana
- Substance use initiation
- Vaping, including nicotine, marijuana, flavoring, sources
- Voting, political activism
NOTE: In 2020, school-based data collection was halted due to COVID-19. BY sample sizes were affected, and data for some questions on forms 2 and 3 were suppressed. The list of variables affected is found in the 2020 12th grade Codebook available through NAHDAP.
Please see the study documentation available on the MTF Panel series page for question-specific details , including content areas included in all survey forms.
NOTE: Researchers are encouraged to begin their work with the "core" data file, NAHDAP study 39223. Please see the User's Guide, IV. Working with the MTF Restricted Panel Data, for details.
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the MTF research team, describing the data collection and trends over time.
Monitoring the Future: Base Year & Follow-Up Form 1 Panel Data, Ages 18-30, United States, 1976-2021 [Restricted-Use] (ICPSR 39282)
The MTF study consists of six different survey forms (five forms from 1976-1988). This study contains the data for Form 1 longitudinal panel participants. The MTF Form 1 Panel dataset includes data for the base year (BY) 12th grade surveys (modal age 18) and their young adult follow-up FU surveys (modal ages 19-30).
In addition to demographic-related questions and questions about lifetime, annual, and 30-day substance use that are included on all survey forms, Form 1 also includes questions covering:
- incidence of first use
- co-use of substances
- sources of obtaining substances
- perceived friends' use
- perceived availability of substances
- when, where, and with who substance use is occurring
- modes of substance use administration
- reasons for use or non-use
- own attitudes about substance use
- perceived risk of use
- substance use advertising
- sources of help and treatment
- free time and activities
- role of citizens in government, confidence in government
- voting and political activism
- attitudes towards discrimination
- satisfaction with life domains
- healthy behaviors
- physical health symptoms
Please see the study documentation available on the MTF Panel series page for question-specific details.
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Monitoring the Future: Base Year & Follow-Up Form 5 Panel Data, Ages 18-30, United States, 1976-2021 [Restricted-Use] (ICPSR 39283)
The MTF study consists of six different survey forms (five forms from 1976-1988). This study contains the data for Form 5 longitudinal panel participants. The MTF Form 5 restricted panel dataset includes data for the base year (BY) 12th grade surveys (modal age 18) and their young adult follow-up FU surveys (modal ages 19-30).
In addition to demographic-related questions and questions about lifetime, annual, and 30-day substance use that are included on all survey forms, Form 5 also includes questions covering:
- Non-prescription substance use, including Ritalin, Adderall, Oxycontin, Vicodin, fentanyl
- Energy drinks/shots
- Flavored alcohol, alcohol+caffeine
- Flavored small and large cigars
- Hookah
- dissolvable tobacco, snus, smokeless tobacco
- Synthetic marijuana use
- Incidence of first use
- Perceived risk of substance use
- Own and others' attitudes and perceptions about substance use
- Exposure to substance use
- Substance use problems
- Reasons for substance use, abstention or stopping use
- Perceived availability of substances
- Expected future substance use
- Sources of help and treatment for substance use
- Job-related substance use testing
- Methods of substance use
- Satisfaction with life domains
- Interpersonal relationships
- Parenthood: status, attitudes, expectations
- Dating, marriage, and family: status, values, attitudes, expectations, sex roles
- Military: plans for service, attitudes toward the military as an institution and occupation
- Working arrangements and settings
- Work ethic/success orientation
- Leisure time: extent, activities, and attitudes
- Community involvement
- Voting and political activism
- Political interest and preference
- Concern for others, locally and globally
- Conservation of resources, ecological concerns, mass transit
- Attitudes towards discrimination
- Expectations concerning societal change
- Reactions to personal and social change
- Personal materialism
- Delinquency and victimization
- Psychosocial domains: boredom, loneliness, self-esteem, depressive affect,social support, self-efficacy, risk taking
- Healthy behaviors, illness, COVID-19
- Post high school: status, plans, characteristics
- High school sport involvement, concussion
- Substance use education in high school
Please see the study documentation available on the MTF Panel series page for question-specific details, including content areas included in all survey forms.
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Individualized Assessment and Treatment for Marijuana Dependence: Treatment Mechanisms, United States, 2013-2016 (ICPSR 39044)
Marijuana is the most commonly used illicit drug in the US, but treatment for marijuana dependence is not fully effective. The most effective treatments to date have employed motivational enhancement (MET) plus cognitive-behavioral coping skills treatment (CB) and contingency management (CM) for abstinence. This study was intended to deliver a treatment to enhance coping and self-efficacy to improve marijuana outcomes in the long term. Researchers are explored the idea that more tailored teaching of coping skills may result in improved outcomes for marijuana-dependence than those seen thus far. The Individualized Assessment and Treatment Program (IATP) for marijuana dependent patients employed experience sampling (ES) to determine the strengths and weaknesses of each patient in drug-use situations so that treatment could be tailored accordingly.
Participants were 198 men and women meeting criteria for marijuana dependence and randomly assigned to 9 sessions of treatment in one of 4 treatment conditions: Standardized MET plus CB (SMET-CB); SMET+ CM (SMET-CB-CM); IATP; or IATP + CM (IATP-CM). Patients in all treatments engaged in ES via cell-phone for two weeks prior to treatment, for a weekly period during treatment, for another week after treatment has ended, and for two weekly periods at months 8 and 14. In the IATP conditions, the information gathered from the pretreatment and during-treatment ES periods provided data for a functional analysis of patients' drug use and urges to use. Therapists used the information to address specific cognitions, affects, and behaviors that were adaptive and maladaptive, and tailored a specific coping skills program with the patient. During-treatment experience sampling allowed monitoring of the treatment goals and procedures, making the treatment adaptive. In the SMET-CB conditions the experience sampling data were not used in therapy, but still provides in-vivo measures of drug use and coping skills.
It was hypothesized that IATP conditions would yield significantly better coping skills acquisition than SMET-CB conditions, both at posttreatment and at extended follow-ups, and that change in coping skills would predict better outcomes for the IATP conditions. It was further predicted that the addition of CM to both IATP and SMET-CB would enhance short-term and long-term outcomes. The results would have implications for improved tailoring of treatment to patients' strength and deficits, and for the validity of the training of coping skills for cannabis relapse prevention. The data collected will shed light on the ways in which patients in treatment use coping skills in real-time contexts. Finally, the use of repeated ES periods will allow researchers to determine how treatment impacts thoughts, feelings and behaviors, and how these in turn affect outcome in the long and short term.
National Comorbidity Survey: Reinterview (NCS-2), 2001-2002 [Restricted-Use] (ICPSR 30921)
The NCS-2 was a re-interview of 5,001 individuals who participated in the Baseline (NCS-1). The study was conducted a decade after the initial baseline survey. The aim was to collect information about changes in mental disorders, substance use disorders, and the predictors and consequences of these changes over the ten years between the two surveys. The collection contains four major sections: the main survey, demographic data, diagnostic data, and state, county, and tract FIPS data.
In the main survey, respondents were asked about general physical and mental health. Questions focused on a variety of health issues, including limitations caused by respondents' health issues, substance use, childhood health, life-threatening illnesses, chronic conditions, medications taken in the past 12 months, level of functioning and symptoms experienced in the past 30 days, and any services used by the respondents since the (NCS-1). Additional questions focused on mental disorders including depression, bipolar disorder, specific and social phobias, generalized anxiety, intermittent explosive disorder, suicidality, post-traumatic stress disorder, neurasthenia, pre-menstrual dysphoric disorder, attention deficit/hyperactivity disorder, oppositional defiant disorder, conduct disorder, and separation anxiety. Respondents were also asked about their lives in general, with topics including employment, finances, marriage, children, their social lives, and stressful life events experienced in the past 12 months. Additionally, two personality assessments were included consisting of respondents' opinions on whether various true/false statements accurately described their personalities. Another focus of the main survey dealt with substance use and abuse, nonmedical use of prescription drugs, and polysubstance use. Interview questions in the NCS-2 Main Survey were customized to each respondent based on previous responses in the Baseline (NCS-1).
The second part contains demographic and other background information including age, education, employment, household composition, household income, marital status, and region.
The third part focuses on whether respondents met diagnostic criteria for psychological disorders asked about in the main survey.
The fourth part contains respondents' state, county, and tract FIPS data.
Developing Methods for Assessing Outcomes of Law and Policy on Drug Trafficking Offenders, Organizations, and Criminal Justice Responses, United States, 2000-2018 (ICPSR 38441)
This project sought to gather and analyze data on the effects of marijuana legalization from primary and secondary data sources that are both local and national in scope, and at both the individual and aggregate level. Since 1996, 37 states have passed statutes legalizing marijuana for medical and/or recreational use, while it has remained illegal under federal law. Jurisdictional and temporal variation in law creates a complex environment and substantial challenges for police and prosecutors charged with enforcement, and little is known about the justice system processing, public safety, and public health outcomes of evolving laws and policies.
Secondary criminal justice and public health data were gathered from federal, state, and local sources. Each source has a sufficiently long time series to provide statistical power and to allow for sometimes gradual implementation. The design exploits geographic and temporal variation in the implementation of marijuana law, using a difference-in-differences design that compares outcomes in states which implemented the policies with states that did not, before and after implementation.
National Evaluation of the Fighting Back Program: General Population Surveys, 1995-1999 (ICPSR 3801)
National Comorbidity Survey: Baseline (NCS-1), 1990-1992 (Restricted Version) (ICPSR 25381)
COEP Replication Package for "Higher education: The impact of recreational marijuana on college applications" (ICPSR 194847)
Twitter Tweets on Non-Tobacco Blunt Wraps in the USA from January 2017 to November 2021 (ICPSR 182001)
Monitoring the Future: Restricted-Use Panel Data, United States, 1976-2019 (ICPSR 37072)
The Monitoring the Future (MTF) project is a long-term epidemiologic and etiologic study of substance use among youth and adults in the United States. It is conducted at the University of Michigan's Institute for Social Research, and funded by a series of investigator-initiated research grants from the National Institute on Drug Abuse. MTF has two components: MTF Main and MTF Panel.
From its inception in 1975, the cross-sectional MTF Main study has collected data annually from nationally representative samples of 12,000-19,000 high school seniors in 12th grade located in approximately 135 schools nationwide. Beginning in 1991, similar annual cross-sectional surveys of nationally representative samples of 8th and 10th graders have been conducted. In all, approximately 45,000 students annually respond to about 100 drug use and demographic questions, as well as to about 200 additional questions divided among multiple survey forms on other topics such as attitudes toward government, social institutions, race relations, changing gender roles, educational aspirations, occupational aims, and marital plans.
The longitudinal MTF Panel study conducts follow-up surveys with representative subsamples of respondents from each 12th grade cohort participating in MTF Main. From each cohort, a sample of about 2,450 students are selected for longitudinal follow-up, with an oversampling of students who reported prior drug use during their 12th grade survey. Longitudinal follow-up currently spans modal ages 19-30 and 35-60. For surveys at modal ages 19-30, the sample is randomly split into two halves (approx. 1,225 each) to be followed every other year. One half-sample begins its first follow-up the year after high school (at modal age 19), and the other half-sample begins its first follow-up in the second year after high school (at modal age 20). Thus, six young adult follow-up (FU) surveys occur between modal ages 19-30, at modal ages 19/20 (FU1), 21/22 (FU2), 23/24 (FU3), 25/26 (FU4), 27/28 (FU5), and 29/30 (FU6). After age 30, respondents are surveyed every five years: 35, 40, 45, 50, 55, and 60 (these are referred to as FZ surveys). The FZ surveys cover many of the same topics as the 12th grade and FU surveys and include additional questions on life events and health.
MTF Panel surveys for the young adults (ages 19-30) were conducted using mailed paper surveys from 1977-2017. In 2018 and 2019, a random half of all those aged 19-30 received a mailed paper survey, while the other half were surveyed using a new procedure that encouraged participation using web surveys (web-push). The FZ surveys (ages 35-60) were conducted using mailed paper surveys through the 2019 data collection.
More information about the MTF project can be accessed through the Monitoring the Future website. Annual reports are published by the research team, describing the data collection and trends over time.
Survey of Consumer Attitudes and Behavior, Winter 1975 (ICPSR 7479)
The Survey of Consumer Attitudes and Behavior series (also known as the Surveys of Consumers) was undertaken to measure changes in consumer attitudes and expectations, to understand why such changes occur, and to evaluate how they relate to consumer decisions to save, borrow, or make discretionary purchases. The data regularly include the Index of Consumer Sentiment, the Index of Current Economic Conditions, and the Index of Consumer Expectations.
This survey was undertaken to assess consumer sentiment and buying plans. Open-ended questions were asked concerning evaluations and expectations about personal finances, employment, recession, price changes, and the national business situation. Additional variables probe respondents' buying intentions for a house, automobiles, appliances, and other consumer durables, and the respondents' appraisals of present market conditions for purchasing houses and other durables. Other variables probe respondents' opinions of the United States government's help to the South Vietnamese government, the seriousness of Arab nations' intentions regarding peace with Israel, women's right to abortion, voting for a woman or a Jew as a presidential candidate, gun permit law, causes of crime and lawlessness, chances of Russian adherence to a nuclear weapons limitation agreement with the United States, and communism in the United States and free speech. Additional topics covered include the proposed government tax returns, a solution to the energy crisis, the relative merits of buying a new or used car and the relative value of small foreign cars and the small American cars, job pay satisfaction, penalties for smoking marijuana, freedom to make uncomplimentary public speeches, monetary drive of lawyers and doctors and the state of the public good, satisfaction with life in the United States, government's expected role in racial integration and relations between white and Black people, vacation plans, and respondents' assessment of their financial status relative to the previous year. Information is also provided on respondents' car ownership and the make and use of it, political party self-identification and party candidate vote preference, self-identified ideological position, the neighborhood and house structure respondents live in, and spending plans for their income tax refunds. Demographic variables provide information on respondents' age, sex, race, marital status, occupation, employment status, religion, and family income.
Survey of Consumer Attitudes and Behavior, Spring 1975 (ICPSR 7480)
The Survey of Consumer Attitudes and Behavior series (also known as the Surveys of Consumers) was undertaken to measure changes in consumer attitudes and expectations, to understand why such changes occur, and to evaluate how they relate to consumer decisions to save, borrow, or make discretionary purchases. The data regularly include the Index of Consumer Sentiment, the Index of Current Economic Conditions, and the Index of Consumer Expectations.
This survey was undertaken to assess consumer sentiment and buying plans. Open-ended questions were asked concerning evaluations and expectations about personal finances, employment, recession, price changes, and the national business situation. Additional variables probe respondents' buying intentions for a house, automobiles, appliances, and other consumer durables, and the respondents' appraisals of present market conditions for purchasing houses and other durables. Other variables probe respondents' opinions of their health relative to that of other people in their age group, the relative merits of small and standard full-size cars as well as of small foreign cars and small American cars, the long-term cost and durability of certain household appliances, their satisfaction with the amount of money they had in savings, their satisfaction with life in the United States and with their lives in general, the United States government's help to the South Vietnamese government, and the seriousness of Arab nations' intentions regarding peace with Israel. Additional topics covered include a solution to the energy crisis, penalties for smoking marijuana, freedom to make uncomplimentary public speeches, communism in the United States and free speech, causes of crime and lawlessness, the role of government in improving the quality of life of the people, job satisfaction, monetary drive of lawyers and doctors and the state of the public good, and unionization of workers, as well as their financial status relative to the previous year and relative to that of their parents at a comparable age. Information is also provided on respondents' car ownership and the make and use of it, religious group affiliation, hobbies, political influence, political party identification, and self-identified ideological position. Demographic variables provide information on respondents' age, sex, race, marital status, education, occupation, employment status, religion, and family income.
Survey of Consumer Attitudes and Behavior, Fall 1973 (ICPSR 7525)
The Survey of Consumer Attitudes and Behavior series (also known as the Surveys of Consumers) was undertaken to measure changes in consumer attitudes and expectations, to understand why such changes occur, and to evaluate how they relate to consumer decisions to save, borrow, or make discretionary purchases. The data regularly include the Index of Consumer Sentiment, the Index of Current Economic Conditions, and the Index of Consumer Expectations.
This survey was undertaken to assess consumer sentiment and buying plans, as well as to provide information on their savings and investment habits and perceptions of government. Open-ended questions were asked concerning evaluations and expectations about personal finances, employment, recession, price changes, and the national business situation. Additional variables probe respondents' buying intentions for a house, automobiles, appliances, and other consumer durables, and respondents' appraisals of present market conditions for purchasing houses and other durables. Other variables probe respondents' assessments of their financial status relative to the previous year, their views of the government in Washington, the need for governmental changes, military spending, government support for Black people, and their satisfaction with their income and their jobs, as well as their opinion of married women working outside the home, women's liberation, and penalties for marijuana use. Information is also provided on respondents' political party identification, time spent with their children, savings accounts, contributions to charitable organizations, and car ownership and plans to buy a new one. Demographic variables provide information on respondents' age, sex, race, ethnic group, marital status, education, occupation, employment status, and family income.
Survey of Consumer Attitudes and Behavior, Fall 1974 (ICPSR 7524)
Effects of Marijuana Legalization on Law Enforcement and Crime, Washington, 2004-2018 (ICPSR 37661)
This study sought to examine the effects of cannabis legalization on crime and law enforcement in Washington State. In 2012 citizens voted to legalize possession of small amounts of cannabis, with the first licensed retail outlets opening on July 1, 2014. Researchers crafted their analysis around two questions. First, how are law enforcement agencies handling crime and offenders, particularly involving marijuana, before and after legalization? Second, what are the effects of marijuana legalization on crime, crime clearance, and other policing activities statewide, as well as in urban, rural, tribal, and border areas?
Research participants and crime data were collected from 14 police organizations across Washington, as well as Idaho police organizations situated by the Washington-Idaho border where marijuana possession is illegal. Additional subjects were recruited from other police agencies across Washington, prosecutors, and officials from the Washington State Department of Fish and Wildlife, Washington State Liquor and Cannabis Board, and the National Association of State Boating Law Administrators for focus groups and individual interviews. Variables included dates of calls for service from 2004 through 2018, circumstances surrounding calls for service, geographic beats, agency, whether calls were dispatch or officer initiated, and whether the agency was in a jurisdiction with legal cannabis.
Raising Healthy Children, Seattle Metropolitan Area, 2004-2011 (ICPSR 37584)
The Great Smoky Mountains Study (GSMS): Alcohol, Cannabis, Depression Disorders, North Carolina, 1992-2003 (ICPSR 37221)
The Great Smoky Mountain Study (GSMS) is a longitudinal epidemiological study of 1,420 children begun in 1992 in 11 rural counties in western North Carolina. Originally, the study had three aims: 1) to estimate the prevalence of common psychiatric disorders; 2) to study their development over time; and 3) to determine the level of mental health service use. The study expanded over time to include correlates and predictors of substance abuse and psychiatric problems. The study continued for over 20 years, with the original participants assessed up to 11 times from ages 9 to 30 (over 11,000 assessments total).
This collection includes data from study modules related to alcohol, cannabis, and depressive disorders in addition to core data on participants. This core data includes demographic variables related to age, sex, socioeconomic status, and race.
Epidemiologic Catchment Area Program Sites 1-4, 1979-1983 with National Death Index Data through 2007 (ICPSR 36621)
The Epidemiologic Catchment Area (ECA) program of research was initiated in response to the 1977 report of the President's Commission on Mental Health. The purpose was to collect data on the prevalence and incidence of mental disorders and on the use of and need for services by the mentally ill. Independent research teams at five universities (Yale University, Johns Hopkins University, Washington University, Duke University, and University of California at Los Angeles), in collaboration with the National Institute for Mental Health, conducted the studies with a core of common questions and sample characteristics. The sites were areas that had previously been designated as Community Mental Health Center catchment areas: New Haven, Connecticut, Baltimore, Maryland, St. Louis, Missouri, Durham, North Carolina, and Los Angeles, California. Each site sampled over 3,000 community residents and 500 residents of institutions, yielding 20,861 respondents overall. The longitudinal ECA design incorporated two waves of personal interviews administered one year apart and a brief telephone interview in between (for the household sample). The diagnostic interview used in the ECA was the NIMH Diagnostic Interview Schedule (DIS), Version III (with the exception of the Yale Wave I survey, which used Version II). Diagnoses were categorized according to the DIAGNOSTIC AND STATISTICAL MANUAL OF MENTAL DISORDERS, 3rd Edition (DSM-III). Diagnoses derived from the DIS include manic episode, dysthymia, bipolar disorder, single episode major depression, recurrent major depression, atypical bipolar disorder, alcohol abuse or dependence, drug abuse or dependence, schizophrenia, schizophreniform, obsessive compulsive disorder, phobia, somatization, panic, antisocial personality, and anorexia nervosa. The DIS uses the Mini-Mental State Examination (MMSE), which measures cognitive functioning, as an indirect measure of the DSM-III Organic Mental Disorders. In the ECA survey, this diagnosis is called cognitive impairment.
This collection features data from 17,327 participants across 2,005 variables. Data from the Los Angeles, California, Catchment (UCLA) are not included. Baseline data (Wave 1) and Wave 2 data were linked to the National Death Index through 2007, which includes primary and contributing causes of death, International Classification of Disease (ICD) codes, and nature of injury variables.
Drug consumption, collected online March 2011 to March 2012, English-speaking countries (ICPSR 36536)
The Simon Poll: Spring 2016 [Illinois Statewide] (ICPSR 100169)
National Survey on Drug Use and Health, 2014 (ICPSR 36361)
General Social Survey, 1972-2014 [Cumulative File] (ICPSR 36319)
Minnesota Adolescent Community Cohort (MACC) Study 2000-2013 (ICPSR 36282)
National Survey on Drug Use and Health, 2008 (ICPSR 26701)
The National Survey on Drug Use and Health (NSDUH) series (formerly titled National Household Survey on Drug Abuse) primarily measures the prevalence and correlates of drug use in the United States. Detailed NSDUH 2008 documentation is available from SAMHSA. The surveys are designed to provide quarterly, as well as annual, estimates. Information is provided on the use of illicit drugs, alcohol, and tobacco among members of United States households aged 12 and older. Questions included age at first use as well as lifetime, annual, and past-month usage for the following drug classes: marijuana, cocaine (and crack), hallucinogens, heroin, inhalants, alcohol, tobacco, and nonmedical use of prescription drugs, including pain relievers, tranquilizers, stimulants, and sedatives. The survey covered substance abuse treatment history and perceived need for treatment, and included questions from the Diagnostic and Statistical Manual (DSM) of Mental Disorders that allow diagnostic criteria to be applied. The survey included questions concerning treatment for both substance abuse and mental health related disorders. Respondents were also asked about personal and family income sources and amounts, health care access and coverage, illegal activities and arrest record, problems resulting from the use of drugs, and needle-sharing. Questions introduced in previous administrations were retained in the 2008 survey, including questions asked only of respondents aged 12 to 17. These "youth experiences" items covered a variety of topics, such as neighborhood environment, illegal activities, drug use by friends, social support, extracurricular activities, exposure to substance abuse prevention and education programs, and perceived adult attitudes toward drug use and activities such as school work. Several measures focused on prevention-related themes in this section. Also retained were questions on mental health and access to care, perceived risk of using drugs, perceived availability of drugs, driving and personal behavior, and cigar smoking. Questions on the tobacco brand used most often were introduced with the 1999 survey. For this 2008 survey, Adult mental health questions were added to measure symptoms of psychological distress in the worst period of distress that a person experienced in the past 30 days and suicidal ideation. A split-sample design also was included to administer separate sets of questions to assess impairment due to mental health problems. Background information includes gender, race, age, ethnicity, marital status, educational level, job status, veteran status, and current household composition.