National Neighborhood Data Archive (NaNDA)
NaNDA is a nationwide collection of data with measures of the demographic, economic, social and physical environments, generally at the level of Zip Code Tabulation Areas (ZCTA) and Census Track. A Zip Code to ZCTA crosswalk is included. More detailed information can be found on the NaNDA webpage and ICPSR. The individual datasets are listed below.
Air Conditioning Probability
This dataset provides estimates of the probability that housing units have central air conditioning (central AC) in 2016 across United States Census Tracts using the 2010 FIPS code. For each tract, it includes the mean, minimum, maximum, median, and interquartile range (25th and 75th percentiles) of central AC probability, as well as the number of valid values used in these calculations.
Arts, Entertainment, and Recreation Organizations
These datasets contain measures of the number and per capita density of select types of arts, entertainment, and recreation organizations—such as museums, libraries, spectator sports organizations, amusement parks, fitness centers, bowling alleys, and casinos per United States census tract or ZCTA from 1990-2021.
Broadband Internet Availability, Speed, and Adoption
These datasets contain measures of broadband internet access and usage per census tract or ZCTA in 2014 through 2020. The data are derived primarily from internet service providers’ Form 477 reports to the Federal Communications Commission. Key variables include the average upload and download speed of fixed broadband connections, the number of internet service providers, and the number of households with broadband.
Civil, Social, And Religious Organizations
This dataset contains measures of the number and density of select types of civic, social, and religious organizations per United States Census Tract or ZIP Code Tabulation Area (ZCTA) from 1990 through 2021. The dataset includes four separate files for four different geographic areas (GIS shapefiles from the United States Census Bureau).
Crimes
This dataset contains county-level totals for the years 2002-2014 for eight types of crime: murder, rape, robbery, aggravated assault, burglary, larceny, motor vehicle theft, and arson. These crimes are classed as Part I criminal offenses by the United States Federal Bureau of Investigations (FBI) in their Uniform Crime Reporting (UCR) program. Each record in the dataset represents the total of each type of criminal offense reported in (or, in the case of missing data, attributed to) the county in a given year.
Dollar Stores
These datasets contain measures of the number and density of dollar stores per United States census tract or ZCTA from 1990-2021.
Eating and Drinking Places
These datasets contain measures of the number and density of eating and drinking places – including fast food restaurants, coffee shops, and bars – per census tract or ZCTA in the United States from 1990-2021.
Education and Training Services
These datasets contain measures of the number and per capita density of education and training services per United States census tract or ZCTA from 2003 through 2017. This includes traditional education establishments such as elementary schools, secondary schools, and colleges, as well as businesses offering specialized training such as art classes, driving instruction, computer training, and standardized test preparation.
Essential Businesses
This dataset contains measures of the number and density of businesses and their employees deemed essential in the first year (2020) of the COVID-19 pandemic by the US Department of Homeland Security’s Cybersecurity & Infrastructure Security Agency (CISA) in versions 3.0 (April 17, 2020) and 4.0 (August 18, 2020) of their advisory guidance on the essential critical infrastructure workforce. Measures are provided for 2020 per United States Census Tract or ZIP Code Tabulation Area (ZCTA).
Essential Workers
During the COVID-19 pandemic, certain occupations and industries were deemed “essential”, and typically included individuals who worked in healthcare, food service, public transportation, etc. However, early on in the pandemic, while these workers faced disproportionately higher risks, they often did not receive adequate personal protective equipment (PPE), were unable to work from home, and were limited in their ability to take other precautions to safeguard their health (Chen et al., 2021). As a result, previous studies have documented higher rates of infection, hospitalization, and death among essential workers compared to their non-essential worker counterparts (Selden & Berdahl, 2021; Wei et al., 2022). This dataset provides users with information on the number and proportion of essential workers in census tracts or ZIP Code tabulation areas (ZCTAs) in the United States over the 2016-2020 period. These data were added to the GLR in June 2024.
Grocery Stores
These datasets contain measures of the number and density of grocery stores – including supermarkets, specialty food stores, and warehouse clubs – per United States census tract or ZCTA from 2003 through 2017. These types of businesses represent places where neighborhood residents can obtain fresh and healthy foods.
Health Care Services
These datasets describe the number and density of health care services in each census tract or ZCTA in the United States from 1990-2021. The data includes counts, per capita densities, and area densities for many types of businesses in the health care sector, including doctors, dentists, mental health providers, nursing homes, and pharmacies.
Home Mortgage
The Home Mortgage Disclosure Act (HMDA) database (Consumer Financial Protection Bureau, 2022) has compiled mortgage lending data since 1981, but the collection and dissemination methods have changed over time (Federal Financial Institutions Examination Council, 2018), creating barriers to conducting longitudinal analyses. This HMDA Longitudinal Dataset (HLD) organizes and standardizes information across different eras of HMDA data collection between 1981 and 2020, enabling such analysis. The data in the GLR are HMDA aggregated data by census tract for each decade. Items for analysis include borrower income values, mortgages by loan type (e.g., conventional, Federal Housing Administration (FHA), Veterans Affairs (VA), refinances), and mortgages by borrower race and gender.
Hospitals
This dataset contains measures of the number and density of hospitals per United States Census Tract or ZIP Code Tabulation Area (ZCTA) in 2023.
Internet Access
These datasets contain measures of internet access per United States census tract or ZCTA from the 2015-2019 American Community Survey five-year estimate. Key variables include the number and percent of households with any type of internet subscription, with broadband internet, and with a computer or smartphone.
Land Cover
These datasets contain measures of land cover (e.g., low-, medium-, or high-density development, forest, wetland, open water) derived from the National Land Cover Database (NLCD) and aggregated by US census tract or ZCTA for 1985 through 2023. Land cover is measured both in total square meters and as a proportion of all land within the tract or ZCTA.
Law Enforcement Organizations
These datasets contain measures of the number and per capita density of law enforcement and safety organizations—such as police and fire departments, courts, jails, and lawyers—per United States census tract or ZCTA from 1990-2021.
Libraries
This dataset contains measures of the number and density of libraries per United States Census Tract or ZIP Code Tabulation Area (ZCTA) from 1992-2021.
Liquor, Tobacco, and Convenience Stores
These datasets contain measures of the number and density of liquor, tobacco, and convenience stores per United States census tract or ZCTA from 1990-2021.
Neighborhood-School Gap
These datasets contain measures of neighborhood-school gap for 2009-2010 and 2015-2016. Neighborhood-school gap (NS gap) refers to the discrepancy between the demographics of a public school and its surrounding community. For example, if 60% of a school’s student body is Black, but 30% of the neighborhood population is Black, the school has a positive Black neighborhood-school gap. The datasets measure gaps in race and poverty between elementary school student populations and the census tracts or ZCTAs that those elementary schools serve. Supplemental data containing component variables used to calculate NS gap at the school and block group level is also available.
Ophthalmologists
This dataset contains measures of the number and density of ophthalmologists per United States Census Tract or ZIP Code Tabulation Area (ZCTA) from 1990 through 2021.
Parks
These datasets describe the number and area of parks in each census tract or ZCTA in the United States for 2018 and 2022. Measures include the total number of parks, park area, and proportion of park area within each census tract or ZCTA.
Personal Care Services and Laundromats
These datasets contain measures of the number and per capita density of personal care services (such as barber shops, hair and nail salons, and spas) and laundromats per United States census tract or ZCTA from 1990-2021.
Polluting Sites
These datasets contain counts of polluting sites in each United States census tract or ZCTA and within a 0.5-mile buffer to capture spillover effects form 1987-2021. Polluting sites are taken from the US Environmental Protection Agency’s (EPA) Toxics Release Inventory. These facilities are typically larger and involved in manufacturing, metal mining, electric power generation, chemical manufacturing, and hazardous waste treatment.
Post Offices and Banks
These datasets contain measures of the number and per capita density of post offices and banks per United States census tract or ZCTA from 1990-2021.
Primary and Secondary Roads
These datasets contain measures of primary and secondary roads (highways and main arteries) per United States census tract or ZCTA in 2010 and 2020. These measures may be used as a proxy for heavy traffic, high traffic speeds, and impediments to walking or biking. Variables include counts of primary, secondary, and all streets per tract or ZCTA; total length of primary, secondary, and all streets per tract or ZCTA; ratio of primary and/or secondary road counts to all roads; and ratio of length of primary/secondary roads to all streets.
PRISM Climate
The PRISM NaNDA dataset provides daily weather data—minimum temperature (tmin), maximum temperature (tmax), and precipitation (ppt)—for all census tracts in the contiguous United States (CONUS) from 1981 to 2024. These data are derived from Oregon State University’s PRISM Climate Group (Northwest Alliance for Computational Science & Engineering & Oregon State University, 2025), which produces high- resolution (4 km x 4 km) gridded climate estimates.
In addition to daily values, the dataset includes two types of annual tract-level summary measures: (1) Percentiles (0.5th, 1st, 5th, 50th, 95th, 99th, and 99.5th), calculated using a rolling 10-year window of historical data, available for tmin, tmax, and ppt. (2) Percents, representing the proportion of days per year that fall above or below these percentile thresholds, available for tmin and tmax only. These features enable robust analyses of long-term environmental trends, extreme weather events, and their potential impacts on population health.
Public Transit Stops
These datasets list the number of public transit stops per United States census tract or ZCTA based on data from the National Transit Map (NTM). Each observation represents the count and density (per capita and square mile) of transit stops within a census tract or ZCTA, as voluntarily reported to NTM between 2016 and 2018 by one of 270 regional transit agencies choosing to participate. Data were also compiled on January 13, 2024 from the Bureau of Transportation Statistics.
Religious, Civic, and Social Organizations
These datasets contain measures of the number and per capita density of select types of religious, civic, and social organizations – such as churches, mosques, synagogues, ethnic associations, and veterans’ associations – per United States census tract or ZCTA from 2003 through 2017.
Retail Establishments
These datasets contain measures of the number and per capita density of select types of retail establishments—such as clothing, department, building and garden, furniture, and thrift stores—per United States census tract or ZCTA from 1990-2021.
School Counts and Characteristics
These datasets contain data on schools and school districts by district, census tract or ZCTA within the United States from 2000 through 2018. Key variables include district-level enrollment by race and ethnicity; numbers of teachers and counselors; teacher-student ratios; counts of public, private, and charter schools within districts; and expenditures and revenue, including per-pupil revenue.
SES and Demography
These datasets contain measures of socioeconomic and demographic characteristics by US census tract or ZCTA for different years through 1990-2022. Example measures include population density; population distribution by race, ethnicity, age, and income; and proportion of population living below the poverty level, receiving public assistance, and female-headed families. The dataset also contains a set of index variables to represent neighborhood disadvantage and affluence.
Social Services
These datasets contain measures of the number and per capita density of social services—such as senior centers, youth centers, food banks, job training programs, and day care centers—per United States census tract or ZCTA from 1990-2021.
Street Connectivity
These datasets contain measures of street connectivity (how well streets connect with one another) within all United States census tracts or ZCTAs for 2010 and 2020. This includes measures of the number of street segments (links) and intersections (nodes) per tract, street length within tracts, and indices representing overall connectivity within the tract or ZCTA.
Traffic Volume
These datasets contain measures of traffic volume per census tract or ZCTA in the United States from 1963 to 2019 (primarily 1997 to 2019). High traffic volume may be used as a proxy for heavy traffic, high traffic speeds, and impediments to walking or biking. The dataset contains measures of the average, maximum, and minimum traffic volume per tract or ZCTA per year. These figures are available for all streets, highways, and non-highways.
Urbanicity
This dataset contains measures of the urban/rural characteristics of each census tract in the United States for 2010. These include proportions of urban and rural population, population density, rural/urban commuting area (RUCA) codes, and RUCA-based four- and seven- category urbanicity scales.
Voter Registration, Turnout, and Partisanship
This dataset contains counts of voter registration and voter turnout for all counties in the United States for the years 2004-2018. It also contains measures of each county’s Democratic and Republican partisanship, including six-year longitudinal partisan indices for 2006-2016.
Weather
These datasets contain measures of weather by county for the years 2003-2016. Measures include average daily temperature, freezing days, cold days, hot days, rainy days, and snowy days.