A significant number of employees within the United States can be considered "informal" or
"off-the-books" workers. These workers, who by definition do not appear in administrative wage
records, are distinct from the larger group of private jobholders who do appear in administrative
records. However, while socioeconomic and spatial information on these individuals is readily
available in standard datasets, such as the 2000 Decennial Census Long Form, it is not possible
to identify the informal workers by only using such data because of the lack of accurate, formal
wage records. This study takes advantage of firm-based data that originates in Unemployment
Insurance administrative wage records linked with the Census Bureau's household-based data in
order to examine informal jobholders by their demographic characteristics as well as their
economic, commuting, and spatial location outcomes. In addition this report evaluates whether
informal jobholders should be included explicitly in future labor-workforce analyses and
transportation modeling. The analyses in this report use the sample of workers who lived in Los
Angeles County, California.
Social, Economic, Spatial, and Commuting Patterns of Self-Employed Jobholders
April 2007
Working Paper Number:
A significant number of employees within the United States identify themselves as selfemployed,
and they are distinct from the larger group identified as private jobholders. While
socioeconomic and spatial information on these individuals is readily available in standard
datasets, such as the 2000 Decennial Census Long Form, it is possible to gain further information
on their wage earnings by using data from administrative wage records. This study takes
advantage of firm-based data from Unemployment Insurance administrative wage records linked
with the Census Bureau's household-based data in order to examine self-employed jobholders -
both as a whole and as subgroups defined according to their earned wage status - by their
demographic characteristics as well as their economic, commuting, and spatial location
outcomes. Additionally, this report evaluates whether self-employed jobholders and the defined
subgroups should be included explicitly in future labor-workforce analyses and transportation
modeling. The analyses in this report use the sample of self-employed workers who lived in Los
Angeles County, California.
View Full
Paper PDF
Social, Economic, Spatial, and Commuting Patterns of Dual Jobholders
April 2007
Working Paper Number:
Individuals who hold multiple jobs have complex working lives and complex commuting
patterns. Economic and spatial information on these individuals is not readily available in
standard datasets, such as the 2000 Decennial Census Long Form, because the survey questions
were not designed to collect details on multiple jobs. This study takes advantage of firm-based
data from the Unemployment Insurance administrative wage records, linked with the Census
Bureau's household-based data, to examine multiple jobholders - and specifically a sentinel
group of dual jobholders. The study uses a sample from Los Angeles County, California and
examines the dual jobholders by their demographic characteristics as well as their economic,
commuting, and spatial location outcomes. In addition this report evaluates whether multiple
jobholders should be included explicitly in future labor-workforce analyses and transportation
View Full
Paper PDF
Further Evidence from Census 2000 About Earnings by Detailed Occupation for Men and Women: The Role of Race and Hispanic Origin
November 2011
Working Paper Number:
A 2004 report by the author reviewed data from Census 2000 and concluded "There is a substantial gap in median earnings between men and women that is unexplained, even after controlling for work experience (to the extent it can be represented by age and presence of children), education, and occupation." This paper extends the analysis and concludes that once those characteristics are controlled for, no further explanatory power is attributable to race or Hispanic origin.
View Full
Paper PDF
Design Comparison of LODES and ACS Commuting Data Products
October 2014
Working Paper Number:
The Census Bureau produces two complementary data products, the American Community Survey (ACS) commuting and workplace data and the Longitudinal Employer-Household Dynamics (LEHD) Origin-Destination Employment Statistics (LODES), which can be used to answer questions about spatial, economic, and demographic questions relating to workplaces and home-to-work flows. The products are complementary in the sense that they measure similar activities but each has important unique characteristics that provide information that the other measure cannot. As a result of questions from data users, the Census Bureau has created this document to highlight the major design differences between these two data products. This report guides users on the relative advantages of each data product for various analyses and helps explain differences that may arise when using the products.2,3
As an overview, these two data products are sourced from different inputs, cover different populations and time periods, are subject to different sets of edits and imputations, are released under different confidentiality protection mechanisms, and are tabulated at different geographic and characteristic levels. As a general rule, the two data products should not be expected to match exactly for arbitrary queries and may differ substantially for some queries.
Within this document, we compare the two data products by the design elements that were deemed most likely to contribute to differences in tabulated data. These elements are: Collection, Coverage, Geographic and Longitudinal Scope, Job Definition and Reference Period, Job and Worker Characteristics, Location Definitions (Workplace and Residence), Completeness of Geographic Information and Edits/Imputations, Geographic Tabulation Levels, Control Totals, Confidentiality Protection and Suppression, and Related
Public-Use Data Products.
An in-depth data analysis'in aggregate or with the microdata'between the two data products will be the subject of a future technical report. The Census Bureau has begun a pilot project to integrate ACS microdata with LEHD administrative data to develop an enhanced frame of employment status, place of work, and commuting. The Census Bureau will publish quality metrics for person match rates, residence and workplace match rates, and commute distance comparisons.
View Full
Paper PDF
September 2013
Working Paper Number:
The Census Bureau collects industry information through surveys and administrative data and creates associated public-use statistics. In this paper, we compare person-reported industry in the American Community Survey (ACS) to employer-reported industry from the Quarterly Census of Employment and Wages (QCEW) that is part of the Census Bureau's Longitudinal Employer-Household Dynamics (LEHD) program. This research provides necessary information on the use of administrative data as a supplement to survey data industry information, and the findings will be useful for anyone using industry information from either source. Our project is part of a larger effort to compare information on jobs from household survey data to employer-reported information. This research is the first to compare ACS job data to firm-based administrative data. We find an overall industry sector match rate of 75 percent, and a 61 percent match rate at the 4-digit Census Industry Code (CIC) level. Industry match rates vary by sector and by whether industry sector is classified using ACS or LEHD industry information. The educational services and health care and social assistance sectors have among the highest match rates. The management of companies and enterprises sector has the lowest match rate, using either ACS-reported or LEHD-reported sector. For individuals with imputed industry data, the industry sector match rate is only 14 percent. Our findings suggest that the industry distribution and the sample in a particular industry sector will differ depending on whether ACS or LEHD data are used.
View Full
Paper PDF
Assimilation and Coverage of the
Foreign-Born Population in Administrative Records
April 2015
Working Paper Number:
The U.S. Census Bureau is researching ways to incorporate administrative data in decennial census and survey operations. Critical to this work is an understanding of the coverage of the population by administrative records. Using federal and third party administrative data linked to the American Community Survey (ACS), we evaluate the extent to which administrative records provide data on foreign-born individuals in the ACS and employ multinomial logistic regression techniques to evaluate characteristics of those who are in administrative records relative to those who are not. We find that overall, administrative records provide high coverage of foreign-born individuals in our sample for whom a match can be determined. The odds of being in administrative records are found to be tied to the processes of immigrant assimilation - naturalization, higher English proficiency, educational attainment, and full-time employment are associated with greater odds of being in administrative records. These findings suggest that as immigrants adapt and integrate into U.S. society, they are more likely to be involved in government and commercial processes and programs for which we are including data. We further explore administrative records coverage for the two largest race/ethnic groups in our sample - Hispanic and non-Hispanic single-race Asian foreign born, finding again that characteristics related to assimilation are associated with administrative records coverage for both groups. However, we observe that neighborhood context impacts Hispanics and Asians differently.
View Full
Paper PDF
Exploring Differences in Employment between Household and Establishment Data
April 2009
Working Paper Number:
Using a large data set that links individual Current Population Survey (CPS) records to employer-reported administrative data, we document substantial discrepancies in basic measures of employment status that persist even after controlling for known definitional differences between the two data sources. We hypothesize that reporting discrepancies should be most prevalent for marginal workers and marginal jobs, and find systematic associations between the incidence of reporting discrepancies and observable person and job characteristics that are consistent with this hypothesis. The paper discusses the implications of the reported findings for both micro and macro labor market analysis
View Full
Paper PDF
Coverage and Agreement of Administrative Records and 2010 American Community Survey Demographic Data
November 2014
Working Paper Number:
The U.S. Census Bureau is researching possible uses of administrative records in decennial census and survey operations. The 2010 Census Match Study and American Community Survey (ACS) Match Study represent recent efforts by the Census Bureau to evaluate the extent to which administrative records provide data on persons and addresses in the 2010 Census and 2010 ACS. The 2010 Census Match Study also examines demographic response data collected in administrative records. Building on this analysis, we match data from the 2010 ACS to federal administrative records and third party data as well as to previous census data and examine administrative records coverage and agreement of ACS age, sex, race, and Hispanic origin responses. We find high levels of coverage and agreement for sex and age responses and variable coverage and agreement across race and Hispanic origin groups. These results are similar to findings from the 2010 Census Match Study.
View Full
Paper PDF
Noncitizen Coverage and Its Effects on U.S. Population Statistics
August 2023
Working Paper Number:
We produce population estimates with the same reference date, April 1, 2020, as the 2020 Census of Population and Housing by combining 31 types of administrative record (AR) and third-party sources, including several new to the Census Bureau with a focus on noncitizens. Our AR census national population estimate is higher than other Census Bureau official estimates: 1.8% greater than the 2020 Demographic Analysis high estimate, 3.0% more than the 2020 Census count, and 3.6% higher than the vintage-2020 Population Estimates Program estimate. Our analysis suggests that inclusion of more noncitizens, especially those with unknown legal status, explains the higher AR census estimate. About 19.8% of AR census noncitizens have addresses that cannot be linked to an address in the 2020 Census collection universe, compared to 5.7% of citizens, raising the possibility that the 2020 Census did not collect data for a significant fraction of noncitizens residing in the United States under the residency criteria used for the census. We show differences in estimates by age, sex, Hispanic origin, geography, and socioeconomic characteristics symptomatic of the differences in noncitizen coverage.
View Full
Paper PDF
A New Measure of Multiple Jobholding in the U.S. Economy
September 2020
Working Paper Number:
We create a measure of multiple jobholding from the U.S. Census Bureau's Longitudinal Employer-Household Dynamics data. This new series shows that 7.8 percent of persons in the U.S. are multiple jobholders, this percentage is pro-cyclical, and has been trending upward during the past twenty years. The data also show that earnings from secondary jobs are, on average, 27.8 percent of a multiple jobholder's total quarterly earnings. Multiple jobholding occurs at all levels of earnings, with both higher- and lower-earnings multiple jobholders earning more than 25 percent of their total earnings from multiple jobs. These new statistics tell us that multiple jobholding is more important in the U.S. economy than we knew.
View Full
Paper PDF