Papers Containing Keywords(s): 'statistical'
The following papers contain search terms that you selected. From the papers listed below, you can navigate to the PDF, the profile page for that working paper, or see all the working papers written by an author. You can also explore tags, keywords, and authors that occur frequently within these papers.
See Working Papers by Tag(s), Keywords(s), Author(s), or Search Text
Click here to search again
Frequently Occurring Concepts within this Search
Viewing papers 71 through 80 of 100
-
Working PaperAccess Methods for United States Microdata
August 2007
Working Paper Number:
CES-07-25
Beyond the traditional methods of tabulations and public-use microdata samples, statistical agencies have developed four key alternatives for providing non-government researchers with access to confidential microdata to improve statistical modeling. The first, licensing, allows qualified researchers access to confidential microdata at their own facilities, provided certain security requirements are met. The second, statistical data enclaves, offer qualified researchers restricted access to confidential economic and demographic data at specific agency-controlled locations. Third, statistical agencies can offer remote access, through a computer interface, to the confidential data under automated or manual controls. Fourth, synthetic data developed from the original data but retaining the correlations in the original data have the potential for allowing a wide range of analyses.View Full Paper PDF
-
Working PaperUsing the P90/P10 Index to Measure U.S. Inequality Trends with Current Population Survey Data: A View From Inside the Census Bureau Vaults
June 2007
Working Paper Number:
CES-07-17
The March Current Population Survey (CPS) is the primary data source for estimation of levels and trends in labor earnings and income inequality in the USA. Time-inconsistency problems related to top coding in theses data have led many researchers to use the ratio of the 90th and 10th percentiles of these distributions (P90/P10) rather than a more traditional summary measure of inequality. With access to public use and restricted-access internal CPS data, and bounding methods, we show that using P90/P10 does not completely obviate time inconsistency problems, especially for household income inequality trends. Using internal data, we create consistent cell mean values for all top-coded public use values that, when used with public use data, closely track inequality trends in labor earnings and household income using internal data. But estimates of longer-term inequality trends with these corrected data based on P90/P10 differ from those based on the Gini coefficient. The choice of inequality measure matters.View Full Paper PDF
-
Working PaperResident Perceptions of Crime: How Similar are They to Official Crime Rates?
March 2007
Working Paper Number:
CES-07-10
This study compares the relationship between official crime rates and residents' perceptions of crime in census tracts. Employing a unique dataset that links household level data from the American Housing Survey metro samples over a period of 25 years (1976-2000) with official crime rate data for census tracts in selected cities during selected years, this large sample provides considerable ability to generalize the findings. I find that residents' perception of crime is most strongly related to official rates of tract violent crime. Models simultaneously taking into account both violent and property crime consistently found that property crime actually has a negative effect on perceived crime. Among types of violent crime, the robbery rate is consistently related to higher levels of perceived crime in the tract, whereas it appears a structural shift occurred in the mid-1980s in which aggravated assault and murder rates now impact perceptions of crime, even when taking into account the robbery rate.View Full Paper PDF
-
Working PaperDistribution Preserving Statistical Disclosure Limitation
September 2006
Working Paper Number:
tp-2006-04
One approach to limiting disclosure risk in public-use microdata is to release multiply-imputed, partially synthetic data sets. These are data on actual respondents, but with confidential data replaced by multiply-imputed synthetic values. A mis-specified imputation model can invalidate inferences because the distribution of synthetic data is completely determined by the model used to generate them. We present two practical methods of generating synthetic values when the imputer has only limited information about the true data generating process. One is applicable when the true likelihood is known up to a monotone transformation. The second requires only limited knowledge of the true likelihood, but nevertheless preserves the conditional distribution of the confidential data, up to sampling error, on arbitrary subdomains. Our method maximizes data utility and minimizes incremental disclosure risk up to posterior uncertainty in the imputation model and sampling error in the estimated transformation. We validate the approach with a simulation and application to a large linked employer-employee database.View Full Paper PDF
-
Working PaperMeasuring Poverty in the United States: History and Current Issues
April 2006
Working Paper Number:
CES-06-11
Formal measurement of poverty in the United States is now about 40 years old. This paper first briefly describes the origins and basis of the official poverty thresholds adopted by the federal government in the late 1960s. Then, it discusses in some detail some of the more current issues that observers suggest must be addressed if changes are to be made. The final sections discuss recent efforts to propose alternates to the current official approach.View Full Paper PDF
-
Working PaperConfidentiality Protection in the Census Bureau Quarterly Workforce Indicators
February 2006
Working Paper Number:
tp-2006-02
The QuarterlyWorkforce Indicators are new estimates developed by the Census Bureau's Longitudinal Employer-Household Dynamics Program as a part of its Local Employment Dynamics partnership with 37 state Labor Market Information offices. These data provide detailed quarterly statistics on employment, accessions, layoffs, hires, separations, full-quarter employment (and related flows), job creations, job destructions, and earnings (for flow and stock categories of workers). The data are released for NAICS industries (and 4-digit SICs) at the county, workforce investment board, and metropolitan area levels of geography. The confidential microdata - unemployment insurance wage records, ES-202 establishment employment, and Title 13 demographic and economic information - are protected using a permanent multiplicative noise distortion factor. This factor distorts all input sums, counts, differences and ratios. The released statistics are analytically valid - measures are unbiased and time series properties are preserved. The confidentiality protection is manifested in the release of some statistics that are flagged as "significantly distorted to preserve confidentiality." These statistics differ from the undistorted statistics by a significant proportion. Even for the significantly distorted statistics, the data remain analytically valid for time series properties. The released data can be aggregated; however, published aggregates are less distorted than custom postrelease aggregates. In addition to the multiplicative noise distortion, confidentiality protection is provided by the estimation process for the QWIs, which multiply imputes all missing data (including missing establishment, given UI account, in the UI wage record data) and dynamically re-weights the establishment data to provide state-level comparability with the BLS's Quarterly Census of Employment and Wages.View Full Paper PDF
-
Working PaperNew Approaches to Confidentiality Protection Synthetic Data, Remote Access and Research Data Centers
June 2004
Working Paper Number:
tp-2004-03
View Full Paper PDF
-
Working PaperSynthetic Data and Confidentiality Protection
September 2003
Working Paper Number:
tp-2003-10
View Full Paper PDF
-
Working PaperUsing Worker Flows in the Analysis of the Firm
August 2003
Working Paper Number:
tp-2003-09
This paper uses a novel approach to measure firm entry and exit, mergers and acquisition. It uses information about the flows of clusters of workers across business units to identify longitudinal linkage relationships in longitudinal business data. These longitudinal relationships may be the result of either administrative or economic changes and we explore both types of newly identified longitudinal relationships. In particular, we develop a set of criteria based on worker flows to identify changes in firm relationships ? such as mergers and acquisitions, administrative identifier changes and outsourcing. We demonstrate how this new data infrastructure and this cluster flow methodology can be used to better differentiate true firm entry/exit and simple changes in administrative identifiers. We explore the role of outsourcing in a variety of ways but in particular the outsourcing of workers to the temporary help industry. While the primary focus is on developing the data infrastructure and the methodology to identify and interpret these clustered flows of workers, we conclude the paper with an analysis of the impact of these changes on the earnings of workers.View Full Paper PDF
-
Working PaperUsing Administrative Earnings Records to Assess Wage Data Quality in the March Current Population Survey and the Survey of Income and Program Participation
November 2002
Working Paper Number:
tp-2002-22
The March Current Population Survey (CPS) and the Survey of Income and Program Participation (SIPP) produce different aggregates and distributions of annual wages. An excess of high wages and shortage of low wages occurs in the March CPS. SIPP shows the opposite, an excess of low wages and shortage of high wages. Exactly-matched Detailed Earnings Records (DER) from the Social Security Administration allow comparing March CPS and SIPP people's wages using data independent of the surveys. Findings include the following. March CPS and SIPP people differ little in their true wage characteristics. March CPS and SIPP represent a worker's percentile rank better than the dollar amount of wages. Workers with one job and low work effort have underestimated March CPS wages. March CPS has a higher level of "underground" wages than SIPP, and increasingly so in the 1990s. March CPS has a higher level of self-employment income "misclassified" as wages than SIPP, and increasingly so in the 1990s. These trends may explain one-third of March CPS's 6-percentage-point increase in aggregate wages relative to independent estimates from 1993 to 1995. Finally, the paper delineates March CPS occupations disproportionately likely to be absent from the administrative data entirely or to "misclassify" self-employment income as wages.View Full Paper PDF