Journal of Statistical Software
Wiley Interdisciplinary Reviews Computational StatisticsPublished 30 June 2009Open access
Jan de Leeuw
Citations3,607
SJR quartileQ1
SJR score1.45
SNIP3.11
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
The history, motivation, and implementation of the Journal of Statistical Software are discussed, including the history of the journal itself and its purpose and scope.
Abstract
Abstract The Journal of Statistical Software is an e‐journal that publishes and reviews open source statistical software. We discuss the history, motivation, and implementation of the journal. Copyright © 2009 John Wiley & Sons, Inc. This article is categorized under: Applications of Computational Statistics > Organizations and Publications
Keywords
Computer ScienceDecision SciencesMathematics
Journal of Statistical Planning and InferenceData-swapping: A technique for disclosure control
347 Citations1982Tore Dalenius, Steven P. Reiss
Data-swapping is a data transformation technique where the underlying statistics of the data are preserved and can be used as a basis for microdata release or to justify the release of tabulations.
Journal of the American Statistical AssociationThe Multiple Adaptations of Multiple Imputation
257 Citations2007Jerome P. Reiter, Trivellore E. Raghunathan
Some of the main adaptations of the multiple-imputation framework, including missing data in large and small samples, data confidentiality, and measurement error, are described, and the combining rules for each setting are reviewed and explained.
Journal of the Royal Statistical Society Series A (Statistics in Society)Releasing Multiply Imputed, Synthetic Public use Microdata: An Illustration and Empirical Study
236 Citations2004Jerome P. Reiter
Simulations based on data from the US Current Population Survey are used to evaluate the potential validity of inferences based on fully synthetic data for a variety of descriptive and analytic estimands and to illustrate the specification of synthetic data imputation models.
Journal of the American Statistical AssociationDisclosure-Limited Data Dissemination
218 Citations1986George T. Duncan, Diane Lambert
Common disclosure control policies, such as requiring released cell relative frequencies to be bounded away from both zero and one, are shown to be equivalent to disclosure rules that allow data release only if specific uncertainty functions at particular predictive distribution are allowed.
The American StatisticianA Framework for Evaluating the Utility of Data Altered to Protect Confidentiality
195 Citations2006Alan F. Karr, Christine N. Kohnen +3 more
Using both genuine and simulated data, utility measures can be used in a decision-theoretic formulation for evaluating disclosure limitation procedures and differences in inferences obtained from the altered data and corresponding inferences from the original data are presented.
Statistical ScienceEnhancing Access to Microdata While Protecting Confidentiality: Prospects for the Future
160 Citations1991George T. Duncan, R. W. Pearson
This article presents a scenario for the future of research access to federally collected microdata, as they relate to improvements in database techniques, computer and analytical method- ologies and legal and administrative arrangements for access to and protection of federal statistics.
Journal of the American Statistical AssociationEstimating Risks of Identification Disclosure in Microdata
131 Citations2005Jerome P. Reiter
Methods tailored specifically to data altered by recoding and topcoding variables, data swapping, or adding random noise that agencies can use to assess threats from intruders who possess information on relationships among variables and the methods of data alteration are described.
Journal of the American Statistical AssociationDisclosure Control of Microdata
125 Citations1990Jelke Bethlehem, Wouter J. Keller +1 more
Journal of the American Statistical AssociationOptimal Disclosure Limitation Strategy in Statistical Databases: Deterring Tracker Attacks through Additive Noise
79 Citations2000George T. Duncan, Sumitra Mukherjee
Conditions under which an attack by a data snooper is better thwarted by a combination of query restriction and data masking than by either disclosure limitation method separately are derived.
Lecture notes in computer scienceOn the Security of Noise Addition for Privacy in Statistical Databases
76 Citations2004Josep Domingo‐Ferrer, Francesc Sebé +1 more
This paper is a critical analysis of the security of the methods in that family of methods used in the protection of the privacy of individual data (microdata) in statistical databases.
Statistics and ComputingDisclosure risk assessment in statistical microdata protection via advanced record linkage
71 Citations2003Josep Domingo‐Ferrer, Vicenç Torra
This paper reviews conventional record linkage, which assumes shared variables between the external and the protected data sets, and then shows that record linkage—and thus disclosure—is still possible without shared variables.
Journal of Statistical Planning and InferenceSignificance tests for multi-component estimands from multiply imputed, synthetic microdata
65 Citations2004Jerome P. Reiter
Lecture notes in computer scienceMasking and Re-identification Methods for Public-Use Microdata: Overview and Research Problems
65 Citations2004William E. Winkler
This paper provides an overview of methods of masking microdata so that the data can be placed in public-use files and indicates where the data may have limitations for most analyses and how re-identification might or can be performed.
Journal of Statistical Planning and InferenceOptimal noise addition for preserving confidentiality in multivariate data
53 Citations1991Patrick Tendick
These measures of confidentiality and data integrity for multivariate data when noise has been added are evaluated in the case that the vectors of added noise have independent components and it is demonstrated that the amount of protection provided could be relatively low for data of high dimension.
CHANCEDisclosure Risk vs. Data Utility: The R-U Confidentiality Map as Applied to Topcoding
51 Citations2004George T. Duncan, S. Lynne Stokes
Maintaining a long tradition, government statistical agencies, such as the U.S. Census Bureau and the National Center for Education Statistics, have established policies to protect the privacy of respondents and maintain the confidentiality of the data they collect.
Lecture notes in computer scienceOutlier Protection in Continuous Microdata Masking
41 Citations2004Josep M. Mateo‐Sanz, Francesc Sebé +1 more
The objective of this work is to compare, for different masking methods, the information loss and disclosure risk related to outliers in the original data set and to evaluate the protection level offered by different masked methods to extreme individuals.
Statistica NeerlandicaStrategies for measuring risk in public use microdata files
41 Citations1992Brian V. Greenberg, Laura Zayatz
Lecture notes in computer scienceFast Generation of Accurate Synthetic Microdata
36 Citations2004Josep M. Mateo‐Sanz, Antoni Martínez-Ballesté +1 more
A new method for generating continuous synthetic microdata is proposed, where the covariance matrix and the univariate statistics of the original data set are exactly preserved.
International Journal of Uncertainty Fuzziness and Knowledge-Based SystemsON THE SECURITY OF MICROAGGREGATION WITH INDIVIDUAL RANKING: ANALYTICAL ATTACKS
33 Citations2002Josep Domingo‐Ferrer, Anna Oganian +2 more
It is shown in this paper how to find interval estimates for the original data based on the microaggregated data, which can be considerably narrower than intervals resulting from subtraction of means, and can be useful to detect lack of security in a microaggRegated data set.
On privacy-preserving access to distributed heterogeneous healthcare information
29 Citations2004Claus Boyens, Ramayya Krishnan +1 more
This paper focuses on mediator-based architectures and the privacy problems that arise in the healthcare context owing to the linkage of information about patients, physicians, and diseases enabled by the mediator and proposes an "audit and aggregate" methodology that chooses the optimal level of aggregation of the data.
International Statistical ReviewSafe Data versus Safe Settings: Access to Microdata from the British Census
27 Citations1994Catherine Marsh, Angela Dale +1 more
Journal of the Royal Statistical Society Series A (Statistics in Society)Proposals for 2001 Samples of Anonymized Records: An Assessment of Disclosure Risk
17 Citations2001Angela Dale, Mark Elliot
Sociological Methods & ResearchIdentification Risks of Microdata
16 Citations1995Walter Müller, Uwe Blien +1 more
The risks of discovering the identity of data providers even in cases where the microdata have been previously anonymized by some procedure are examined, and the results show that the identification risk is smaller than has been assumed by previous research.
Elsevier eBooksConfidentiality and Statistical Disclosure Limitation
11 Citations2001George T. Duncan
