Computational disclosure control: a primer on data privacy protection
DSpace@MIT (Massachusetts Institute of Technology)Published 1 January 2001Open access
Latanya Sweeney, Hal Abelson
Citations139
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
Abstract
Thesis (Ph. D.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 2001.
Keywords
Computer Science
Choice Reviews OnlineArtificial intelligence: a modern approach
22,205 Citations1995Stuart Russell, Peter Norvig +2 more
IEEE Transactions on Information TheoryNearest neighbor pattern classification
16,134 Citations1967Thomas M. Cover, Peter E. Hart
The nearest neighbor decision rule assigns to an unclassified sample point the classification of the nearest of a set of previously classified points, so it may be said that half the classification information in an infinite sample set is contained in the nearest neighbor.
Pattern classification and scene analysis
12,643 Citations1973Richard O. Duda, Peter E. Hart
Choice Reviews OnlinePrinciples of database and knowledge-base systems
2,701 Citations1989
This book goes into the details of database conception and use, it tells you everything on relational databases from theory to the actual used algorithms.
Calhoun: The Naval Postgraduate School Institutional Archive (Naval Postgraduate School)Cryptography and data security
1,911 Citations1982Dorothy E. Denning
The goal of this book is to introduce the mathematical principles of data security and to show how these principles apply to operating systems, database systems, and computer networks.
Artificial Intelligence ReviewLocally Weighted Learning
1,684 Citations1997Christopher G. Atkeson, Andrew Moore +1 more
The survey discusses distance functions, smoothing parameters, weighting functions, local model structures, regularization of the estimates and bias, assessing predictions, handling noisy data and outliers, improving the quality of predictions by tuning fit parameters, and applications of locally weighted learning.
Elsevier eBooksEfficient Algorithms for Minimizing Cross Validation Error
227 Citations1994Andrew Moore, Mary S. Lee
It is shown how experimental design methods can achieve this, using a technique similar to a Bayesian version of Kaelbling's Interval Estimation, to reduce the computational burden of large scale cross validation searches.
Artificial Intelligence in MedicineAn evaluation of machine-learning methods for predicting pneumonia mortality
212 Citations1997Gregory Cooper, Constantin Aliferis +13 more
The models are distinguished more by the number of variables and parameters that they contain than by their error rates; these differences suggest which models may be the most amenable to future implementation as paper-based guidelines.
ACM Transactions on Database SystemsThe tracker
207 Citations1979Dorothy E. Denning, Peter J. Denning
It is shown that the compromise of small query sets can in fact almost always be accomplished with the help of characteristic formulas called trackers, and security is not guaranteed by the lack of a general tracker.
Journal of the American Statistical AssociationOn the Question of Statistical Confidentiality
174 Citations1972Ivan P. Fellegi
The nature of statistical confidentiality is explored, its essential role in the collection of data by statistical offices, its relationship to privacy and the need for increased attention to potential statistical disclosures because of the increased tabulation and dissemination capabilities of statistical offices are explored.
A Multilevel Relational Data Model
165 Citations1987Dorothy E. Denning, Teresa F. Lunt
The model is defined in terms of the standard relational model, but lends itself to a design and implementation that offers a high level of assurance for mandatory security.
Statistical ScienceEnhancing Access to Microdata While Protecting Confidentiality: Prospects for the Future
160 Citations1991George T. Duncan, R. W. Pearson
This article presents a scenario for the future of research access to federally collected microdata, as they relate to improvements in database techniques, computer and analytical method- ologies and legal and administrative arrangements for access to and protection of federal statistics.
Security and inference in multilevel database and knowledge-base systems
118 Citations1987Matthew Morgenstern
This paper establishes a framework for studying these inference control problems, describes a representation for relevant semantics of the application, develops criteria for safety and security of a system to prevent these problems, and outlines algorithms for enforcing these criteria.
JAMAPrivacy and Security of Personal Information in a New Health Care System
110 Citations1993Lawrence O. Gostin
A COMPLEX health care information infrastructure will exist under a reformed health care system as proposed in the American Health Security Act of 1993 and the success of the new system will depend in part on the accuracy, correctness, and trustworthiness of the information and the privacy rights of individuals to control the disclosure of personal information.
Inference aggregation detection in database management systems
110 Citations2003Thomas H. Hinke
An algorithm for processing the semantic relationship graph to discover whether potential inference aggregation problems exist and the use of set theory and the addition of set operations to the DBMS to permit the description of aggregation detection queries are presented.
Aggregation and inference: facts and fallacies
76 Citations2003Teresa F. Lunt
It is shown that sensitive associations among entities of different types are best treated by representing the sensitive association separately and classifying the individual entities low and the relationship high, and the suggested approaches allow the mandatory reference monitor to protect the sensitive associations.
IEEE Transactions on Knowledge and Data EngineeringControlling FD and MVD inferences in multilevel relational database systems
71 Citations1991Tzong-An Su, Gültekin Özsoyoğlu
It is proven that incurring minimum information loss to prevent compromises is an NP-complete problem and an exact algorithm to adjust the attribute levels so that no compromise due to functional dependencies occurs is given.
Detection and elimination of inference channels in multilevel relational database systems
55 Citations2002Xiaolei Qian, M.E. Stickel +3 more
A global optimization approach to upgrading is suggested to block a set of inference problems that allows upgrade costs to be considered, and supports security categories as well as levels.
PubMedSharing electronic medical records across multiple heterogeneous and competing institutions.
55 Citations1996Isaac S. Kohane, F J van Wingerde +9 more
W3-EMRS, a multi-institutional architecture, and its implementation are described and Thorny problems in data sharing underlined by the W3- EMRS project are reviewed.
Modeling security-relevant data semantics
39 Citations1990Gary W. Smith
The use of an extended data model that represents both integrity and secrecy aspects of data is presented and can be used as a database design tool and as a vehicle by which domain experts, database designers, and security officers can precisely define the security requirements for an application domain.
Microdata disclosure limitation in statistical databases: query size and random sample query control
24 Citations2002George T. Duncan, Sumitra Mukherjee
A probabilistic framework is used to assess the strengths and weaknesses of two existing disclosure control mechanisms and an alternative scheme combining query set size restriction and random sample query control results in a significant decrease in the risk of disclosure.
Catalytic inference analysis: detecting inference threats due to knowledge discovery
23 Citations2002John Hale, Sujeet Shenoi
This paper presents a formalism for modeling and analyzing catalytic inference in "mixed" databases containing various precise, imprecise and fuzzy relations, which is flexible and robust, and well-suited to implementation.
Abductive and approximate reasoning models for characterizing inference channels
17 Citations2002Thomas D. Garvey, Teresa F. Lunt +1 more
The authors introduce abductive reasoning, which provides a framework for reasoning with approximate and uncertain information, which enables them to extend the model for inference channels by taking into account the likelihood that a person might believe some statement of interest.
