Ontology highlight
ABSTRACT: Motivation
Classification of individuals into disease or clinical categories from high-dimensional biological data with low prediction error is an important challenge of statistical learning in bioinformatics. Feature selection can improve classification accuracy but must be incorporated carefully into cross-validation to avoid overfitting. Recently, feature selection methods based on differential privacy, such as differentially private random forests and reusable holdout sets, have been proposed. However, for domains such as bioinformatics, where the number of features is much larger than the number of observations p≫n , these differential privacy methods are susceptible to overfitting.Methods
We introduce private Evaporative Cooling, a stochastic privacy-preserving machin
SUBMITTER: Le TT
PROVIDER: S-EPMC5870708 | biostudies-literature | 2017 Sep
REPOSITORIES: biostudies-literature