Unknown

Dataset Information

0

Scalable Bayesian variable selection for structured high-dimensional data.


ABSTRACT: Variable selection for structured covariates lying on an underlying known graph is a problem motivated by practical applications, and has been a topic of increasing interest. However, most of the existing methods may not be scalable to high-dimensional settings involving tens of thousands of variables lying on known pathways such as the case in genomics studies. We propose an adaptive Bayesian shrinkage approach which incorporates prior network information by smoothing the shrinkage parameters for connected variables in the graph, so that the corresponding coefficients have a similar degree of shrinkage. We fit our model via a computationally efficient expectation maximization algorithm which scalable to high-dimensional settings ( p ? 100 , 000 ). Theoretical properties for fixed as well as increasing dimensions are established, even when the number of variables increases faster than the sample size. We demonstrate the advantages of our approach in terms of variable selection, prediction, and computational scalability via a simulation study, and apply the method to a cancer genomics study.

SUBMITTER: Chang C 

PROVIDER: S-EPMC6222001 | biostudies-literature | 2018 Dec

REPOSITORIES: biostudies-literature

altmetric image

Publications

Scalable Bayesian variable selection for structured high-dimensional data.

Chang Changgee C   Kundu Suprateek S   Long Qi Q  

Biometrics 20180508 4


Variable selection for structured covariates lying on an underlying known graph is a problem motivated by practical applications, and has been a topic of increasing interest. However, most of the existing methods may not be scalable to high-dimensional settings involving tens of thousands of variables lying on known pathways such as the case in genomics studies. We propose an adaptive Bayesian shrinkage approach which incorporates prior network information by smoothing the shrinkage parameters f  ...[more]

Similar Datasets

| S-EPMC5885321 | biostudies-literature
| S-EPMC10843621 | biostudies-literature
| S-EPMC5891168 | biostudies-literature
| S-EPMC3587767 | biostudies-literature
| S-EPMC7487595 | biostudies-literature
| S-EPMC7133715 | biostudies-literature
| S-EPMC9789572 | biostudies-literature
| S-EPMC3478096 | biostudies-literature
| S-EPMC4848399 | biostudies-literature
| S-EPMC9545322 | biostudies-literature