Unknown

Dataset Information

0

Selecting SNPs informative for African, American Indian and European Ancestry: application to the Family Investigation of Nephropathy and Diabetes (FIND).


ABSTRACT: BACKGROUND:The presence of population structure in a sample may confound the search for important genetic loci associated with disease. Our four samples in the Family Investigation of Nephropathy and Diabetes (FIND), European Americans, Mexican Americans, African Americans, and American Indians are part of a genome- wide association study in which population structure might be particularly important. We therefore decided to study in detail one component of this, individual genetic ancestry (IGA). From SNPs present on the Affymetrix 6.0 Human SNP array, we identified 3 sets of ancestry informative markers (AIMs), each maximized for the information in one the three contrasts among ancestral populations: Europeans (HAPMAP, CEU), Africans (HAPMAP, YRI and LWK), and Native Americans (full heritage Pima Indians). We estimate IGA and present an algorithm for their standard errors, compare IGA to principal components, emphasize the importance of balancing information in the ancestry informative markers (AIMs), and test the association of IGA with diabetic nephropathy in the combined sample. RESULTS:A fixed parental allele maximum likelihood algorithm was applied to the FIND to estimate IGA in four samples: 869 American Indians; 1385 African Americans; 1451 Mexican Americans; and 826 European Americans. When the information in the AIMs is unbalanced, the estimates are incorrect with large error. Individual genetic admixture is highly correlated with principle components for capturing population structure. It takes ~700 SNPs to reduce the average standard error of individual admixture below 0.01. When the samples are combined, the resulting population structure creates associations between IGA and diabetic nephropathy. CONCLUSIONS:The identified set of AIMs, which include American Indian parental allele frequencies, may be particularly useful for estimating genetic admixture in populations from the Americas. Failure to balance information in maximum likelihood, poly-ancestry models creates biased estimates of individual admixture with large error. This also occurs when estimating IGA using the Bayesian clustering method as implemented in the program STRUCTURE. Odds ratios for the associations of IGA with disease are consistent with what is known about the incidence and prevalence of diabetic nephropathy in these populations.

SUBMITTER: Williams RC 

PROVIDER: S-EPMC4855449 | biostudies-literature | 2016 May

REPOSITORIES: biostudies-literature

altmetric image

Publications

Selecting SNPs informative for African, American Indian and European Ancestry: application to the Family Investigation of Nephropathy and Diabetes (FIND).

Williams Robert C RC   Elston Robert C RC   Kumar Pankaj P   Knowler William C WC   Abboud Hanna E HE   Adler Sharon S   Bowden Donald W DW   Divers Jasmin J   Freedman Barry I BI   Igo Robert P RP   Ipp Eli E   Iyengar Sudha K SK   Kimmel Paul L PL   Klag Michael J MJ   Kohn Orly O   Langefeld Carl D CD   Leehey David J DJ   Nelson Robert G RG   Nicholas Susanne B SB   Pahl Madeleine V MV   Parekh Rulan S RS   Rotter Jerome I JI   Schelling Jeffrey R JR   Sedor John R JR   Shah Vallabh O VO   Smith Michael W MW   Taylor Kent D KD   Thameem Farook F   Thornley-Brown Denyse D   Winkler Cheryl A CA   Guo Xiuqing X   Zager Phillip P   Hanson Robert L RL  

BMC genomics 20160504


<h4>Background</h4>The presence of population structure in a sample may confound the search for important genetic loci associated with disease. Our four samples in the Family Investigation of Nephropathy and Diabetes (FIND), European Americans, Mexican Americans, African Americans, and American Indians are part of a genome- wide association study in which population structure might be particularly important. We therefore decided to study in detail one component of this, individual genetic ancest  ...[more]

Similar Datasets

2022-02-16 | PXD029323 | Pride
| S-EPMC3141729 | biostudies-literature
| S-EPMC7860026 | biostudies-literature
| S-EPMC8753123 | biostudies-literature
| S-EPMC3476707 | biostudies-literature
| S-EPMC5495141 | biostudies-literature
| S-EPMC3072384 | biostudies-literature
| S-EPMC3046166 | biostudies-literature
| S-EPMC3899836 | biostudies-literature
2012-05-23 | GSE28000 | GEO