Unknown

Dataset Information

0

Genetic structure of the Han Chinese population revealed by genome-wide SNP variation.


ABSTRACT: Population stratification is a potential problem for genome-wide association studies (GWAS), confounding results and causing spurious associations. Hence, understanding how allele frequencies vary across geographic regions or among subpopulations is an important prelude to analyzing GWAS data. Using over 350,000 genome-wide autosomal SNPs in over 6000 Han Chinese samples from ten provinces of China, our study revealed a one-dimensional "north-south" population structure and a close correlation between geography and the genetic structure of the Han Chinese. The north-south population structure is consistent with the historical migration pattern of the Han Chinese population. Metropolitan cities in China were, however, more diffused "outliers," probably because of the impact of modern migration of peoples. At a very local scale within the Guangdong province, we observed evidence of population structure among dialect groups, probably on account of endogamy within these dialects. Via simulation, we show that empirical levels of population structure observed across modern China can cause spurious associations in GWAS if not properly handled. In the Han Chinese, geographic matching is a good proxy for genetic matching, particularly in validation and candidate-gene studies in which population stratification cannot be directly accessed and accounted for because of the lack of genome-wide data, with the exception of the metropolitan cities, where geographical location is no longer a good indicator of ancestral origin. Our findings are important for designing GWAS in the Chinese population, an activity that is expected to intensify greatly in the near future.

SUBMITTER: Chen J 

PROVIDER: S-EPMC2790583 | biostudies-literature | 2009 Dec

REPOSITORIES: biostudies-literature

altmetric image

Publications

Genetic structure of the Han Chinese population revealed by genome-wide SNP variation.

Chen Jieming J   Zheng Houfeng H   Bei Jin-Xin JX   Sun Liangdan L   Jia Wei-hua WH   Li Tao T   Zhang Furen F   Seielstad Mark M   Zeng Yi-Xin YX   Zhang Xuejun X   Liu Jianjun J  

American journal of human genetics 20091201 6


Population stratification is a potential problem for genome-wide association studies (GWAS), confounding results and causing spurious associations. Hence, understanding how allele frequencies vary across geographic regions or among subpopulations is an important prelude to analyzing GWAS data. Using over 350,000 genome-wide autosomal SNPs in over 6000 Han Chinese samples from ten provinces of China, our study revealed a one-dimensional "north-south" population structure and a close correlation b  ...[more]

Similar Datasets

| S-EPMC2735092 | biostudies-literature
| S-EPMC2652362 | biostudies-literature
| S-EPMC5498554 | biostudies-literature
| S-EPMC10919046 | biostudies-literature
| S-EPMC10445094 | biostudies-literature
| S-EPMC3567019 | biostudies-literature
| S-EPMC8291712 | biostudies-literature
| S-EPMC7731653 | biostudies-literature
| S-EPMC3840132 | biostudies-literature
| S-EPMC9580327 | biostudies-literature