Article,

Space is the Place: Effects of Continuous Spatial Structure on Analysis of Population Genetic Data

C. Battey, P. Ralph, and A. Kern.
Genetics, 215 (1): 193-214 (May 2020)
DOI: 10.1534/genetics.120.303143

Abstract

Real geography is continuous, but standard models in population genetics are based on discrete, well-mixed populations. As a result, many methods of analyzing genetic data assume that samples are a random draw from a well-mixed population, but are applied to clustered samples from populations that are structured clinally over space. Here, we use simulations of populations living in continuous geography to study the impacts of dispersal and sampling strategy on population genetic summary statistics, demographic inference, and genome-wide association studies (GWAS). We find that most common summary statistics have distributions that differ substantially from those seen in well-mixed populations, especially when Wright’s neighborhood size is \\< 100 and sampling is spatially clustered. “Stepping-stone” models reproduce some of these effects, but discretizing the landscape introduces artifacts that in some cases are exacerbated at higher resolutions. The combination of low dispersal and clustered sampling causes demographic inference from the site frequency spectrum to infer more turbulent demographic histories, but averaged results across multiple simulations revealed surprisingly little systematic bias. We also show that the combination of spatially autocorrelated environments and limited dispersal causes GWAS to identify spurious signals of genetic association with purely environmentally determined phenotypes, and that this bias is only partially corrected by regressing out principal components of ancestry. Last, we discuss the relevance of our simulation results for inference from genetic variation in real organisms.

BibTeX key: battey2020space
entry type: article
year: 2020
month: 05
journal: Genetics
number: 1
pages: 193-214
volume: 215
eprint: https://academic.oup.com/genetics/article-pdf/215/1/193/42100884/genetics0193.pdf
issn: 1943-2631
DOI: 10.1534/genetics.120.303143
url: https://doi.org/10.1534/genetics.120.303143

Users

Comments and Reviewsshow / hide

Please log in to take part in the discussion (add own reviews or comments).

Cite this publication

%0 Journal Article %1 battey2020space %A Battey, C J %A Ralph, Peter L %A Kern, Andrew D %D 2020 %J Genetics %K myown spatial_population_genetics spatial_simulations %N 1 %P 193-214 %R 10.1534/genetics.120.303143 %T Space is the Place: Effects of Continuous Spatial Structure on Analysis of Population Genetic Data %U https://doi.org/10.1534/genetics.120.303143 %V 215 %X Real geography is continuous, but standard models in population genetics are based on discrete, well-mixed populations. As a result, many methods of analyzing genetic data assume that samples are a random draw from a well-mixed population, but are applied to clustered samples from populations that are structured clinally over space. Here, we use simulations of populations living in continuous geography to study the impacts of dispersal and sampling strategy on population genetic summary statistics, demographic inference, and genome-wide association studies (GWAS). We find that most common summary statistics have distributions that differ substantially from those seen in well-mixed populations, especially when Wright’s neighborhood size is \\< 100 and sampling is spatially clustered. “Stepping-stone” models reproduce some of these effects, but discretizing the landscape introduces artifacts that in some cases are exacerbated at higher resolutions. The combination of low dispersal and clustered sampling causes demographic inference from the site frequency spectrum to infer more turbulent demographic histories, but averaged results across multiple simulations revealed surprisingly little systematic bias. We also show that the combination of spatially autocorrelated environments and limited dispersal causes GWAS to identify spurious signals of genetic association with purely environmentally determined phenotypes, and that this bias is only partially corrected by regressing out principal components of ancestry. Last, we discuss the relevance of our simulation results for inference from genetic variation in real organisms.

BibSonomy

Space is the Place: Effects of Continuous Spatial Structure on Analysis of Population Genetic Data

Abstract

Tags

Users

Comments and Reviewsshow / hide

Cite this publication

More citation styles

search on