Conflation of short identity-by-descent segments bias their inferred   length distribution

Charleston W.K. Chiang; Peter Ralph; John Novembre

arXiv:1410.5313·q-bio.PE·December 24, 2020

Conflation of short identity-by-descent segments bias their inferred length distribution

Charleston W.K. Chiang, Peter Ralph, John Novembre

PDF

TL;DR

This study reveals that a significant portion of inferred IBD segments are conflations of shorter segments, which biases length distribution estimates and impacts downstream genetic analyses, emphasizing the need for correction methods.

Contribution

The paper quantifies the conflation effect in IBD detection, demonstrating its prevalence across methods and its impact on genetic inference accuracy.

Findings

01

Nearly 40% of 1-2cM IBD segments are conflations of shorter segments

02

Segments longer than 2cM are less affected by conflation (<5%)

03

Conflation biases estimates of mutation rates from IBD data

Abstract

Identity-by-descent (IBD) is a fundamental concept in genetics with many applications. In a common definition, two haplotypes are said to contain an IBD segment if they share a segment that is inherited from a recent shared common ancestor without intervening recombination. Long IBD segments (> 1cM) can be efficiently detected by a number of algorithms using high-density SNP array data from a population sample. However, these approaches detect IBD based on contiguous segments of identity-by-state, and such segments may exist due to the conflation of smaller, nearby IBD segments. We quantified this effect using coalescent simulations, finding that nearly 40% of inferred segments 1-2cM long are results of conflations of two or more shorter segments, under demographic scenarios typical for modern humans. This biases the inferred IBD segment length distribution, and so can affect downstream…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.