Functional constraint and small insertions and deletions in the ENCODE regions of the human genome.


Clark, TG; Andrew, T; Cooper, GM; Margulies, EH; Mullikin, JC; Balding, DJ; (2007) Functional constraint and small insertions and deletions in the ENCODE regions of the human genome. Genome biology, 8 (9). R180. ISSN 1465-6906 DOI: https://doi.org/10.1186/gb-2007-8-9-r180

[img]
Preview
Text - Published Version
License:

Download (395kB) | Preview

Abstract

BACKGROUND: We describe the distribution of indels in the 44 Encyclopedia of DNA Elements (ENCODE) regions (about 1% of the human genome) and evaluate the potential contributions of small insertion and deletion polymorphisms (indels) to human genetic variation. We relate indels to known genomic annotation features and measures of evolutionary constraint. RESULTS: Indel rates are observed to be reduced approximately 20-fold to 60-fold in exonic regions, 5-fold to 10-fold in sequence that exhibits high evolutionary constraint in mammals, and up to 2-fold in some classes of regulatory elements (for instance, formaldehyde assisted isolation of regulatory elements [FAIRE] and hypersensitive sites). In addition, some noncoding transcription and other chromatin mediated regulatory sites also have reduced indel rates. Overall indel rates for these data are estimated to be smaller than single nucleotide polymorphism (SNP) rates by a factor of approximately 2, with both rates measured as base pairs per 100 kilobases to facilitate comparison. CONCLUSION: Indel rates exhibit a broadly similar distribution across genomic features compared with SNP density rates, with a reduction in rates in coding transcription and evolutionarily constrained sequence. However, unlike indels, SNP rates do not appear to be reduced in some noncoding functional sequences, such as pseudo-exons, and FAIRE and hypersensitive sites. We conclude that indel rates are greatly reduced in transcribed and evolutionarily constrained DNA, and discuss why indel (but not SNP) rates appear to be constrained at some regulatory sites.

Item Type: Article
Faculty and Department: Faculty of Epidemiology and Population Health > Dept of Infectious Disease Epidemiology
Faculty of Infectious and Tropical Diseases > Dept of Pathogen Molecular Biology
PubMed ID: 17784950
Web of Science ID: 252100800006
URI: http://researchonline.lshtm.ac.uk/id/eprint/4446

Statistics


Download activity - last 12 months
Downloads since deposit
275Downloads
297Hits
Accesses by country - last 12 months
Accesses by referrer - last 12 months
Impact and interest
Additional statistics for this record are available via IRStats2

Actions (login required)

Edit Item Edit Item