Locating protein-coding sequences under selection for additional, overlapping functions in 29 mammalian genomes

Michael F Lin, Pouya Kheradpour, Stefan Washietl, Brian J Parker, Jakob S Pedersen, Manolis Kellis

64 Citations (Scopus)

Abstract

The degeneracy of the genetic code allows protein-coding DNA and RNA sequences to simultaneously encode additional, overlapping functional elements. A sequence in which both protein-coding and additional overlapping functions have evolved under purifying selection should show increased evolutionary conservation compared to typical protein-coding genes-especially at synonymous sites. In this study, we use genome alignments of 29 placental mammals to systematically locate short regions within human ORFs that show conspicuously low estimated rates of synonymous substitution across these species. The 29-species alignment provides statistical power to locate more than 10,000 such regions with resolution down to nine-codon windows, which are found within more than a quarter of all human protein-coding genes and contain ~2% of their synonymous sites. We collect numerous lines of evidence that the observed synonymous constraint in these regions reflects selection on overlapping functional elements including splicing regulatory elements, dual-coding genes, RNA secondary structures, microRNA target sites, and developmental enhancers. Our results show that overlapping functional elements are common in mammalian genes, despite the vast genomic landscape.
Original languageEnglish
JournalGenome Research
Volume21
Issue number11
Pages (from-to)1916-28
Number of pages13
ISSN1088-9051
DOIs
Publication statusPublished - Nov 2011

Fingerprint

Dive into the research topics of 'Locating protein-coding sequences under selection for additional, overlapping functions in 29 mammalian genomes'. Together they form a unique fingerprint.

Cite this