Lambert, Samuel A, Yang, Ally, Sasse, Alexander, Cowley, Gwendolyn ORCID: 0000-0002-8505-1354, Albu, Mihai, Caddick, Mark X, Morris, Quaid D, Weirauch, Matthew T and Hughes, Timothy R
(2019)
Similarity Regression predicts evolution of transcription factor sequence specificity.
Nature Genetics.
There is a more recent version of this item available. |
Text
53077_1_merged_1549744850.pdf - Author Accepted Manuscript Access to this file is restricted: awaiting official publication and publisher embargo. Download (10MB) |
Abstract
Transcription factor (TF) binding specificities (motifs) are essential to the analysis of noncoding DNA and gene regulation. Accurate prediction of the sequence specificities of TFs is critical, because the hundreds of sequenced eukaryotic genomes encompass hundreds of thousands of TFs, and assaying each is currently infeasible. There is ongoing controversy regarding the efficacy of motif prediction methods, as well as the degree of motif diversification among related species. Here, we describe Similarity Regression (SR), a significantly improved method for predicting motifs. We have updated and expanded the Cis-BP database using SR, and validate its predictive capacity with new data from diverse eukaryotic TFs. SR inherently quantifies TF motif evolution, and we show that previous claims of near-complete conservation of motifs between human and Drosophila are grossly inflated, with nearly half the motifs in each species absent from the other. We conclude that diversification in DNA binding motifs is pervasive, and present a new tool and updated resource to study TF diversity and gene regulation across eukaryotes.
Item Type: | Article |
---|---|
Depositing User: | Symplectic Admin |
Date Deposited: | 12 Mar 2019 10:36 |
Last Modified: | 19 Jan 2023 00:57 |
URI: | https://livrepository.liverpool.ac.uk/id/eprint/3034076 |
Available Versions of this Item
- Similarity Regression predicts evolution of transcription factor sequence specificity. (deposited 12 Mar 2019 10:36) [Currently Displayed]