Quantitative modeling of transcription factor binding specificities using DNA shape
Abstract
Genomes provide an abundance of putative binding sites for each transcription factor (TF). However, only small subsets of these potential targets are functional. TFs of the same protein family bind to target sites that are very similar but not identical. This distinction allows closely related TFs to regulate different genes and thus execute distinct functions. Because the nucleotide sequence of the core motif is often not sufficient for identifying a genomic target, we refined the description of TF binding sites by introducing a combination of DNA sequence and shape features, which consistently improved the modeling of in vitro TF-DNA binding specificities. Although additional factors affect TF binding in vivo, shape-augmented models reveal binding specificity mechanisms that are not apparent from sequence alone.
- Publication:
-
Proceedings of the National Academy of Science
- Pub Date:
- April 2015
- DOI:
- 10.1073/pnas.1422023112
- Bibcode:
- 2015PNAS..112.4654Z