>
Fa   |   Ar   |   En
   A machine learning approach for identifying novel cell type-specific transcriptional regulators of myogenesis  
   
نویسنده busser b.w. ,taher l. ,kim y. ,tansey t. ,bloom m.j. ,ovcharenko i. ,michelson a.m.
منبع plos genetics - 2012 - دوره : 8 - شماره : 3
چکیده    Transcriptional enhancers integrate the contributions of multiple classes of transcription factors (tfs) to orchestrate the myriad spatio-temporal gene expression programs that occur during development. a molecular understanding of enhancers with similar activities requires the identification of both their unique and their shared sequence features. to address this problem,we combined phylogenetic profiling with a dna-based enhancer sequence classifier that analyzes the tf binding sites (tfbss) governing the transcription of a co-expressed gene set. we first assembled a small number of enhancers that are active in drosophila melanogaster muscle founder cells (fcs) and other mesodermal cell types. using phylogenetic profiling,we increased the number of enhancers by incorporating orthologous but divergent sequences from other drosophila species. functional assays revealed that the diverged enhancer orthologs were active in largely similar patterns as their d. melanogaster counterparts,although there was extensive evolutionary shuffling of known tfbss. we then built and trained a classifier using this enhancer set and identified additional related enhancers based on the presence or absence of known and putative tfbss. predicted fc enhancers were over-represented in proximity to known fc genes; and many of the tfbss learned by the classifier were found to be critical for enhancer activity,including pou homeodomain,myb,ets,forkhead,and t-box motifs. empirical testing also revealed that the t-box tf encoded by org-1 is a previously uncharacterized regulator of muscle cell identity. finally,we found extensive diversity in the composition of tfbss within known fc enhancers,suggesting that motif combinatorics plays an essential role in the cellular specificity exhibited by such enhancers. in summary,machine learning combined with evolutionary sequence analysis is useful for recognizing novel tfbss and for facilitating the identification of cognate tfs that coordinate cell type-specific developmental gene expression patterns.
آدرس laboratory of developmental systems biology,national heart,lung,and blood institute,national institutes of health,bethesda,md, United States, computational biology branch,national center for biotechnology information,national library of medicine,national institutes of health,bethesda,md, United States, laboratory of developmental systems biology,national heart,lung,and blood institute,national institutes of health,bethesda,md, United States, laboratory of developmental systems biology,national heart,lung,and blood institute,national institutes of health,bethesda,md, United States, laboratory of developmental systems biology,national heart,lung,and blood institute,national institutes of health,bethesda,md, United States, computational biology branch,national center for biotechnology information,national library of medicine,national institutes of health,bethesda,md, United States, laboratory of developmental systems biology,national heart,lung,and blood institute,national institutes of health,bethesda,md, United States
 
     
   
Authors
  
 
 

Copyright 2023
Islamic World Science Citation Center
All Rights Reserved