Sequence Determinants of Polyadenylation-mediated Regulation
Overview
Affiliations
The cleavage and polyadenylation reaction is a crucial step in transcription termination and pre-mRNA maturation in human cells. Despite extensive research, the encoding of polyadenylation-mediated regulation of gene expression within the DNA sequence is not well understood. Here, we utilized a massively parallel reporter assay to inspect the effect of over 12,000 rationally designed polyadenylation sequences (PASs) on reporter gene expression and cleavage efficiency. We find that the PAS sequence can modulate gene expression by over five orders of magnitude. By using a uniquely designed scanning mutagenesis data set, we gain mechanistic insight into various modes of action by which the cleavage efficiency affects the sensitivity or robustness of the PAS to mutation. Furthermore, we employ motif discovery to identify both known and novel sequence motifs associated with PAS-mediated regulation. By leveraging the large scale of our data, we train a deep learning model for the highly accurate prediction of RNA levels from DNA sequence alone ( = 0.83). Moreover, we devise unique approaches for predicting exact cleavage sites for our reporter constructs and for endogenous transcripts. Taken together, our results expand our understanding of PAS-mediated regulation, and provide an unprecedented resource for analyzing and predicting PAS for regulatory genomics applications.
Predicting RNA-seq coverage from DNA sequence as a unifying model of gene regulation.
Linder J, Srivastava D, Yuan H, Agarwal V, Kelley D Nat Genet. 2025; .
PMID: 39779956 DOI: 10.1038/s41588-024-02053-6.
Decoding biology with massively parallel reporter assays and machine learning.
La Fleur A, Shi Y, Seelig G Genes Dev. 2024; 38(17-20):843-865.
PMID: 39362779 PMC: 11535156. DOI: 10.1101/gad.351800.124.
Delineating yeast cleavage and polyadenylation signals using deep learning.
Stroup E, Ji Z Genome Res. 2024; 34(7):1066-1080.
PMID: 38914436 PMC: 11368178. DOI: 10.1101/gr.278606.123.
Stroup E, Ji Z Nat Commun. 2023; 14(1):7378.
PMID: 37968271 PMC: 10651852. DOI: 10.1038/s41467-023-43266-3.
Stress responses of plants through transcriptome plasticity by mRNA alternative polyadenylation.
Zhou J, Li Q Mol Hortic. 2023; 3(1):19.
PMID: 37789388 PMC: 10536700. DOI: 10.1186/s43897-023-00066-z.