GEO series
Iterative deep learning-design of human enhancers exploits condensed sequence grammar to achieve cell type-specificity [RNA-Seq]
GSE269036
Homo sapiens; synthetic construct
Expression profiling by high throughput sequencing; Other
16 samples
2024/06/14
GPL27609GPL21697
Summary
An important and largely unsolved problem in synthetic biology is how to target gene expression to specific cell types. Here, we apply iterative deep learning to design synthetic enhancers with strong differential activity between two human cell lines. We initially train models on published datasets of enhancer activity and chromatin accessibility and use them to guide the design of synthetic enhancers that maximize predicted specificity. We experimentally validate these sequences, use the measurements to re-optimize the predictor, and design a second generation of enhancers with improved specificity. Our design methods embed relevant transcription factor binding site (TFBS) motifs with higher frequencies than comparable endogenous enhancers while using a more selective motif vocabulary, and we show that enhancer activity is correlated with transcription factor expression at the single cell level. Finally, we characterize causal features of top enhancers via perturbation experiments and show enhancers as short as 50bp can maintain specificity.
Download
NCBI GEO page ↗
Paper (PMID 40472848) ↗
{# Names what the click gives you. "Open in finder" meant nothing to a
visitor who arrived from a search engine and has never seen the tool. #}
Find more
RNA-seq datasets →
Similar datasets
- GSE272887 In vivo double knockout CAR-T screen identifies synergistic gene pairs that enhance anti-tumor immunity. 67 samples
- GSE243312 Pateamine A mediates RNA sequence-selective translation repression by anchoring eIF4A and DDX3 to GNG motifs 36 samples
- GSE241674 Immune signature and monocytic leukemia cells are resistant to MDM2 inhibition 14 samples
- GSE236530 Enzyme-mediated methylation and alkynylation enables transcriptome-wide identification of pseudouridine modifications 50 samples
- GSE306225 Distinct 5′ and 3′ Coverage Biases Shape Transcriptome Interpretation in Nanopore Direct RNA versus PCR-cDNA Sequencing 14 samples
Share this dataset
Metadata from NCBI GEO, cached and refreshed periodically — the NCBI page above is authoritative. Downloads link straight to NCBI/ENA; nothing is proxied through BioTransfer.