← BioTransfer GEO Dataset Finder
GEO series

CREsted: modeling genomic and synthetic cell type-specific enhancers across tissues and species

GSE292617 Homo sapiens Genome binding/occupancy profiling by high throughput sequencing 6 samples Submitted 2025/03/28 Platform GPL34281
Summary
Sequence-based deep learning models have become the state of the art for the analysis of the genomic regulatory code. Particularly for transcriptional enhancers, deep learning models excel at deciphering sequence features and grammar that underlie their spatiotemporal activity. To enable end-to-end enhancer modeling and design, we developed a software and modeling package, called CREsted. It combines preprocessing starting from single-cell ATAC-seq data; modeling with a choice of several architectures for training classification and regression models on either topics or pseudobulk peak heights; sequence design using multiple strategies; and downstream analysis through a collection of tools to locate transcription factor (TF) binding sites, infer the effect of a TF (activating or repressing) on enhancer accessibility, decipher enhancer grammar, and score gene loci. We demonstrate CREsted using a mouse cortex model that we validate using the BICCN collection of in vivo validated mouse brain enhancers. Classical enhancers in immune cells, including the IFN-β enhanceosome are revisited using a PBMC model, and we assess the accuracy of TF binding site predictions with ChIP-seq. Additionally, we use CREsted to compare mesenchymal-like cancer cell states between tumor types; and we investigate different fine-tuning strategies of Borzoi within CREsted, comparing their performance and explainability with CREsted models trained from scratch. Finally, we train a CREsted model on a scATAC-seq atlas of zebrafish development, and use this to design and in vivo validate cell type-specific synthetic enhancers in 3 tissues. For varying datasets we demonstrate that CREsted facilitates efficient training and analyses, enabling scrutinization of the enhancer logic and design of synthetic enhancers across tissues and species. CREsted is available at https://crested.readthedocs.io.
This dataset
Download

Direct links to NCBI, no account and no request form: the whole study as GSE292617_RAW.tar, processed values as the series matrix, the supplementary file directory, and per-sample supplementary files for any of the 6 samples. Raw sequencing reads are also available from ENA.

Also filed as BioProject PRJNA1240307 and SRA study SRP574092. Searching any of these in the dataset finder brings you back here.

Samples in this study

The sample list for this study is not cached yet. Press Sort into groups and it will be fetched from NCBI.

+ 6 more — browse all 6 samples with per-sample file links →

Similar datasets

Search all human ChIP / ATAC / CUT&Tag datasets in GEO →

Share this dataset

Metadata from NCBI GEO, cached and refreshed periodically — the NCBI page above is authoritative. Downloads link straight to NCBI/ENA; nothing is proxied through BioTransfer.