← BioTransfer GEO Dataset Finder
GEO series

Cell-free DNA comprises an in vivo, genome-wide nucleosome footprint that informs its tissue(s)-of-origin

GSE71378 Homo sapiens Genome binding/occupancy profiling by high throughput sequencing 60 samples Submitted 2015/11/04 Platform GPL11154
Summary
Nucleosomes are the basic unit of packaging of eukaryotic chromatin, and nucleosome positioning can differ substantially between cell types. Here, we sequence 14.5 billion plasma-borne cell-free DNA (cfDNA) fragments (700-fold coverage) to generate genome-wide maps of in vivo nucleosome occupancy. We identify 13 million local maxima of nucleosome protection, spanning 2.53 gigabases (Gb) of the human genome, whose positions and spacings correlate with nuclear architecture, gene structure and gene expression. We further show that short cfDNA fragments - poorly recovered by standard protocols - directly footprint the in vivo occupancy of DNA-bound transcription factors such as CTCF. The sequence composition of cfDNA has previously been used to noninvasively monitor cancer, pregnancy and organ transplantation, but a key limitation of this paradigm is its dependence on genotypic differences to distinguish between contributing tissues. We show that nucleosome spacing in gene bodies and cis-regulatory elements, inferred from cfDNA in healthy individuals, correlates most strongly with transcriptional and epigenetic features of lymphoid and myeloid cells, consistent with hematopoietic cell death as the normal source of cfDNA. We build on this observation to show how in vivo nucleosome footprints can be used to infer the cell types that contribute to circulating cfDNA in pathological states such as cancer. Because it does not rely on genotypic differences, this strategy may enable the noninvasive cfDNA-based monitoring of a much broader set of clinical conditions than is currently possible.
This dataset
Download

Direct links to NCBI, no account and no request form: the whole study as GSE71378_RAW.tar, processed values as the series matrix, the supplementary file directory, and per-sample supplementary files for any of the 60 samples. Raw sequencing reads are also available from ENA.

Also filed as BioProject PRJNA291063 and SRA study SRP061633. Searching any of these in the dataset finder brings you back here.

Samples in this study

The sample list for this study is not cached yet. Press Sort into groups and it will be fetched from NCBI.

+ 60 more — browse all 60 samples with per-sample file links →

Similar datasets

Search all human ChIP / ATAC / CUT&Tag datasets in GEO →

Share this dataset

Metadata from NCBI GEO, cached and refreshed periodically — the NCBI page above is authoritative. Downloads link straight to NCBI/ENA; nothing is proxied through BioTransfer.