← BioTransfer GEO Dataset Finder
GEO series

WBT-DC Pipeline: Whole Blood Transcriptomics data-based Disease Classification

GSE282218 Homo sapiens Expression profiling by high throughput sequencing 165 samples 2026/05/20 GPL24676
Summary
Machine learning together with cell/tissue transcriptomics data has been widely used for disease classification. However, obtaining transcriptomics data for human tissues require invasive procedures, making it challenging for widespread application in the clinic. In this study, we developed the WBT-DC (Whole Blood Transcriptomics (WBT) data based Disease Classification). We utilized gene rank-based methods for feature extraction to mitigate issues associated with batch effects and gene noise. We applied the ensemble machine learning model, Random Forest, and performed cross-validation and model tuning. We evaluated our methods on four different diseases including crohn's disease (CD), ulcerative colitis (UC) and amyotrophic lateral sclerosis (ALS) and rheumatoid arthritis (RA) datasets, using data from seven independent cohorts and 2,452 participants, across RNA-Sequencing and microarrays. Our machine learning based WBT-DC pipeline demonstrated a robust performance across various disease datasets and different transcriptomics platforms, establishing itself as a valuable non-invasive tool for future disease classification and prediction.
Download
NCBI GEO page ↗ Paper (PMID 42116144) ↗ {# Names what the click gives you. "Open in finder" meant nothing to a visitor who arrived from a search engine and has never seen the tool. #} Find more human RNA-seq datasets →
Similar datasets

Search all human RNA-seq datasets in GEO →

Share this dataset

Metadata from NCBI GEO, cached and refreshed periodically — the NCBI page above is authoritative. Downloads link straight to NCBI/ENA; nothing is proxied through BioTransfer.