Deep Learning of Representations for Transcriptomics-based Phenotype Prediction

The ability to predict health outcomes from gene expression would catalyze a revolution in molecular diagnostics. This task is complicated because expression data are high dimensional whereas each experiment is usually small (e.g., ~20,000 genes may be measured for ~100 subjects). However, thousands of transcriptomics experiments with hundreds of thousands of samples are available in public repositories. Can representation learning techniques leverage these public data to improve predictive performance on other tasks? Here, we report a comprehensive analysis using different gene sets, normalization schemes, and machine learning methods on a set of 24 binary and multiclass prediction problems and 26 survival analysis tasks. Methods that combine large numbers of genes outperformed single gene methods, but neither unsupervised nor semi-supervised representation learning techniques yielded consistent improvements in out-of-sample performance across datasets. Our findings suggest that using l2-regularized regression methods applied to centered log-ratio transformed transcript abundances provide the best predictive analyses.

Enter your email address to download paper.

Click the link to begin download.
Oops! Something went wrong while submitting the form.

Enter your email address to watch the webinar.

Click the link to watch webinar.
Oops! Something went wrong while submitting the form.
Webinars

How will AI transform the future of medicine?

White Papers

Evaluating Digital Twins for Alzheimer’s Disease using Data from a Completed Phase 2 Clinical Trial

White Papers

Prognostic digital twins overcome the limitations of external control arms in RCTs

Both methods reduce control arm sizes, but only digital twins control for bias.
A Phase 2 study on crenezumab in mild-to-moderate AD was used to retrospectively assess the validity of Unlearn's approach for AD clinical trials.
Hear from Charles Fisher, founder and CEO of Unlearn, in this on-demand webinar about how AI will transform the medical landscape of tomorrow.