Published January 3, 2019 | Version v1

Processed TCGA pan-cancer data set used in the I-Boost paper (Wong et al. 2019)

  • 1. The Hong Kong Polytechnic University
  • 2. University of North Carolina at Chapel Hill

Description

This data set contains the clinical and genomics data for 1,420 subjects analyzed in the paper: Wong KY, Fan C, Tanioka M, Parker JS, Nobel AB, Zeng D, Lin DY, Perou CM. I-Boost: an integrative boosting approach for predicting survival time with multiple genomics platforms. Genome Biology. 2019. It contains data on time to death, cancer type, 4 clinical variables, expression of 12,434 genes, somatic mutation of 130 genes, expression of 305 miRNA, expression of 136 proteins or phospho-proteins, copy number of 216 DNA segments, and 497 gene expression modules. Data on time to death, clinical variables, somatic mutation, copy number variation, mRNA expression, and miRNA expression were derived from the pan-cancer data set at Synapse (syn2468297 at https://www.synapse.org/#!Synapse:syn2468297). The protein expression data were obtained from Broad GDAC Firehose (https://gdac.broadinstitute.org/runs/stddata__2016_01_28/).

Files

TCGA_8cancer_rmmis.csv

Files (123.3 MB)

Name Size Download all
md5:b820e29a176cd8ffcffe751fadfdb72f
123.3 MB Preview Download