What are the relationships between the different omics datasets?
How well do the omics datasets, individually or combined, predict outcomes?
Supervised or unsupervised?
Supervised we have some kind of outcome variable (example: disease status) and we are trying to figure out which combinations of features in the different omics datasets predict it
Unsupervised we don’t have an outcome, we are just exploring relationships between the datasets
Descriptive or predictive? (within supervised methods)
Descriptive we want to find weights for the variables to optimally separate the classes
many different criteria can be used to decide what constitutes “optimal”
Predictive we want to predict the class of new samples if we know its variables