Bidimensional linked matrix factorization for pan-omics pan-cancer analysis

02/07/2020
by   Eric F. Lock, et al.
0

Several modern applications require the integration of multiple large data matrices that have shared rows and/or columns. For example, cancer studies that integrate multiple omics platforms across multiple types of cancer, pan-omics pan-cancer analysis, have extended our knowledge of molecular heterogenity beyond what was observed in single tumor and single platform studies. However, these studies have been limited by available statistical methodology. We propose a flexible approach to the simultaneous factorization and decomposition of variation across such bidimensionally linked matrices, BIDIFAC+. This decomposes variation into a series of low-rank components that may be shared across any number of row sets (e.g., omics platforms) or column sets (e.g., cancer types). This builds on a growing literature for the factorization and decomposition of linked matrices, which has primarily focused on multiple matrices that are linked in one dimension (rows or columns) only. Our objective function extends nuclear norm penalization, is motivated by random matrix theory, gives an identifiable decomposition under relatively mild conditions, and can be shown to give the mode of a Bayesian posterior distribution. We apply BIDIFAC+ to pan-omics pan-cancer data from TCGA, identifying shared and specific modes of variability across 4 different omics platforms and 29 different cancer types.

READ FULL TEXT
research
06/09/2019

Integrative Factorization of Bidimensionally Linked Matrices

Advances in molecular "omics'" technologies have motivated new methodolo...
research
08/30/2023

Multiple Augmented Reduced Rank Regression for Pan-Cancer Analysis

Statistical approaches that successfully combine multiple datasets are m...
research
02/28/2021

A Hierarchical Spike-and-Slab Model for Pan-Cancer Survival Using Pan-Omic Data

Pan-omics, pan-cancer analysis has advanced our understanding of the mol...
research
02/19/2021

A Higher-Order Generalized Singular Value Decomposition for Rank Deficient Matrices

The higher-order generalized singular value decomposition (HO-GSVD) is a...
research
09/13/2018

MSc Dissertation: Exclusive Row Biclustering for Gene Expression Using a Combinatorial Auction Approach

The availability of large microarray data has led to a growing interest ...
research
01/09/2018

GIFT: Guided and Interpretable Factorization for Tensors - An Application to Large-Scale Multi-platform Cancer Analysis

Given multi-platform genome data with prior knowledge of functional gene...
research
06/29/2021

Meta-learning for Matrix Factorization without Shared Rows or Columns

We propose a method that meta-learns a knowledge on matrix factorization...

Please sign up or login with your details

Forgot password? Click here to reset