A geometric view of Biodiversity: scaling to metagenomics

03/06/2018
by   Pierre Blanchard, et al.
0

We have designed a new efficient dimensionality reduction algorithm in order to investigate new ways of accurately characterizing the biodiversity, namely from a geometric point of view, scaling with large environmental sets produced by NGS (∼ 10^5 sequences). The approach is based on Multidimensional Scaling (MDS) that allows for mapping items on a set of n points into a low dimensional euclidean space given the set of pairwise distances. We compute all pairwise distances between reads in a given sample, run MDS on the distance matrix, and analyze the projection on first axis, by visualization tools. We have circumvented the quadratic complexity of computing pairwise distances by implementing it on a hyperparallel computer (Turing, a Blue Gene Q), and the cubic complexity of the spectral decomposition by implementing a dense random projection based algorithm. We have applied this data analysis scheme on a set of 10^5 reads, which are amplicons of a diatom environmental sample from Lake Geneva. Analyzing the shape of the point cloud paves the way for a geometric analysis of biodiversity, and for accurately building OTUs (Operational Taxonomic Units), when the data set is too large for implementing unsupervised, hierarchical, high-dimensional clustering.

READ FULL TEXT
research
02/22/2023

nSimplex Zen: A Novel Dimensionality Reduction for Euclidean and Hilbert Spaces

Dimensionality reduction techniques map values from a high dimensional s...
research
12/08/2018

Letting symmetry guide visualization: multidimensional scaling on groups

Multidimensional scaling (MDS) is a fundamental tool for both data visua...
research
09/11/2017

Subspace Least Squares Multidimensional Scaling

Multidimensional Scaling (MDS) is one of the most popular methods for di...
research
08/01/2023

Self-supervised Multidimensional Scaling with F-ratio: Improving Microbiome Visualization

Multidimensional scaling (MDS) is an unsupervised learning technique tha...
research
10/26/2022

Bayesian Hyperbolic Multidimensional Scaling

Multidimensional scaling (MDS) is a widely used approach to representing...
research
06/21/2019

Visualizing Representational Dynamics with Multidimensional Scaling Alignment

Representational similarity analysis (RSA) has been shown to be an effec...
research
03/10/2023

A dual basis approach to multidimensional scaling: spectral analysis and graph regularity

Classical multidimensional scaling (CMDS) is a technique that aims to em...

Please sign up or login with your details

Forgot password? Click here to reset