DeepAI AI Chat
Log In Sign Up

Scalable Topological Data Analysis and Visualization for Evaluating Data-Driven Models in Scientific Applications

by   Shusen Liu, et al.

With the rapid adoption of machine learning techniques for large-scale applications in science and engineering comes the convergence of two grand challenges in visualization. First, the utilization of black box models (e.g., deep neural networks) calls for advanced techniques in exploring and interpreting model behaviors. Second, the rapid growth in computing has produced enormous datasets that require techniques that can handle millions or more samples. Although some solutions to these interpretability challenges have been proposed, they typically do not scale beyond thousands of samples, nor do they provide the high-level intuition scientists are looking for. Here, we present the first scalable solution to explore and analyze high-dimensional functions often encountered in the scientific data analysis pipeline. By combining a new streaming neighborhood graph construction, the corresponding topology computation, and a novel data aggregation scheme, namely topology aware datacubes, we enable interactive exploration of both the topological and the geometric aspect of high-dimensional data. Following two use cases from high-energy-density (HED) physics and computational biology, we demonstrate how these capabilities have led to crucial new insights in both applications.


page 4

page 7

page 9


Opening the black-box of Neighbor Embedding with Hotelling's T2 statistic and Q-residuals

In contrast to classical techniques for exploratory analysis of high-dim...

Forming IDEAS Interactive Data Exploration & Analysis System

Modern cyber security operations collect an enormous amount of logging a...

ShapeVis: High-dimensional Data Visualization at Scale

We present ShapeVis, a scalable visualization technique for point cloud ...

Pheno-Mapper: An Interactive Toolbox for the Visual Exploration of Phenomics Data

High-throughput technologies to collect field data have made observation...

Geometric and Topological Inference for Deep Representations of Complex Networks

Understanding the deep representations of complex networks is an importa...

High dimensionality: The latest challenge to data analysis

The advent of modern technology, permitting the measurement of thousands...

Code Repositories


Robust neural network surrogate for inertial confinement fusion

view repo