Distribution-based Sketching of Single-Cell Samples

06/30/2022
by   Vishal Athreya Baskaran, et al.
0

Modern high-throughput single-cell immune profiling technologies, such as flow and mass cytometry and single-cell RNA sequencing can readily measure the expression of a large number of protein or gene features across the millions of cells in a multi-patient cohort. While bioinformatics approaches can be used to link immune cell heterogeneity to external variables of interest, such as, clinical outcome or experimental label, they often struggle to accommodate such a large number of profiled cells. To ease this computational burden, a limited number of cells are typically sketched or subsampled from each patient. However, existing sketching approaches fail to adequately subsample rare cells from rare cell-populations, or fail to preserve the true frequencies of particular immune cell-types. Here, we propose a novel sketching approach based on Kernel Herding that selects a limited subsample of all cells while preserving the underlying frequencies of immune cell-types. We tested our approach on three flow and mass cytometry datasets and on one single-cell RNA sequencing dataset and demonstrate that the sketched cells (1) more accurately represent the overall cellular landscape and (2) facilitate increased performance in downstream analysis tasks, such as classifying patients according to their clinical outcome. An implementation of sketching with Kernel Herding is publicly available at <https://github.com/vishalathreya/Set-Summarization>.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/18/2022

Interpretable Single-Cell Set Classification with Kernel Mean Embeddings

Modern single-cell flow and mass cytometry technologies measure the expr...
research
09/26/2016

Connecting the dots across time: Reconstruction of single cell signaling trajectories using time-stamped data

Single cell responses are shaped by the geometry of signaling kinetic tr...
research
07/17/2023

Kernel-Based Testing for Single-Cell Differential Analysis

Single-cell technologies have provided valuable insights into the distri...
research
08/23/2021

Automated Identification of Cell Populations in Flow Cytometry Data with Transformers

Acute Lymphoblastic Leukemia (ALL) is the most frequent hematologic mali...
research
10/13/2021

CloudPred: Predicting Patient Phenotypes From Single-cell RNA-seq

Single-cell RNA sequencing (scRNA-seq) has the potential to provide powe...
research
10/11/2012

Inferring clonal evolution of tumors from single nucleotide somatic mutations

High-throughput sequencing allows the detection and quantification of fr...
research
08/11/2022

Interpretable cytometry cell-type annotation with flow-based deep generative models

Cytometry enables precise single-cell phenotyping within heterogeneous p...

Please sign up or login with your details

Forgot password? Click here to reset