Privacy-Preserving Tensor Factorization for Collaborative Health Data Analysis

08/26/2019
by   Jing Ma, et al.
0

Tensor factorization has been demonstrated as an efficient approach for computational phenotyping, where massive electronic health records (EHRs) are converted to concise and meaningful clinical concepts. While distributing the tensor factorization tasks to local sites can avoid direct data sharing, it still requires the exchange of intermediary results which could reveal sensitive patient information. Therefore, the challenge is how to jointly decompose the tensor under rigorous and principled privacy constraints, while still support the model's interpretability. We propose DPFact, a privacy-preserving collaborative tensor factorization method for computational phenotyping using EHR. It embeds advanced privacy-preserving mechanisms with collaborative learning. Hospitals can keep their EHR database private but also collaboratively learn meaningful clinical concepts by sharing differentially private intermediary results. Moreover, DPFact solves the heterogeneous patient population using a structured sparsity term. In our framework, each hospital decomposes its local tensors, and sends the updated intermediary results with output perturbation every several iterations to a semi-trusted server which generates the phenotypes. The evaluation on both real-world and synthetic datasets demonstrated that under strict privacy constraints, our method is more accurate and communication-efficient than state-of-the-art baseline methods.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
04/11/2017

Federated Tensor Factorization for Computational Phenotyping

Tensor factorization models offer an effective approach to convert massi...
research
09/03/2021

Communication Efficient Tensor Factorization for Decentralized Healthcare Networks

Tensor factorization has been proved as an efficient unsupervised learni...
research
08/09/2023

Collaborative Learning From Distributed Data With Differentially Private Synthetic Twin Data

Consider a setting where multiple parties holding sensitive data aim to ...
research
10/02/2018

A New Approach to Privacy-Preserving Clinical Decision Support Systems for HIV Treatment

Background: HIV treatment prescription is a complex process; clinical de...
research
11/11/2019

Privacy-Preserving Multiple Tensor Factorization for Synthesizing Large-Scale Location Traces

With the widespread use of LBSs (Location-based Services), synthesizing ...
research
08/08/2018

PIVETed-Granite: Computational Phenotypes through Constrained Tensor Factorization

It has been recently shown that sparse, nonnegative tensor factorization...
research
11/21/2018

Privacy-Preserving Collaborative Prediction using Random Forests

We study the problem of privacy-preserving machine learning (PPML) for e...

Please sign up or login with your details

Forgot password? Click here to reset