Facilitating Federated Genomic Data Analysis by Identifying Record Correlations while Ensuring Privacy

03/10/2022
by   Leonard Dervishi, et al.
0

With the reduction of sequencing costs and the pervasiveness of computing devices, genomic data collection is continually growing. However, data collection is highly fragmented and the data is still siloed across different repositories. Analyzing all of this data would be transformative for genomics research. However, the data is sensitive, and therefore cannot be easily centralized. Furthermore, there may be correlations in the data, which if not detected, can impact the analysis. In this paper, we take the first step towards identifying correlated records across multiple data repositories in a privacy-preserving manner. The proposed framework, based on random shuffling, synthetic record generation, and local differential privacy, allows a trade-off of accuracy and computational efficiency. An extensive evaluation on real genomic data from the OpenSNP dataset shows that the proposed solution is efficient and effective.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
03/23/2016

On the Theory and Practice of Privacy-Preserving Bayesian Data Analysis

Bayesian inference has great promise for the privacy-preserving analysis...
research
04/01/2022

LDP-IDS: Local Differential Privacy for Infinite Data Streams

Streaming data collection is essential to real-time data analytics in va...
research
06/17/2021

Interval Privacy: A Framework for Data Collection

The emerging public awareness and government regulations of data privacy...
research
04/25/2023

(Local) Differential Privacy has NO Disparate Impact on Fairness

In recent years, Local Differential Privacy (LDP), a robust privacy-pres...
research
02/02/2023

Fed-GLOSS-DP: Federated, Global Learning using Synthetic Sets with Record Level Differential Privacy

This work proposes Fed-GLOSS-DP, a novel approach to privacy-preserving ...
research
04/04/2023

Privacy-Preserving Federated Discovery of DNA Motifs with Differential Privacy

DNA motif discovery is an important issue in gene research, which aims t...
research
12/19/2018

Preventing Attacks on Anonymous Data Collection

Anonymous data collection systems allow users to contribute the data nec...

Please sign up or login with your details

Forgot password? Click here to reset