A Framework for Mediation Analysis with Massive Data

02/14/2023
by   Haixiang Zhang, et al.
0

During the past few years, mediation analysis has gained increasing popularity across various research fields. The primary objective of mediation analysis is to examine the direct impact of exposure on outcome, as well as the indirect effects that occur along the pathways from exposure to outcome. There has been a great number of articles that applied mediation analysis to data from hundreds or thousands of individuals. With the rapid development of technology, the volume of avaliable data increases exponentially, which brings new challenges to researchers. Directly conducting statistical analysis for large datasets is often computationally infeasible. Nonetheless, there is a paucity of findings regarding mediation analysis in the context of big data. In this paper, we propose utilizing subsampled double bootstrap and divide-and-conquer algorithms to conduct statistical mediation analysis on large-scale datasets. The proposed algorithms offer a significant enhancement in computational efficiency over traditional bootstrap confidence interval and Sobel test, while simultaneously ensuring desirable confidence interval coverage and power. We conducted extensive numerical simulations to evaluate the performance of our method. The practical applicability of our approach is demonstrated through two real-world data examples.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/27/2012

The Big Data Bootstrap

The bootstrap provides a simple and powerful means of assessing the qual...
research
02/06/2023

A Fast Bootstrap Algorithm for Causal Inference with Large Data

Estimating causal effects from large experimental and observational data...
research
08/19/2021

Estimating the natural indirect effect and the mediation proportion via the product method

The natural indirect effect (NIE) and mediation proportion (MP) are two ...
research
04/09/2015

Robust, scalable and fast bootstrap method for analyzing large scale data

In this paper we address the problem of performing statistical inference...
research
12/21/2011

A Scalable Bootstrap for Massive Data

The bootstrap provides a simple and powerful means of assessing the qual...
research
04/09/2020

Confidence interval for the AUC of SROC curve and some related methods using bootstrap for meta-analysis of diagnostic accuracy studies

Background: The area under the curve (AUC) of summary receiver operating...
research
09/25/2019

Exact confidence interval for generalized Flajolet-Martin algorithms

This paper develop a deep mathematical-statistical approach to analyze a...

Please sign up or login with your details

Forgot password? Click here to reset