A causal inference framework for cancer cluster investigations using publicly available data

11/14/2018
by   Rachel C. Nethery, et al.
0

Often, a community becomes alarmed when high rates of cancer are noticed, and residents suspect that the cancer cases could be caused by a known source of hazard. In response, the CDC recommends that departments of health perform a standardized incidence ratio (SIR) analysis to determine whether the observed cancer incidence is higher than expected. This approach has several limitations that are well documented in the literature. In this paper we propose a novel causal inference approach to cancer cluster investigations, rooted in the potential outcomes framework. Assuming that a source of hazard representing a potential cause of increased cancer rates in the community is identified a priori, we introduce a new estimand called the causal SIR (cSIR). The cSIR is a ratio defined as the expected cancer incidence in the exposed population divided by the expected cancer incidence under the (counterfactual) scenario of no exposure. To estimate the cSIR we need to overcome two main challenges: 1) identify unexposed populations that are as similar as possible to the exposed one to inform estimation under the counterfactual scenario of no exposure, and 2) make inference on cancer incidence in these unexposed populations using publicly available data that are often available at a much higher level of spatial aggregation than what is desired. We overcome the first challenge by relying on matching. We overcome the second challenge by developing a Bayesian hierarchical model that borrows information from other sources to impute cancer incidence at the desired finer level of spatial aggregation. We apply our proposed approach to determine whether trichloroethylene vapor exposure has caused increased cancer incidence in Endicott, NY.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
03/17/2020

ParKCa: Causal Inference with Partially Known Causes

Causal Inference methods based on observational data are an alternative ...
research
02/26/2019

A Source-Oriented Approach to Coal Power Plant Emissions Health Effects

There is increasing focus on whether air pollution originating from diff...
research
07/18/2023

Spatio-temporal quasi-experimental methods for rare disease outcomes: The impact of reformulated gasoline on childhood hematologic cancer

Although some pollutants emitted in vehicle exhaust, such as benzene, ar...
research
06/05/2023

Methods for Estimating the Exposure-Response Curve to Inform the New Safety Standards for Fine Particulate Matter

Exposure to fine particulate matter (PM_2.5) poses significant health ri...
research
06/09/2023

Longitudinal Position and Cancer Risk in the United States Revisited

The debate over whether to keep daylight savings time has gained attenti...
research
11/08/2020

Adversarial Counterfactual Learning and Evaluation for Recommender System

The feedback data of recommender systems are often subject to what was e...
research
09/12/2018

Access to Population-Level Signaling as a Source of Inequality

We identify and explore differential access to population-level signalin...

Please sign up or login with your details

Forgot password? Click here to reset