A Binary Regression Adaptive Goodness-of-fit Test (BAGofT)

11/08/2019
by   Jiawei Zhang, et al.
24

The Pearson's χ^2 test and residual deviance test are two classical goodness-of-fit tests for binary regression models such as logistic regression. These two tests cannot be applied when we have one or more continuous covariates in the data, a quite common situation in practice. In that case, the most widely used approach is the Hosmer-Lemeshow test, which partitions the covariate space into groups according to quantiles of the fitted probabilities from all the observations. However, its grouping scheme is not flexible enough to explore how to adversarially partition the data space in order to enhance the power. In this work, we propose a new methodology, named binary regression adaptive grouping goodness-of-fit test (BAGofT), to address the above concern. It is a two-stage solution where the first stage adaptively selects candidate partitions using "training" data, and the second stage performs χ^2 tests with necessary corrections based on "test" data. A proper data splitting ensures that the test has desirable size and power properties. From our experimental results, BAGofT performs much better than Hosmer-Lemeshow test in many situations.

READ FULL TEXT

page 24

page 25

research
02/25/2021

Improving the Hosmer-Lemeshow Goodness-of-Fit Test in Large Models with Replicated Trials

The Hosmer-Lemeshow (HL) test is a commonly used global goodness-of-fit ...
research
08/09/2019

Goodness-of-fit testing in high-dimensional generalized linear models

We propose a family of tests to assess the goodness-of-fit of a high-dim...
research
10/27/2018

Informative Features for Model Comparison

Given two candidate models, and a set of target observations, we address...
research
09/30/2020

Testing for linearity in boundary regression models with application to maximal life expectancies

We consider a regression model with errors that are a.s. negative. Thus ...
research
11/03/2019

Variable Grouping Based Bayesian Additive Regression Tree

Using ensemble methods for regression has been a large success in obtain...
research
08/11/2023

A Plot is Worth a Thousand Tests: Assessing Residual Diagnostics with the Lineup Protocol

Regression experts consistently recommend plotting residuals for model d...

Please sign up or login with your details

Forgot password? Click here to reset