The good, the bad, and the ugly: Bayesian model selection produces spurious posterior probabilities for phylogenetic trees

10/12/2018
by   Ziheng Yang, et al.
0

The Bayesian method is noted to produce spuriously high posterior probabilities for phylogenetic trees in analysis of large datasets, but the precise reasons for this over-confidence are unknown. In general, the performance of Bayesian selection of misspecified models is poorly understood, even though this is of great scientific interest since models are never true in real data analysis. Here we characterize the asymptotic behavior of Bayesian model selection and show that when the competing models are equally wrong, Bayesian model selection exhibits surprising and polarized behaviors in large datasets, supporting one model with full force while rejecting the others. If one model is slightly less wrong than the other, the less wrong model will eventually win when the amount of data increases, but the method may become overconfident before it becomes reliable. We suggest that this extreme behavior may be a major factor for the spuriously high posterior probabilities for evolutionary trees. The philosophical implications of our results to the application of Bayesian model selection to evaluate opposing scientific hypotheses are yet to be explored, as are the behaviors of non-Bayesian methods in similar situations.

READ FULL TEXT
research
07/24/2020

Robust and Reproducible Model Selection Using Bagged Posteriors

Bayesian model selection is premised on the assumption that the data are...
research
10/04/2018

Bayesian Model Selection for a Class of Spatially-Explicit Capture Recapture Models

A vast amount of ecological knowledge generated recently has hinged upon...
research
11/14/2022

Scalable Model Selection for Staged Trees: Mean-posterior Clustering and Binary Trees

Several structure-learning algorithms for staged trees, asymmetric exten...
research
03/17/2018

Large-Scale Model Selection with Misspecification

Model selection is crucial to high-dimensional learning and inference fo...
research
10/19/2016

Robust and Parallel Bayesian Model Selection

Effective and accurate model selection is an important problem in modern...
research
06/24/2014

Reliable ABC model choice via random forests

Approximate Bayesian computation (ABC) methods provide an elaborate appr...
research
05/08/2018

Bayesian models in geographic profiling

We consider the problem of geographic profiling and offer an approach to...

Please sign up or login with your details

Forgot password? Click here to reset