The Prediction-Explanation Fallacy: A Pervasive Problem in Scientific Applications of Machine Learning

Marco Del Giudice

doi:10.5964/meth.11235

The Prediction-Explanation Fallacy: A Pervasive Problem in Scientific Applications of Machine Learning

Marco Del Giudice
Department of Life Sciences, University of Trieste, Trieste, Italy

Abstract

I highlight a problem that has become ubiquitous in scientific applications of machine learning and can lead to seriously distorted inferences. I call it the Prediction-Explanation Fallacy. The fallacy occurs when researchers use prediction-optimized models for explanatory purposes, without considering the relevant tradeoffs. This is a problem for at least two reasons. First, prediction-optimized models are often deliberately biased and unrealistic in order to prevent overfitting. In other cases, they have an exceedingly complex structure that is hard or impossible to interpret. Second, different predictive models trained on the same or similar data can be biased in different ways, so that they may predict equally well but suggest conflicting explanations. Here I introduce the tradeoffs between prediction and explanation in a non-technical fashion, present illustrative examples from neuroscience, and end by discussing some mitigating factors and methods that can be used to limit the problem.

PDF HTML XML

Published at

22. March 2024
https://doi.org/10.5964/meth.11235
Issue:

Vol. 20 No. 1 (2024)
Section:

Original Article
Keywords:

bias-variance tradeoff machine learning prediction Rashomon effect
Share:

Del Giudice, M. (2024). The Prediction-Explanation Fallacy: A Pervasive Problem in Scientific Applications of Machine Learning. Methodology, 20(1), 22-46. https://doi.org/10.5964/meth.11235

Download Citation

This work is licensed under a Creative Commons Attribution (CC BY) 4.0 International License.

PlumX

Dimensions

Views:

Total	Abstract	PDF	HTML	XML
974	513	381	50	30

Authors

Abstract