conference-paper Open access

Questionable Practices in Methodological Deep Learning Research

  • Proceedings of the Northern Lights Deep Learning Workshop
Research footprint

At a glance

Citations
3
References
17
Comments
0
Paper overview

Abstract

Evaluation of new methodology in deep learning (DL) research is typically done by reporting point estimates of a few performance metrics, calculated from a single training run. This paper argues that this frequently used evaluation protocol in DL is fundamentally flawed -- presenting 8 questionable practices that are widely adopted in the evaluation of new DL methods. The questionable practices are derived from violations of statistical principles of the scientific method, and from Hansson's definition of pseudoscience. A survey of recent publications from a top-tier DL conference indicates the widespread adoption of these practices in state-of-the-art DL research. Lastly, arguments in favor of the questionable practices, possible reasons for their adoption, and measures that have been taken to remove them, are discussed.

Record transparency

Publication details

DOI
10.7557/18.6804
OpenAlex
W4317905935
Document type
conference-paper
Language
EN
Source
Proceedings of the Northern Lights Deep Learning Workshop
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.