On admissibility in post-hoc hypothesis testing

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Chugg, Ben, Lardy, Tyron, Ramdas, Aaditya, Grünwald, Peter
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911385660686336
author Chugg, Ben
Lardy, Tyron
Ramdas, Aaditya
Grünwald, Peter
author_facet Chugg, Ben
Lardy, Tyron
Ramdas, Aaditya
Grünwald, Peter
contents The validity of classical hypothesis testing requires the significance level $α$ be fixed before any statistical analysis takes place. This is a stringent requirement. For instance, it prohibits updating $α$ during (or after) an experiment due to changing concern about the cost of false positives, or to reflect unexpectedly strong evidence against the null. Perhaps most disturbingly, witnessing a p-value $p\llα$ vs $p= α- ε$ for tiny $ε> 0$ has no (statistical) relevance for any downstream decision-making. Following recent work of Grünwald (2024), we develop a theory of post-hoc hypothesis testing, enabling $α$ to be chosen after seeing and analyzing the data. To study "good" post-hoc tests we introduce $Γ$-admissibility, where $Γ$ is a set of adversaries which map the data to a significance level. We classify the set of $Γ$-admissible rules for various sets $Γ$, showing they must be based on e-values, and recover the Neyman-Pearson lemma when $Γ$ is the constant map.
format Preprint
id arxiv_https___arxiv_org_abs_2508_00770
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle On admissibility in post-hoc hypothesis testing
Chugg, Ben
Lardy, Tyron
Ramdas, Aaditya
Grünwald, Peter
Statistics Theory
Methodology
The validity of classical hypothesis testing requires the significance level $α$ be fixed before any statistical analysis takes place. This is a stringent requirement. For instance, it prohibits updating $α$ during (or after) an experiment due to changing concern about the cost of false positives, or to reflect unexpectedly strong evidence against the null. Perhaps most disturbingly, witnessing a p-value $p\llα$ vs $p= α- ε$ for tiny $ε> 0$ has no (statistical) relevance for any downstream decision-making. Following recent work of Grünwald (2024), we develop a theory of post-hoc hypothesis testing, enabling $α$ to be chosen after seeing and analyzing the data. To study "good" post-hoc tests we introduce $Γ$-admissibility, where $Γ$ is a set of adversaries which map the data to a significance level. We classify the set of $Γ$-admissible rules for various sets $Γ$, showing they must be based on e-values, and recover the Neyman-Pearson lemma when $Γ$ is the constant map.
title On admissibility in post-hoc hypothesis testing
topic Statistics Theory
Methodology
url https://arxiv.org/abs/2508.00770