Saved in:
Bibliographic Details
Main Authors: Wan, Chuang, Sun, Jiajun, Xu, Xingbai
Format: Preprint
Published: 2025
Subjects:
Online Access:https://arxiv.org/abs/2511.04466
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915602646433792
author Wan, Chuang
Sun, Jiajun
Xu, Xingbai
author_facet Wan, Chuang
Sun, Jiajun
Xu, Xingbai
contents We consider the problem of testing for differences in group-specific slopes between the selected groups in panel data identified via k-means clustering. In this setting, the classical Wald-type test statistic is problematic because it produces an extremely inflated type I error probability. The underlying reason is that the same dataset is used to identify the group structure and construct the test statistic, simultaneously. This creates dependence between the selection and inference stages. To address this issue, we propose a valid selective inference approach conditional on the selection event to account for the selection effect. We formally define the selective type I error and describe how to efficiently compute the correct p-values for clusters obtained using k-means clustering. Furthermore, the same idea can be extended to test for differences in coefficients due to a single covariate and can be incorporated into the GMM estimation framework. Simulation studies show that our method has satisfactory finite sample performance. We apply this method to explore the heterogeneous relationships between economic growth and the $CO_2$ emission across countries for which some new findings are discovered. An R package TestHomoPanel is provided to implement the proposed selective inference framework for panel data.
format Preprint
id arxiv_https___arxiv_org_abs_2511_04466
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Conditional Selective Inference for the Selected Groups in Panel Data
Wan, Chuang
Sun, Jiajun
Xu, Xingbai
Methodology
We consider the problem of testing for differences in group-specific slopes between the selected groups in panel data identified via k-means clustering. In this setting, the classical Wald-type test statistic is problematic because it produces an extremely inflated type I error probability. The underlying reason is that the same dataset is used to identify the group structure and construct the test statistic, simultaneously. This creates dependence between the selection and inference stages. To address this issue, we propose a valid selective inference approach conditional on the selection event to account for the selection effect. We formally define the selective type I error and describe how to efficiently compute the correct p-values for clusters obtained using k-means clustering. Furthermore, the same idea can be extended to test for differences in coefficients due to a single covariate and can be incorporated into the GMM estimation framework. Simulation studies show that our method has satisfactory finite sample performance. We apply this method to explore the heterogeneous relationships between economic growth and the $CO_2$ emission across countries for which some new findings are discovered. An R package TestHomoPanel is provided to implement the proposed selective inference framework for panel data.
title Conditional Selective Inference for the Selected Groups in Panel Data
topic Methodology
url https://arxiv.org/abs/2511.04466