An Empirical Study of Pre-trained Model Selection for Out-of-Distribution Generalization and Calibration
Fuente:
arXiv
Saved in:
| Main Authors: | Naganuma, Hiroki, Hataya, Ryuichiro, Yoshida, Kotaro, Mitliagkas, Ioannis |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Revisiting Generalization Measures Beyond IID: An Empirical Study under Distributional Shift
by: Nakai, Sora, et al.
Published: (2026)
by: Nakai, Sora, et al.
Published: (2026)
Empirical Analysis of Model Selection for Heterogeneous Causal Effect Estimation
by: Mahajan, Divyat, et al.
Published: (2022)
by: Mahajan, Divyat, et al.
Published: (2022)
Towards Understanding Variants of Invariant Risk Minimization through the Lens of Calibration
by: Yoshida, Kotaro, et al.
Published: (2024)
by: Yoshida, Kotaro, et al.
Published: (2024)
Geometric Insights into Focal Loss: Reducing Curvature for Enhanced Model Calibration
by: Kimura, Masanari, et al.
Published: (2024)
by: Kimura, Masanari, et al.
Published: (2024)
Adaptive Batch Sizes Using Non-Euclidean Gradient Noise Scales for Stochastic Sign and Spectral Descent
by: Naganuma, Hiroki, et al.
Published: (2026)
by: Naganuma, Hiroki, et al.
Published: (2026)
Expecting The Unexpected: Towards Broad Out-Of-Distribution Detection
by: Guille-Escuret, Charles, et al.
Published: (2023)
by: Guille-Escuret, Charles, et al.
Published: (2023)
Feature learning as alignment: a structural property of gradient descent in non-linear neural networks
by: Beaglehole, Daniel, et al.
Published: (2024)
by: Beaglehole, Daniel, et al.
Published: (2024)
Provable Target Sample Complexity Improvements as Pre-Trained Models Scale
by: Fukuchi, Kazuto, et al.
Published: (2026)
by: Fukuchi, Kazuto, et al.
Published: (2026)
Towards efficient representation identification in supervised learning
by: Ahuja, Kartik, et al.
Published: (2022)
by: Ahuja, Kartik, et al.
Published: (2022)
Performance Control in Early Exiting to Deploy Large Models at the Same Cost of Smaller Ones
by: Mofakhami, Mehrnaz, et al.
Published: (2024)
by: Mofakhami, Mehrnaz, et al.
Published: (2024)
Causal Pre-training Under the Fairness Lens: An Empirical Study of TabPFN
by: Liu, Qinyi, et al.
Published: (2026)
by: Liu, Qinyi, et al.
Published: (2026)
Glocal Hypergradient Estimation with Koopman Operator
by: Hataya, Ryuichiro, et al.
Published: (2024)
by: Hataya, Ryuichiro, et al.
Published: (2024)
Your Pre-trained LLM is Secretly an Unsupervised Confidence Calibrator
by: Luo, Beier, et al.
Published: (2025)
by: Luo, Beier, et al.
Published: (2025)
Reinforcement Learning for Out-of-Distribution Reasoning in LLMs: An Empirical Study on Diagnosis-Related Group Coding
by: Wang, Hanyin, et al.
Published: (2025)
by: Wang, Hanyin, et al.
Published: (2025)
Compositional Risk Minimization
by: Mahajan, Divyat, et al.
Published: (2024)
by: Mahajan, Divyat, et al.
Published: (2024)
Transfer Learning with Pre-trained Conditional Generative Models
by: Yamaguchi, Shin'ya, et al.
Published: (2022)
by: Yamaguchi, Shin'ya, et al.
Published: (2022)
Noncommutative $C^*$-algebra Net: Learning Neural Networks with Powerful Product Structure in $C^*$-algebra
by: Hataya, Ryuichiro, et al.
Published: (2023)
by: Hataya, Ryuichiro, et al.
Published: (2023)
Quantum Circuit $C^*$-algebra Net
by: Hashimoto, Yuka, et al.
Published: (2024)
by: Hashimoto, Yuka, et al.
Published: (2024)
Modeling the Data-Generating Process is Necessary for Out-of-Distribution Generalization
by: Kaur, Jivat Neet, et al.
Published: (2022)
by: Kaur, Jivat Neet, et al.
Published: (2022)
Sparsity and Out-of-Distribution Generalization
by: Aaronson, Scott, et al.
Published: (2026)
by: Aaronson, Scott, et al.
Published: (2026)
Navigating Potholes with Geometry-Aware Sharpness Minimization
by: Dufort-Labbé, Simon, et al.
Published: (2026)
by: Dufort-Labbé, Simon, et al.
Published: (2026)
Graph Generative Pre-trained Transformer
by: Chen, Xiaohui, et al.
Published: (2025)
by: Chen, Xiaohui, et al.
Published: (2025)
Subgraph Generation for Generalizing on Out-of-Distribution Links
by: Revolinsky, Jay, et al.
Published: (2025)
by: Revolinsky, Jay, et al.
Published: (2025)
Learning to Defer for Causal Discovery with Imperfect Experts
by: Clivio, Oscar, et al.
Published: (2025)
by: Clivio, Oscar, et al.
Published: (2025)
Decoupling Weighing and Selecting for Integrating Multiple Graph Pre-training Tasks
by: Fan, Tianyu, et al.
Published: (2024)
by: Fan, Tianyu, et al.
Published: (2024)
Generative Risk Minimization for Out-of-Distribution Generalization on Graphs
by: Wang, Song, et al.
Published: (2025)
by: Wang, Song, et al.
Published: (2025)
Utilizing Strategic Pre-training to Reduce Overfitting: Baguan -- A Pre-trained Weather Forecasting Model
by: Niu, Peisong, et al.
Published: (2025)
by: Niu, Peisong, et al.
Published: (2025)
Green MLOps to Green GenOps: An Empirical Study of Energy Consumption in Discriminative and Generative AI Operations
by: Sánchez-Mompó, Adrián, et al.
Published: (2025)
by: Sánchez-Mompó, Adrián, et al.
Published: (2025)
Enhancing Distribution and Label Consistency for Graph Out-of-Distribution Generalization
by: Wang, Song, et al.
Published: (2025)
by: Wang, Song, et al.
Published: (2025)
Out-of-Distribution Generalization for Neural Physics Solvers
by: Wei, Zhao, et al.
Published: (2026)
by: Wei, Zhao, et al.
Published: (2026)
Self-attention Networks Localize When QK-eigenspectrum Concentrates
by: Bao, Han, et al.
Published: (2024)
by: Bao, Han, et al.
Published: (2024)
Automatic Domain Adaptation by Transformers in In-Context Learning
by: Hataya, Ryuichiro, et al.
Published: (2024)
by: Hataya, Ryuichiro, et al.
Published: (2024)
Provable Data Scaling Law for Meta Learning via Complexity Minimization
by: Fukuchi, Kazuto, et al.
Published: (2026)
by: Fukuchi, Kazuto, et al.
Published: (2026)
Beyond Multi-Token Prediction: Pretraining LLMs with Future Summaries
by: Mahajan, Divyat, et al.
Published: (2025)
by: Mahajan, Divyat, et al.
Published: (2025)
Self-Calibrated Tuning of Vision-Language Models for Out-of-Distribution Detection
by: Yu, Geng, et al.
Published: (2024)
by: Yu, Geng, et al.
Published: (2024)
Onboard Out-of-Calibration Detection of Deep Learning Models using Conformal Prediction
by: Bhattacharjee, Protim, et al.
Published: (2024)
by: Bhattacharjee, Protim, et al.
Published: (2024)
DIVE: Subgraph Disagreement for Graph Out-of-Distribution Generalization
by: Sun, Xin, et al.
Published: (2024)
by: Sun, Xin, et al.
Published: (2024)
Investigating Out-of-Distribution Generalization of GNNs: An Architecture Perspective
by: Guo, Kai, et al.
Published: (2024)
by: Guo, Kai, et al.
Published: (2024)
Out-of-Distribution Generalization in Time Series: A Survey
by: Wu, Xin, et al.
Published: (2025)
by: Wu, Xin, et al.
Published: (2025)
Generalizing Graph Neural Networks on Out-Of-Distribution Graphs
by: Fan, Shaohua, et al.
Published: (2021)
by: Fan, Shaohua, et al.
Published: (2021)
Similar Items
-
Revisiting Generalization Measures Beyond IID: An Empirical Study under Distributional Shift
by: Nakai, Sora, et al.
Published: (2026) -
Empirical Analysis of Model Selection for Heterogeneous Causal Effect Estimation
by: Mahajan, Divyat, et al.
Published: (2022) -
Towards Understanding Variants of Invariant Risk Minimization through the Lens of Calibration
by: Yoshida, Kotaro, et al.
Published: (2024) -
Geometric Insights into Focal Loss: Reducing Curvature for Enhanced Model Calibration
by: Kimura, Masanari, et al.
Published: (2024) -
Adaptive Batch Sizes Using Non-Euclidean Gradient Noise Scales for Stochastic Sign and Spectral Descent
by: Naganuma, Hiroki, et al.
Published: (2026)