All Models Are Miscalibrated, But Some Less So: Comparing Calibration with Conditional Mean Operators
Fuente:
arXiv
Saved in:
| Main Authors: | Moskvichev, Peter, Sejdinovic, Dino |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Measuring Differences between Conditional Distributions using Kernel Embeddings
by: Moskvichev, Peter, et al.
Published: (2026)
by: Moskvichev, Peter, et al.
Published: (2026)
Neural-Kernel Conditional Mean Embeddings
by: Shimizu, Eiki, et al.
Published: (2024)
by: Shimizu, Eiki, et al.
Published: (2024)
An Overview of Causal Inference using Kernel Embeddings
by: Sejdinovic, Dino
Published: (2024)
by: Sejdinovic, Dino
Published: (2024)
ActiveCQ: Active Estimation of Causal Quantities
by: Gao, Erdun, et al.
Published: (2025)
by: Gao, Erdun, et al.
Published: (2025)
Bayesian Adaptive Calibration and Optimal Design
by: Oliveira, Rafael, et al.
Published: (2024)
by: Oliveira, Rafael, et al.
Published: (2024)
Causal-EPIG: A Prediction-Oriented Active Learning Framework for CATE Estimation
by: Gao, Erdun, et al.
Published: (2025)
by: Gao, Erdun, et al.
Published: (2025)
Label Distribution Learning using the Squared Neural Family on the Probability Simplex
by: Zhang, Daokun, et al.
Published: (2024)
by: Zhang, Daokun, et al.
Published: (2024)
FaIRGP: A Bayesian Energy Balance Model for Surface Temperatures Emulation
by: Bouabid, Shahine, et al.
Published: (2023)
by: Bouabid, Shahine, et al.
Published: (2023)
The Generalised Kernel Covariance Measure
by: Bergen, Luca, et al.
Published: (2026)
by: Bergen, Luca, et al.
Published: (2026)
Exact, Fast and Expressive Poisson Point Processes via Squared Neural Families
by: Tsuchida, Russell, et al.
Published: (2024)
by: Tsuchida, Russell, et al.
Published: (2024)
A Kernel Test for Causal Association via Noise Contrastive Backdoor Adjustment
by: Hu, Robert, et al.
Published: (2021)
by: Hu, Robert, et al.
Published: (2021)
Near-Optimal Approximations for Bayesian Inference in Function Space
by: Wild, Veit, et al.
Published: (2025)
by: Wild, Veit, et al.
Published: (2025)
When Individually Calibrated Models Become Collectively Miscalibrated
by: Wang, Zhaohui
Published: (2026)
by: Wang, Zhaohui
Published: (2026)
Mapping from Meaning: Addressing the Miscalibration of Prompt-Sensitive Language Models
by: Cox, Kyle, et al.
Published: (2025)
by: Cox, Kyle, et al.
Published: (2025)
Squared families: Searching beyond regular probability models
by: Tsuchida, Russell, et al.
Published: (2025)
by: Tsuchida, Russell, et al.
Published: (2025)
Indirect Query Bayesian Optimization with Integrated Feedback
by: Zhang, Mengyan, et al.
Published: (2024)
by: Zhang, Mengyan, et al.
Published: (2024)
Instrumental and Proximal Causal Inference with Gaussian Processes
by: Zhang, Yuqi, et al.
Published: (2026)
by: Zhang, Yuqi, et al.
Published: (2026)
Credal Two-Sample Tests of Epistemic Uncertainty
by: Chau, Siu Lun, et al.
Published: (2024)
by: Chau, Siu Lun, et al.
Published: (2024)
Gaussian Processes and Reproducing Kernels: Connections and Equivalences
by: Kanagawa, Motonobu, et al.
Published: (2025)
by: Kanagawa, Motonobu, et al.
Published: (2025)
Understanding and Mitigating Miscalibration in Prompt Tuning for Vision-Language Models
by: Wang, Shuoyuan, et al.
Published: (2024)
by: Wang, Shuoyuan, et al.
Published: (2024)
Giga-scale Kernel Matrix Vector Multiplication on GPU
by: Hu, Robert, et al.
Published: (2022)
by: Hu, Robert, et al.
Published: (2022)
Discovery of Hidden Miscalibration Regimes
by: Kobalczyk, Katarzyna, et al.
Published: (2026)
by: Kobalczyk, Katarzyna, et al.
Published: (2026)
Large Language Models are Miscalibrated In-Context Learners
by: Li, Chengzu, et al.
Published: (2023)
by: Li, Chengzu, et al.
Published: (2023)
Observationally Informed Adaptive Causal Experimental Design
by: Gao, Erdun, et al.
Published: (2026)
by: Gao, Erdun, et al.
Published: (2026)
Less Memory Means smaller GPUs: Backpropagation with Compressed Activations
by: Barley, Daniel, et al.
Published: (2024)
by: Barley, Daniel, et al.
Published: (2024)
A Course Correction in Steerability Evaluation: Revealing Miscalibration and Side Effects in LLMs
by: Chang, Trenton, et al.
Published: (2025)
by: Chang, Trenton, et al.
Published: (2025)
On Overcoming Miscalibrated Conversational Priors in LLM-based Chatbots
by: Herlihy, Christine, et al.
Published: (2024)
by: Herlihy, Christine, et al.
Published: (2024)
Do not trust what you trust: Miscalibration in Semi-supervised Learning
by: Mishra, Shambhavi, et al.
Published: (2024)
by: Mishra, Shambhavi, et al.
Published: (2024)
All AI Models are Wrong, but Some are Optimal
by: Anand, Akhil S, et al.
Published: (2025)
by: Anand, Akhil S, et al.
Published: (2025)
False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compression in LLMs
by: Okutomi, Akira
Published: (2025)
by: Okutomi, Akira
Published: (2025)
Adjusting Regression Models for Conditional Uncertainty Calibration
by: Gao, Ruijiang, et al.
Published: (2024)
by: Gao, Ruijiang, et al.
Published: (2024)
Some Attention is All You Need for Retrieval
by: Michalak, Felix, et al.
Published: (2025)
by: Michalak, Felix, et al.
Published: (2025)
Pairwise Optimal Transports for Training All-to-All Flow-Based Condition Transfer Model
by: Ikeda, Kotaro, et al.
Published: (2025)
by: Ikeda, Kotaro, et al.
Published: (2025)
Soft Mean Expected Calibration Error (SMECE): A Calibration Metric for Probabilistic Labels
by: Leznik, Michael
Published: (2026)
by: Leznik, Michael
Published: (2026)
Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
by: Plaut, Benjamin, et al.
Published: (2024)
by: Plaut, Benjamin, et al.
Published: (2024)
Modeling All Response Surfaces in One for Conditional Search Spaces
by: Li, Jiaxing, et al.
Published: (2025)
by: Li, Jiaxing, et al.
Published: (2025)
Empirically Calibrated Conditional Independence Tests
by: Pan, Milleno, et al.
Published: (2026)
by: Pan, Milleno, et al.
Published: (2026)
Distributional Bellman Operators over Mean Embeddings
by: Wenliang, Li Kevin, et al.
Published: (2023)
by: Wenliang, Li Kevin, et al.
Published: (2023)
Coarse-to-Fine Concept Bottleneck Models
by: Panousis, Konstantinos P., et al.
Published: (2023)
by: Panousis, Konstantinos P., et al.
Published: (2023)
Say Less, Mean More: Leveraging Pragmatics in Retrieval-Augmented Generation
by: Riaz, Haris, et al.
Published: (2025)
by: Riaz, Haris, et al.
Published: (2025)
Similar Items
-
Measuring Differences between Conditional Distributions using Kernel Embeddings
by: Moskvichev, Peter, et al.
Published: (2026) -
Neural-Kernel Conditional Mean Embeddings
by: Shimizu, Eiki, et al.
Published: (2024) -
An Overview of Causal Inference using Kernel Embeddings
by: Sejdinovic, Dino
Published: (2024) -
ActiveCQ: Active Estimation of Causal Quantities
by: Gao, Erdun, et al.
Published: (2025) -
Bayesian Adaptive Calibration and Optimal Design
by: Oliveira, Rafael, et al.
Published: (2024)