Believe Your Model: Distribution-Guided Confidence Calibration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Xizhong, Zhang, Haotian, Wang, Huiming, Song, Mofei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From the Inside Out: Progressive Distribution Refinement for Confidence Calibration
von: Yang, Xizhong, et al.
Veröffentlicht: (2026)
von: Yang, Xizhong, et al.
Veröffentlicht: (2026)
Semantic Bridging Domains: Pseudo-Source as Test-Time Connector
von: Yang, Xizhong, et al.
Veröffentlicht: (2026)
von: Yang, Xizhong, et al.
Veröffentlicht: (2026)
To Believe or Not to Believe Your LLM
von: Yadkori, Yasin Abbasi, et al.
Veröffentlicht: (2024)
von: Yadkori, Yasin Abbasi, et al.
Veröffentlicht: (2024)
Similarity and Dissimilarity Guided Co-association Matrix Construction for Ensemble Clustering
von: Zhang, Xu, et al.
Veröffentlicht: (2024)
von: Zhang, Xu, et al.
Veröffentlicht: (2024)
Your Pre-trained LLM is Secretly an Unsupervised Confidence Calibrator
von: Luo, Beier, et al.
Veröffentlicht: (2025)
von: Luo, Beier, et al.
Veröffentlicht: (2025)
Balancing the Reasoning Load: Difficulty-Differentiated Policy Optimization with Length Redistribution for Efficient and Robust Reinforcement Learning
von: Xia, Yinan, et al.
Veröffentlicht: (2026)
von: Xia, Yinan, et al.
Veröffentlicht: (2026)
Confidence-Aware Multi-Field Model Calibration
von: Zhao, Yuang, et al.
Veröffentlicht: (2024)
von: Zhao, Yuang, et al.
Veröffentlicht: (2024)
DP$^2$-NILM: A Distributed and Privacy-preserving Framework for Non-intrusive Load Monitoring
von: Dai, Shuang, et al.
Veröffentlicht: (2022)
von: Dai, Shuang, et al.
Veröffentlicht: (2022)
Confidence Calibration in Large Language Models
von: Michael, Noam, et al.
Veröffentlicht: (2026)
von: Michael, Noam, et al.
Veröffentlicht: (2026)
Confidence Calibration in Vision-Language-Action Models
von: Zollo, Thomas P, et al.
Veröffentlicht: (2025)
von: Zollo, Thomas P, et al.
Veröffentlicht: (2025)
CHARM: Calibrating Reward Models With Chatbot Arena Scores
von: Zhu, Xiao, et al.
Veröffentlicht: (2025)
von: Zhu, Xiao, et al.
Veröffentlicht: (2025)
Enhance GNNs with Reliable Confidence Estimation via Adversarial Calibration Learning
von: Wang, Yilong, et al.
Veröffentlicht: (2025)
von: Wang, Yilong, et al.
Veröffentlicht: (2025)
Prior Distribution and Model Confidence
von: Kazanskii, Maksim, et al.
Veröffentlicht: (2025)
von: Kazanskii, Maksim, et al.
Veröffentlicht: (2025)
QA-Calibration of Language Model Confidence Scores
von: Manggala, Putra, et al.
Veröffentlicht: (2024)
von: Manggala, Putra, et al.
Veröffentlicht: (2024)
Efficient RLVR Training via Weighted Mutual Information Data Selection
von: Zhou, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2026)
Show Your Work with Confidence: Confidence Bands for Tuning Curves
von: Lourie, Nicholas, et al.
Veröffentlicht: (2023)
von: Lourie, Nicholas, et al.
Veröffentlicht: (2023)
Confidence over Time: Confidence Calibration with Temporal Logic for Large Language Model Reasoning
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
Expectation Consistency Loss: Rethink Confidence Calibration under Covariate Shift
von: Dong, Jinzong, et al.
Veröffentlicht: (2026)
von: Dong, Jinzong, et al.
Veröffentlicht: (2026)
Conformal Calibration of Statistical Confidence Sets
von: Cabezas, Luben M. C., et al.
Veröffentlicht: (2024)
von: Cabezas, Luben M. C., et al.
Veröffentlicht: (2024)
DexSim2Real: Foundation Model-Guided Sim-to-Real Transfer for Generalizable Dexterous Manipulation
von: Zeng, Zijian, et al.
Veröffentlicht: (2026)
von: Zeng, Zijian, et al.
Veröffentlicht: (2026)
Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation
von: Zollo, Thomas, et al.
Veröffentlicht: (2026)
von: Zollo, Thomas, et al.
Veröffentlicht: (2026)
Confidence Calibration of Classifiers with Many Classes
von: LeCoz, Adrien, et al.
Veröffentlicht: (2024)
von: LeCoz, Adrien, et al.
Veröffentlicht: (2024)
Learn to Guide Your Diffusion Model
von: Galashov, Alexandre, et al.
Veröffentlicht: (2025)
von: Galashov, Alexandre, et al.
Veröffentlicht: (2025)
Confidence Calibration in Large Language Model-Based Entity Matching
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
Graph-based Confidence Calibration for Large Language Models
von: Li, Yukun, et al.
Veröffentlicht: (2024)
von: Li, Yukun, et al.
Veröffentlicht: (2024)
Process Supervision of Confidence Margin for Calibrated LLM Reasoning
von: Wang, Liaoyaqi, et al.
Veröffentlicht: (2026)
von: Wang, Liaoyaqi, et al.
Veröffentlicht: (2026)
Confidence Calibration under Ambiguous Ground Truth
von: Tao, Linwei, et al.
Veröffentlicht: (2026)
von: Tao, Linwei, et al.
Veröffentlicht: (2026)
MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination
von: Armstrong, Joss
Veröffentlicht: (2026)
von: Armstrong, Joss
Veröffentlicht: (2026)
Calibration Attacks: A Comprehensive Study of Adversarial Attacks on Model Confidence
von: Obadinma, Stephen, et al.
Veröffentlicht: (2024)
von: Obadinma, Stephen, et al.
Veröffentlicht: (2024)
A Confidence Interval for the $\ell_2$ Expected Calibration Error
von: Sun, Yan, et al.
Veröffentlicht: (2024)
von: Sun, Yan, et al.
Veröffentlicht: (2024)
Semantic-Aware Confidence Calibration for Automated Audio Captioning
von: Dunker, Lucas, et al.
Veröffentlicht: (2025)
von: Dunker, Lucas, et al.
Veröffentlicht: (2025)
Is Your Explanation Reliable: Confidence-Aware Explanation on Graph Neural Networks
von: Zhang, Jiaxing, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaxing, et al.
Veröffentlicht: (2025)
Identifiable Deep Latent Variable Models for MNAR Data
von: Xie, Huiming, et al.
Veröffentlicht: (2026)
von: Xie, Huiming, et al.
Veröffentlicht: (2026)
Improving the Throughput of Diffusion-based Large Language Models via a Training-Free Confidence-Aware Calibration
von: Shen, Jucheng, et al.
Veröffentlicht: (2025)
von: Shen, Jucheng, et al.
Veröffentlicht: (2025)
CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models
von: Zhu, Xiao, et al.
Veröffentlicht: (2026)
von: Zhu, Xiao, et al.
Veröffentlicht: (2026)
Combining Priors with Experience: Confidence Calibration Based on Binomial Process Modeling
von: Dong, Jinzong, et al.
Veröffentlicht: (2024)
von: Dong, Jinzong, et al.
Veröffentlicht: (2024)
Calibrating Generative Models to Distributional Constraints
von: Smith, Henry D., et al.
Veröffentlicht: (2025)
von: Smith, Henry D., et al.
Veröffentlicht: (2025)
MICE for CATs: Model-Internal Confidence Estimation for Calibrating Agents with Tools
von: Subramani, Nishant, et al.
Veröffentlicht: (2025)
von: Subramani, Nishant, et al.
Veröffentlicht: (2025)
Performance Estimation in Binary Classification Using Calibrated Confidence
von: Kivimäki, Juhani, et al.
Veröffentlicht: (2025)
von: Kivimäki, Juhani, et al.
Veröffentlicht: (2025)
SpecBound: Adaptive Bounded Self-Speculation with Layer-wise Confidence Calibration
von: Wen, Zhuofan, et al.
Veröffentlicht: (2026)
von: Wen, Zhuofan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
From the Inside Out: Progressive Distribution Refinement for Confidence Calibration
von: Yang, Xizhong, et al.
Veröffentlicht: (2026) -
Semantic Bridging Domains: Pseudo-Source as Test-Time Connector
von: Yang, Xizhong, et al.
Veröffentlicht: (2026) -
To Believe or Not to Believe Your LLM
von: Yadkori, Yasin Abbasi, et al.
Veröffentlicht: (2024) -
Similarity and Dissimilarity Guided Co-association Matrix Construction for Ensemble Clustering
von: Zhang, Xu, et al.
Veröffentlicht: (2024) -
Your Pre-trained LLM is Secretly an Unsupervised Confidence Calibrator
von: Luo, Beier, et al.
Veröffentlicht: (2025)