Self-ensemble: Mitigating Confidence Mis-calibration for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Zicheng, Wang, Guanchu, Zheng, Guangyao, Chuang, Yu-Neng, Szalay, Alexander, Hu, Xia, Braverman, Vladimir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DTS: Enhancing Large Reasoning Models via Decoding Tree Sketching
von: Xu, Zicheng, et al.
Veröffentlicht: (2025)
von: Xu, Zicheng, et al.
Veröffentlicht: (2025)
Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models
von: Luo, Feng, et al.
Veröffentlicht: (2026)
von: Luo, Feng, et al.
Veröffentlicht: (2026)
Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2025)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2025)
FaithLM: Towards Faithful Explanations for Large Language Models
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
AutoL2S: Auto Long-Short Reasoning for Efficient Large Language Models
von: Luo, Feng, et al.
Veröffentlicht: (2025)
von: Luo, Feng, et al.
Veröffentlicht: (2025)
Taylor Unswift: Secured Weight Release for Large Language Models via Taylor Expansion
von: Wang, Guanchu, et al.
Veröffentlicht: (2024)
von: Wang, Guanchu, et al.
Veröffentlicht: (2024)
Learning to Route LLMs with Confidence Tokens
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
Self-Training Large Language Models with Confident Reasoning
von: Jang, Hyosoon, et al.
Veröffentlicht: (2025)
von: Jang, Hyosoon, et al.
Veröffentlicht: (2025)
ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models
von: Li, Chen, et al.
Veröffentlicht: (2026)
von: Li, Chen, et al.
Veröffentlicht: (2026)
Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
von: Sui, Yang, et al.
Veröffentlicht: (2025)
von: Sui, Yang, et al.
Veröffentlicht: (2025)
Learning to Compress Prompt in Natural Language Formats
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024)
Digger: Detecting Copyright Content Mis-usage in Large Language Model Training
von: Li, Haodong, et al.
Veröffentlicht: (2024)
von: Li, Haodong, et al.
Veröffentlicht: (2024)
Investigating the Feasibility of Mitigating Potential Copyright Infringement via Large Language Model Unlearning
von: Dou, Guangyao
Veröffentlicht: (2024)
von: Dou, Guangyao
Veröffentlicht: (2024)
Confidence in the Reasoning of Large Language Models
von: Pawitan, Yudi, et al.
Veröffentlicht: (2024)
von: Pawitan, Yudi, et al.
Veröffentlicht: (2024)
Assessing and Enhancing Large Language Models in Rare Disease Question-answering
von: Wang, Guanchu, et al.
Veröffentlicht: (2024)
von: Wang, Guanchu, et al.
Veröffentlicht: (2024)
Winner-Take-All Column Row Sampling for Memory Efficient Adaptation of Language Model
von: Liu, Zirui, et al.
Veröffentlicht: (2023)
von: Liu, Zirui, et al.
Veröffentlicht: (2023)
C$^2$GSPG: Confidence-calibrated Group Sequence Policy Gradient towards Self-aware Reasoning
von: Liu, Haotian, et al.
Veröffentlicht: (2025)
von: Liu, Haotian, et al.
Veröffentlicht: (2025)
Can Large Language Models Solve Robot Routing?
von: Huang, Zhehui, et al.
Veröffentlicht: (2024)
von: Huang, Zhehui, et al.
Veröffentlicht: (2024)
SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
Confidence-Aware Sub-Structure Beam Search (CABS): Mitigating Hallucination in Structured Data Generation with Large Language Models
von: Wei, Chengwei, et al.
Veröffentlicht: (2024)
von: Wei, Chengwei, et al.
Veröffentlicht: (2024)
Confidence over Time: Confidence Calibration with Temporal Logic for Large Language Model Reasoning
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
Large Language Models for Code Summarization
von: Szalontai, Balázs, et al.
Veröffentlicht: (2024)
von: Szalontai, Balázs, et al.
Veröffentlicht: (2024)
Lookback Lens: Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2024)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2024)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache
von: Liu, Zirui, et al.
Veröffentlicht: (2024)
von: Liu, Zirui, et al.
Veröffentlicht: (2024)
Confidence Calibration in Large Language Model-Based Entity Matching
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
Think Before You Prune: Self-Reflective Structured Pruning for Reasoning Language Models
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
von: Mündler, Niels, et al.
Veröffentlicht: (2023)
Confidence-Modulated Speculative Decoding for Large Language Models
von: Sen, Jaydip, et al.
Veröffentlicht: (2025)
von: Sen, Jaydip, et al.
Veröffentlicht: (2025)
Generating with Confidence: Uncertainty Quantification for Black-box Large Language Models
von: Lin, Zhen, et al.
Veröffentlicht: (2023)
von: Lin, Zhen, et al.
Veröffentlicht: (2023)
Mitigating Hallucinations in Large Language Models via Causal Reasoning
von: Li, Yuangang, et al.
Veröffentlicht: (2025)
von: Li, Yuangang, et al.
Veröffentlicht: (2025)
CARE-RFT: Confidence-Anchored Reinforcement Finetuning for Reliable Reasoning in Large Language Models
von: Li, Shuozhe, et al.
Veröffentlicht: (2026)
von: Li, Shuozhe, et al.
Veröffentlicht: (2026)
Confidence Geometry Reveals Trace-Level Correctness in Large Language Model Reasoning
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
Recurrent Confidence Chain: Temporal-Aware Uncertainty Quantification in Large Language Models
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
Mitigating Selection Bias in Large Language Models via Permutation-Aware GRPO
von: Zheng, Jinquan, et al.
Veröffentlicht: (2026)
von: Zheng, Jinquan, et al.
Veröffentlicht: (2026)
Fast and Effective Weight Update for Pruned Large Language Models
von: Boža, Vladimír
Veröffentlicht: (2024)
von: Boža, Vladimír
Veröffentlicht: (2024)
Towards Fair Medical AI: Adversarial Debiasing of 3D CT Foundation Embeddings
von: Zheng, Guangyao, et al.
Veröffentlicht: (2025)
von: Zheng, Guangyao, et al.
Veröffentlicht: (2025)
The Role of Deductive and Inductive Reasoning in Large Language Models
von: Cai, Chengkun, et al.
Veröffentlicht: (2024)
von: Cai, Chengkun, et al.
Veröffentlicht: (2024)
Alignment Pretraining: AI Discourse Causes Self-Fulfilling (Mis)alignment
von: Tice, Cameron, et al.
Veröffentlicht: (2026)
von: Tice, Cameron, et al.
Veröffentlicht: (2026)
Bias in Large Language Models: Origin, Evaluation, and Mitigation
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
von: Guo, Yufei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DTS: Enhancing Large Reasoning Models via Decoding Tree Sketching
von: Xu, Zicheng, et al.
Veröffentlicht: (2025) -
Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models
von: Luo, Feng, et al.
Veröffentlicht: (2026) -
Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2025) -
FaithLM: Towards Faithful Explanations for Large Language Models
von: Chuang, Yu-Neng, et al.
Veröffentlicht: (2024) -
AutoL2S: Auto Long-Short Reasoning for Efficient Large Language Models
von: Luo, Feng, et al.
Veröffentlicht: (2025)