Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hager, Sophia, Mueller, David, Duh, Kevin, Andrews, Nicholas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Inducing Artificial Uncertainty in Language Models
von: Hager, Sophia, et al.
Veröffentlicht: (2026)
von: Hager, Sophia, et al.
Veröffentlicht: (2026)
Learning Extrapolative Sequence Transformations from Markov Chains
von: Hager, Sophia, et al.
Veröffentlicht: (2025)
von: Hager, Sophia, et al.
Veröffentlicht: (2025)
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
von: Xu, Tianyang, et al.
Veröffentlicht: (2024)
von: Xu, Tianyang, et al.
Veröffentlicht: (2024)
DOTResize: Reducing LLM Width via Discrete Optimal Transport-based Neuron Merging
von: Verma, Neha, et al.
Veröffentlicht: (2025)
von: Verma, Neha, et al.
Veröffentlicht: (2025)
Merging Feed-Forward Sublayers for Compressed Transformers
von: Verma, Neha, et al.
Veröffentlicht: (2025)
von: Verma, Neha, et al.
Veröffentlicht: (2025)
Where does In-context Translation Happen in Large Language Models
von: Sia, Suzanna, et al.
Veröffentlicht: (2024)
von: Sia, Suzanna, et al.
Veröffentlicht: (2024)
Generating with Confidence: Uncertainty Quantification for Black-box Large Language Models
von: Lin, Zhen, et al.
Veröffentlicht: (2023)
von: Lin, Zhen, et al.
Veröffentlicht: (2023)
Uncertainty in Semantic Language Modeling with PIXELS
von: Radu, Stefania, et al.
Veröffentlicht: (2025)
von: Radu, Stefania, et al.
Veröffentlicht: (2025)
Recurrent Confidence Chain: Temporal-Aware Uncertainty Quantification in Large Language Models
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
Confidence-aware Self-Semantic Distillation on Knowledge Graph Embedding
von: Liu, Yichen, et al.
Veröffentlicht: (2022)
von: Liu, Yichen, et al.
Veröffentlicht: (2022)
LLMs Should Express Uncertainty Explicitly
von: Guo, Junyu, et al.
Veröffentlicht: (2026)
von: Guo, Junyu, et al.
Veröffentlicht: (2026)
Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language
von: Kumar, Vinayshekhar Bannihatti, et al.
Veröffentlicht: (2026)
von: Kumar, Vinayshekhar Bannihatti, et al.
Veröffentlicht: (2026)
Confidence Preservation Property in Knowledge Distillation Abstractions
von: Vengertsev, Dmitry, et al.
Veröffentlicht: (2024)
von: Vengertsev, Dmitry, et al.
Veröffentlicht: (2024)
Show Your Work with Confidence: Confidence Bands for Tuning Curves
von: Lourie, Nicholas, et al.
Veröffentlicht: (2023)
von: Lourie, Nicholas, et al.
Veröffentlicht: (2023)
Confidence in the Reasoning of Large Language Models
von: Pawitan, Yudi, et al.
Veröffentlicht: (2024)
von: Pawitan, Yudi, et al.
Veröffentlicht: (2024)
Teaching Language Models to Faithfully Express their Uncertainty
von: Eikema, Bryan, et al.
Veröffentlicht: (2025)
von: Eikema, Bryan, et al.
Veröffentlicht: (2025)
Confidence over Time: Confidence Calibration with Temporal Logic for Large Language Model Reasoning
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
QA-Calibration of Language Model Confidence Scores
von: Manggala, Putra, et al.
Veröffentlicht: (2024)
von: Manggala, Putra, et al.
Veröffentlicht: (2024)
Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models
von: Eleraqi, Amr, et al.
Veröffentlicht: (2026)
von: Eleraqi, Amr, et al.
Veröffentlicht: (2026)
Data Kernel Perspective Space Performance Guarantees for Synthetic Data from Transformer Models
von: Browder, Michael, et al.
Veröffentlicht: (2026)
von: Browder, Michael, et al.
Veröffentlicht: (2026)
Benchmarking Uncertainty Calibration in Large Language Model Long-Form Question Answering
von: Müller, Philip, et al.
Veröffentlicht: (2026)
von: Müller, Philip, et al.
Veröffentlicht: (2026)
Self-Training Large Language Models with Confident Reasoning
von: Jang, Hyosoon, et al.
Veröffentlicht: (2025)
von: Jang, Hyosoon, et al.
Veröffentlicht: (2025)
Can Authorship Attribution Models Distinguish Speakers in Speech Transcripts?
von: Aggazzotti, Cristina, et al.
Veröffentlicht: (2023)
von: Aggazzotti, Cristina, et al.
Veröffentlicht: (2023)
Language Models Optimized to Fool Detectors Still Have a Distinct Style (And How to Change It)
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2025)
von: Soto, Rafael Rivera, et al.
Veröffentlicht: (2025)
What Formal Languages Can Transformers Express? A Survey
von: Strobl, Lena, et al.
Veröffentlicht: (2023)
von: Strobl, Lena, et al.
Veröffentlicht: (2023)
Neuron-Level Knowledge Attribution in Large Language Models
von: Yu, Zeping, et al.
Veröffentlicht: (2023)
von: Yu, Zeping, et al.
Veröffentlicht: (2023)
Confidence Calibration in Large Language Model-Based Entity Matching
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
Using Source-Side Confidence Estimation for Reliable Translation into Unfamiliar Languages
von: Sible, Kenneth J., et al.
Veröffentlicht: (2025)
von: Sible, Kenneth J., et al.
Veröffentlicht: (2025)
Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models
von: Zhao, Siyan, et al.
Veröffentlicht: (2026)
von: Zhao, Siyan, et al.
Veröffentlicht: (2026)
Semantic Token Clustering for Efficient Uncertainty Quantification in Large Language Models
von: Cao, Qi, et al.
Veröffentlicht: (2026)
von: Cao, Qi, et al.
Veröffentlicht: (2026)
Improving Uncertainty Quantification in Large Language Models via Semantic Embeddings
von: Grewal, Yashvir S., et al.
Veröffentlicht: (2024)
von: Grewal, Yashvir S., et al.
Veröffentlicht: (2024)
Benchmarking Distilled Language Models: Performance and Efficiency in Resource-Constrained Settings
von: Wani, Sachin Gopal, et al.
Veröffentlicht: (2026)
von: Wani, Sachin Gopal, et al.
Veröffentlicht: (2026)
Confidence Regulation Neurons in Language Models
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
von: Stolfo, Alessandro, et al.
Veröffentlicht: (2024)
Self-ensemble: Mitigating Confidence Mis-calibration for Large Language Models
von: Xu, Zicheng, et al.
Veröffentlicht: (2025)
von: Xu, Zicheng, et al.
Veröffentlicht: (2025)
ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models
von: Li, Chen, et al.
Veröffentlicht: (2026)
von: Li, Chen, et al.
Veröffentlicht: (2026)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
DiLM: Distilling Dataset into Language Model for Text-level Dataset Distillation
von: Maekawa, Aru, et al.
Veröffentlicht: (2024)
von: Maekawa, Aru, et al.
Veröffentlicht: (2024)
Function Vectors in Large Language Models
von: Todd, Eric, et al.
Veröffentlicht: (2023)
von: Todd, Eric, et al.
Veröffentlicht: (2023)
Dist2ill: Distributional Distillation for One-Pass Uncertainty Estimation in Large Language Models
von: Zhao, Yicong, et al.
Veröffentlicht: (2025)
von: Zhao, Yicong, et al.
Veröffentlicht: (2025)
Confidence Geometry Reveals Trace-Level Correctness in Large Language Model Reasoning
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Inducing Artificial Uncertainty in Language Models
von: Hager, Sophia, et al.
Veröffentlicht: (2026) -
Learning Extrapolative Sequence Transformations from Markov Chains
von: Hager, Sophia, et al.
Veröffentlicht: (2025) -
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
von: Xu, Tianyang, et al.
Veröffentlicht: (2024) -
DOTResize: Reducing LLM Width via Discrete Optimal Transport-based Neuron Merging
von: Verma, Neha, et al.
Veröffentlicht: (2025) -
Merging Feed-Forward Sublayers for Compressed Transformers
von: Verma, Neha, et al.
Veröffentlicht: (2025)