Thermometer: Towards Universal Calibration for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Maohao, Das, Subhro, Greenewald, Kristjan, Sattigeri, Prasanna, Wornell, Gregory, Ghosh, Soumya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Are Uncertainty Quantification Capabilities of Evidential Deep Learning a Mirage?
von: Shen, Maohao, et al.
Veröffentlicht: (2024)
von: Shen, Maohao, et al.
Veröffentlicht: (2024)
Large Language Model Confidence Estimation via Black-Box Access
von: Pedapati, Tejaswini, et al.
Veröffentlicht: (2024)
von: Pedapati, Tejaswini, et al.
Veröffentlicht: (2024)
When in Doubt, Cascade: Towards Building Efficient and Capable Guardrails
von: Nagireddy, Manish, et al.
Veröffentlicht: (2024)
von: Nagireddy, Manish, et al.
Veröffentlicht: (2024)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
von: Jiang, Mingjian, et al.
Veröffentlicht: (2024)
von: Jiang, Mingjian, et al.
Veröffentlicht: (2024)
Extending Beacon to Hindi: Cultural Adaptation Drives Cross-Lingual Sycophancy
von: Sattigeri, Sarthak
Veröffentlicht: (2026)
von: Sattigeri, Sarthak
Veröffentlicht: (2026)
Decocted Experience Improves Test-Time Inference in LLM Agents
von: Shen, Maohao, et al.
Veröffentlicht: (2026)
von: Shen, Maohao, et al.
Veröffentlicht: (2026)
Distributional Process Reward Models: Calibrated Prediction of Future Rewards via Conditional Optimal Transport
von: Ma, Rachel, et al.
Veröffentlicht: (2026)
von: Ma, Rachel, et al.
Veröffentlicht: (2026)
Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search
von: Shen, Maohao, et al.
Veröffentlicht: (2025)
von: Shen, Maohao, et al.
Veröffentlicht: (2025)
On the Entropy Calibration of Language Models
von: Cao, Steven, et al.
Veröffentlicht: (2025)
von: Cao, Steven, et al.
Veröffentlicht: (2025)
Value Alignment from Unstructured Text
von: Padhi, Inkit, et al.
Veröffentlicht: (2024)
von: Padhi, Inkit, et al.
Veröffentlicht: (2024)
CTG-KrEW: Generating Synthetic Structured Contextually Correlated Content by Conditional Tabular GAN with K-Means Clustering and Efficient Word Embedding
von: Samanta, Riya, et al.
Veröffentlicht: (2024)
von: Samanta, Riya, et al.
Veröffentlicht: (2024)
Calibrated Large Language Models for Binary Question Answering
von: Giovannotti, Patrizio, et al.
Veröffentlicht: (2024)
von: Giovannotti, Patrizio, et al.
Veröffentlicht: (2024)
RL Tango: Reinforcing Generator and Verifier Together for Language Reasoning
von: Zha, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zha, Kaiwen, et al.
Veröffentlicht: (2025)
HierRouter: Coordinated Routing of Specialized Large Language Models via Reinforcement Learning
von: Gupta, Nikunj, et al.
Veröffentlicht: (2025)
von: Gupta, Nikunj, et al.
Veröffentlicht: (2025)
SLOT: Structuring the Output of Large Language Models
von: Wang, Darren Yow-Bang, et al.
Veröffentlicht: (2025)
von: Wang, Darren Yow-Bang, et al.
Veröffentlicht: (2025)
OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models
von: Shao, Wenqi, et al.
Veröffentlicht: (2023)
von: Shao, Wenqi, et al.
Veröffentlicht: (2023)
Confidence Calibration in Large Language Model-Based Entity Matching
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
von: Kamsteeg, Iris, et al.
Veröffentlicht: (2025)
Compress then Serve: Serving Thousands of LoRA Adapters with Little Overhead
von: Brüel-Gabrielsson, Rickard, et al.
Veröffentlicht: (2024)
von: Brüel-Gabrielsson, Rickard, et al.
Veröffentlicht: (2024)
Fine-tuning vs. In-context Learning in Large Language Models: A Formal Language Learning Perspective
von: Ghosh, Bishwamittra, et al.
Veröffentlicht: (2026)
von: Ghosh, Bishwamittra, et al.
Veröffentlicht: (2026)
Efficient Alignment of Large Language Models via Data Sampling
von: Khera, Amrit, et al.
Veröffentlicht: (2024)
von: Khera, Amrit, et al.
Veröffentlicht: (2024)
Satori-SWE: Evolutionary Test-Time Scaling for Sample-Efficient Software Engineering
von: Zeng, Guangtao, et al.
Veröffentlicht: (2025)
von: Zeng, Guangtao, et al.
Veröffentlicht: (2025)
Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models
von: Jain, Neel, et al.
Veröffentlicht: (2024)
von: Jain, Neel, et al.
Veröffentlicht: (2024)
Benchmarking Uncertainty Calibration in Large Language Model Long-Form Question Answering
von: Müller, Philip, et al.
Veröffentlicht: (2026)
von: Müller, Philip, et al.
Veröffentlicht: (2026)
Beyond Mode-Seeking RL: Trajectory-Balance Post-Training for Diffusion Language Models
von: Ahmadi, Saba, et al.
Veröffentlicht: (2026)
von: Ahmadi, Saba, et al.
Veröffentlicht: (2026)
Efficient Multi-Adapter LLM Serving via Cross-Model KV-Cache Reuse with Activated LoRA
von: Li, Allison, et al.
Veröffentlicht: (2025)
von: Li, Allison, et al.
Veröffentlicht: (2025)
Alignment Studio: Aligning Large Language Models to Particular Contextual Regulations
von: Achintalwar, Swapnaja, et al.
Veröffentlicht: (2024)
von: Achintalwar, Swapnaja, et al.
Veröffentlicht: (2024)
Do Large Language Models Know How Much They Know?
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
Partially Observed Trajectory Inference using Optimal Transport and a Dynamics Prior
von: Gu, Anming, et al.
Veröffentlicht: (2024)
von: Gu, Anming, et al.
Veröffentlicht: (2024)
Private Continuous-Time Synthetic Trajectory Generation via Mean-Field Langevin Dynamics
von: Gu, Anming, et al.
Veröffentlicht: (2025)
von: Gu, Anming, et al.
Veröffentlicht: (2025)
Confidence over Time: Confidence Calibration with Temporal Logic for Large Language Model Reasoning
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
von: Mao, Zhenjiang, et al.
Veröffentlicht: (2026)
Beware of Calibration Data for Pruning Large Language Models
von: Ji, Yixin, et al.
Veröffentlicht: (2024)
von: Ji, Yixin, et al.
Veröffentlicht: (2024)
Calibrating Large Language Models Using Their Generations Only
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
On Calibration of Large Language Models: From Response To Capability
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026)
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
von: Park, Young-Jin, et al.
Veröffentlicht: (2025)
von: Park, Young-Jin, et al.
Veröffentlicht: (2025)
SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning
von: Hu, Chenzhi, et al.
Veröffentlicht: (2026)
von: Hu, Chenzhi, et al.
Veröffentlicht: (2026)
Beyond Size: How Gradients Shape Pruning Decisions in Large Language Models
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2023)
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2023)
Fundamental Safety-Capability Trade-offs in Fine-tuning Large Language Models
von: Chen, Pin-Yu, et al.
Veröffentlicht: (2025)
von: Chen, Pin-Yu, et al.
Veröffentlicht: (2025)
QA-Calibration of Language Model Confidence Scores
von: Manggala, Putra, et al.
Veröffentlicht: (2024)
von: Manggala, Putra, et al.
Veröffentlicht: (2024)
Geometry-Calibrated Conformal Abstention for Language Models
von: Xu, Rui, et al.
Veröffentlicht: (2026)
von: Xu, Rui, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Are Uncertainty Quantification Capabilities of Evidential Deep Learning a Mirage?
von: Shen, Maohao, et al.
Veröffentlicht: (2024) -
Large Language Model Confidence Estimation via Black-Box Access
von: Pedapati, Tejaswini, et al.
Veröffentlicht: (2024) -
When in Doubt, Cascade: Towards Building Efficient and Capable Guardrails
von: Nagireddy, Manish, et al.
Veröffentlicht: (2024) -
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
von: Jiang, Mingjian, et al.
Veröffentlicht: (2024) -
Extending Beacon to Hindi: Cultural Adaptation Drives Cross-Lingual Sycophancy
von: Sattigeri, Sarthak
Veröffentlicht: (2026)