Gespeichert in:
| Hauptverfasser: | Cao, Steven, Valiant, Gregory, Liang, Percy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2511.11966 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Eliciting Language Model Behaviors with Investigator Agents
von: Li, Xiang Lisa, et al.
Veröffentlicht: (2025)
von: Li, Xiang Lisa, et al.
Veröffentlicht: (2025)
Enhanced and Efficient Reasoning in Large Learning Models
von: Valiant, Leslie G.
Veröffentlicht: (2026)
von: Valiant, Leslie G.
Veröffentlicht: (2026)
Asymmetric Advantage Modulation Calibrates Entropy Dynamics in RLVR
von: Gu, Hengrui, et al.
Veröffentlicht: (2026)
von: Gu, Hengrui, et al.
Veröffentlicht: (2026)
MixCE: Training Autoregressive Language Models by Mixing Forward and Reverse Cross-Entropies
von: Zhang, Shiyue, et al.
Veröffentlicht: (2023)
von: Zhang, Shiyue, et al.
Veröffentlicht: (2023)
Reliable and Efficient Amortized Model-based Evaluation
von: Truong, Sang, et al.
Veröffentlicht: (2025)
von: Truong, Sang, et al.
Veröffentlicht: (2025)
Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
von: Dubois, Yann, et al.
Veröffentlicht: (2024)
von: Dubois, Yann, et al.
Veröffentlicht: (2024)
Extracting books from production language models
von: Ahmed, Ahmed, et al.
Veröffentlicht: (2026)
von: Ahmed, Ahmed, et al.
Veröffentlicht: (2026)
Calibrating Language Models with Adaptive Temperature Scaling
von: Xie, Johnathan, et al.
Veröffentlicht: (2024)
von: Xie, Johnathan, et al.
Veröffentlicht: (2024)
Prismatic VLMs: Investigating the Design Space of Visually-Conditioned Language Models
von: Karamcheti, Siddharth, et al.
Veröffentlicht: (2024)
von: Karamcheti, Siddharth, et al.
Veröffentlicht: (2024)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
Beware of Calibration Data for Pruning Large Language Models
von: Ji, Yixin, et al.
Veröffentlicht: (2024)
von: Ji, Yixin, et al.
Veröffentlicht: (2024)
On Calibration of Large Language Models: From Response To Capability
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026)
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026)
Calibrating Large Language Models Using Their Generations Only
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
Uncertainty in Language Models: Assessment through Rank-Calibration
von: Huang, Xinmeng, et al.
Veröffentlicht: (2024)
von: Huang, Xinmeng, et al.
Veröffentlicht: (2024)
The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
von: Cui, Ganqu, et al.
Veröffentlicht: (2025)
von: Cui, Ganqu, et al.
Veröffentlicht: (2025)
UQ: Assessing Language Models on Unsolved Questions
von: Nie, Fan, et al.
Veröffentlicht: (2025)
von: Nie, Fan, et al.
Veröffentlicht: (2025)
Calibrating Long-form Generations from Large Language Models
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
Language Models May Verbatim Complete Text They Were Not Explicitly Trained On
von: Liu, Ken Ziyu, et al.
Veröffentlicht: (2025)
von: Liu, Ken Ziyu, et al.
Veröffentlicht: (2025)
Entropy Aware Reward Guidance for Diffusion Language Model Alignment
von: Tejaswi, Atula, et al.
Veröffentlicht: (2026)
von: Tejaswi, Atula, et al.
Veröffentlicht: (2026)
How Alignment Routes: Localizing, Scaling, and Controlling Policy Circuits in Language Models
von: Frank, Gregory N.
Veröffentlicht: (2026)
von: Frank, Gregory N.
Veröffentlicht: (2026)
Merino: Entropy-driven Design for Generative Language Models on IoT Devices
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
Calibration vs Decision Making: Revisiting the Reliability Paradox in Unlearned Language Models
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2026)
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2026)
Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning
von: Krishnan, Ranganath, et al.
Veröffentlicht: (2024)
von: Krishnan, Ranganath, et al.
Veröffentlicht: (2024)
The Reliability Paradox: Exploring How Shortcut Learning Undermines Language Model Calibration
von: Bihani, Geetanjali, et al.
Veröffentlicht: (2024)
von: Bihani, Geetanjali, et al.
Veröffentlicht: (2024)
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation
von: Wang, Ziyu, et al.
Veröffentlicht: (2024)
von: Wang, Ziyu, et al.
Veröffentlicht: (2024)
More Bang for the Buck: Process Reward Modeling with Entropy-Driven Uncertainty
von: Cao, Lang, et al.
Veröffentlicht: (2025)
von: Cao, Lang, et al.
Veröffentlicht: (2025)
Stacking Small Language Models for Generalizability
von: Liang, Laurence
Veröffentlicht: (2024)
von: Liang, Laurence
Veröffentlicht: (2024)
MoxE: Mixture of xLSTM Experts with Entropy-Aware Routing for Efficient Language Modeling
von: Thiombiano, Abdoul Majid O., et al.
Veröffentlicht: (2025)
von: Thiombiano, Abdoul Majid O., et al.
Veröffentlicht: (2025)
Data-Driven Calibration of Prediction Sets in Large Vision-Language Models Based on Inductive Conformal Prediction
von: Ye, Yuanchang, et al.
Veröffentlicht: (2025)
von: Ye, Yuanchang, et al.
Veröffentlicht: (2025)
On the Limitations of Language Targeted Pruning: Investigating the Calibration Language Impact in Multilingual LLM Pruning
von: Kurz, Simon, et al.
Veröffentlicht: (2024)
von: Kurz, Simon, et al.
Veröffentlicht: (2024)
Graph-based Confidence Calibration for Large Language Models
von: Li, Yukun, et al.
Veröffentlicht: (2024)
von: Li, Yukun, et al.
Veröffentlicht: (2024)
E-Sparse: Boosting the Large Language Model Inference through Entropy-based N:M Sparsity
von: Li, Yun, et al.
Veröffentlicht: (2023)
von: Li, Yun, et al.
Veröffentlicht: (2023)
Large Language and Reasoning Models are Shallow Disjunctive Reasoners
von: Khalid, Irtaza, et al.
Veröffentlicht: (2025)
von: Khalid, Irtaza, et al.
Veröffentlicht: (2025)
Towards Reasoning Ability of Small Language Models
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2025)
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2025)
FVEL: Interactive Formal Verification Environment with Large Language Models via Theorem Proving
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
Rethinking Entropy Regularization in Large Reasoning Models
von: Jiang, Yuxian, et al.
Veröffentlicht: (2025)
von: Jiang, Yuxian, et al.
Veröffentlicht: (2025)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
Fill In The Gaps: Model Calibration and Generalization with Synthetic Data
von: Ba, Yang, et al.
Veröffentlicht: (2024)
von: Ba, Yang, et al.
Veröffentlicht: (2024)
Delta Knowledge Distillation for Large Language Models
von: Cao, Yihan, et al.
Veröffentlicht: (2025)
von: Cao, Yihan, et al.
Veröffentlicht: (2025)
Batch Calibration: Rethinking Calibration for In-Context Learning and Prompt Engineering
von: Zhou, Han, et al.
Veröffentlicht: (2023)
von: Zhou, Han, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Eliciting Language Model Behaviors with Investigator Agents
von: Li, Xiang Lisa, et al.
Veröffentlicht: (2025) -
Enhanced and Efficient Reasoning in Large Learning Models
von: Valiant, Leslie G.
Veröffentlicht: (2026) -
Asymmetric Advantage Modulation Calibrates Entropy Dynamics in RLVR
von: Gu, Hengrui, et al.
Veröffentlicht: (2026) -
MixCE: Training Autoregressive Language Models by Mixing Forward and Reverse Cross-Entropies
von: Zhang, Shiyue, et al.
Veröffentlicht: (2023) -
Reliable and Efficient Amortized Model-based Evaluation
von: Truong, Sang, et al.
Veröffentlicht: (2025)