Uncertainty Quantification for LLM Function-Calling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ye, Zihuiwen, Aichberger, Lukas, Kirchhof, Michael, Williamson, Sinead, Zappella, Luca, Gal, Yarin, Blaas, Arno, Golinski, Adam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revisiting Uncertainty Quantification Evaluation in Language Models: Spurious Interactions with Response Length Bias Results
von: Santilli, Andrea, et al.
Veröffentlicht: (2025)
von: Santilli, Andrea, et al.
Veröffentlicht: (2025)
SelfReflect: Can LLMs Communicate Their Internal Answer Distribution?
von: Kirchhof, Michael, et al.
Veröffentlicht: (2025)
von: Kirchhof, Michael, et al.
Veröffentlicht: (2025)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Step-wise Verification with Generative Reward Models
von: Ye, Zihuiwen, et al.
Veröffentlicht: (2025)
von: Ye, Zihuiwen, et al.
Veröffentlicht: (2025)
BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design
von: Choudhury, Deepro, et al.
Veröffentlicht: (2025)
von: Choudhury, Deepro, et al.
Veröffentlicht: (2025)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
von: Nikitin, Alexander, et al.
Veröffentlicht: (2024)
von: Nikitin, Alexander, et al.
Veröffentlicht: (2024)
LinEAS: End-to-end Learning of Activation Steering with a Distributional Loss
von: Rodriguez, Pau, et al.
Veröffentlicht: (2025)
von: Rodriguez, Pau, et al.
Veröffentlicht: (2025)
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
von: Kirchhof, Michael, et al.
Veröffentlicht: (2025)
von: Kirchhof, Michael, et al.
Veröffentlicht: (2025)
Considerations for Distribution Shift Robustness of Diagnostic Models in Healthcare
von: Blaas, Arno, et al.
Veröffentlicht: (2024)
von: Blaas, Arno, et al.
Veröffentlicht: (2024)
Unlocking the Working Memory of Large Language Models for Latent Reasoning
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
Interpreting CLIP: Insights on the Robustness to ImageNet Distribution Shifts
von: Crabbé, Jonathan, et al.
Veröffentlicht: (2023)
von: Crabbé, Jonathan, et al.
Veröffentlicht: (2023)
Controlling Language and Diffusion Models by Transporting Activations
von: Rodriguez, Pau, et al.
Veröffentlicht: (2024)
von: Rodriguez, Pau, et al.
Veröffentlicht: (2024)
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025)
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025)
Annotations Mitigate Post-Training Mode Collapse
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2026)
von: Springer, Jacob Mitchell, et al.
Veröffentlicht: (2026)
Sparse Autoencoders are Capable LLM Jailbreak Mitigators
von: Assogba, Yannick, et al.
Veröffentlicht: (2026)
von: Assogba, Yannick, et al.
Veröffentlicht: (2026)
Language Models Change Facts Based on the Way You Talk
von: Kearney, Matthew, et al.
Veröffentlicht: (2025)
von: Kearney, Matthew, et al.
Veröffentlicht: (2025)
An LLM Compiler for Parallel Function Calling
von: Kim, Sehoon, et al.
Veröffentlicht: (2023)
von: Kim, Sehoon, et al.
Veröffentlicht: (2023)
Do Multilingual LLMs Think In English?
von: Schut, Lisa, et al.
Veröffentlicht: (2025)
von: Schut, Lisa, et al.
Veröffentlicht: (2025)
In-Context Learning Learns Label Relationships but Is Not Conventional Learning
von: Kossen, Jannik, et al.
Veröffentlicht: (2023)
von: Kossen, Jannik, et al.
Veröffentlicht: (2023)
Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification
von: Bouchard, Dylan, et al.
Veröffentlicht: (2026)
von: Bouchard, Dylan, et al.
Veröffentlicht: (2026)
Posterior Uncertainty Quantification in Neural Networks using Data Augmentation
von: Wu, Luhuan, et al.
Veröffentlicht: (2024)
von: Wu, Luhuan, et al.
Veröffentlicht: (2024)
Asynchronous LLM Function Calling
von: Gim, In, et al.
Veröffentlicht: (2024)
von: Gim, In, et al.
Veröffentlicht: (2024)
Uncertainty Quantification and Decomposition for LLM-based Recommendation
von: Kweon, Wonbin, et al.
Veröffentlicht: (2025)
von: Kweon, Wonbin, et al.
Veröffentlicht: (2025)
Benchmarking LLMs via Uncertainty Quantification
von: Ye, Fanghua, et al.
Veröffentlicht: (2024)
von: Ye, Fanghua, et al.
Veröffentlicht: (2024)
Uncertainties of Latent Representations in Computer Vision
von: Kirchhof, Michael
Veröffentlicht: (2024)
von: Kirchhof, Michael
Veröffentlicht: (2024)
Deep Bayesian Active Learning for Preference Modeling in Large Language Models
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
Improving Reward Models with Synthetic Critiques
von: Ye, Zihuiwen, et al.
Veröffentlicht: (2024)
von: Ye, Zihuiwen, et al.
Veröffentlicht: (2024)
Estimating Semantic Alphabet Size for LLM Uncertainty Quantification
von: McCabe, Lucas H., et al.
Veröffentlicht: (2025)
von: McCabe, Lucas H., et al.
Veröffentlicht: (2025)
Fairness Dynamics During Training
von: Patel, Krishna, et al.
Veröffentlicht: (2025)
von: Patel, Krishna, et al.
Veröffentlicht: (2025)
CallNavi, A Challenge and Empirical Study on LLM Function Calling and Routing
von: Song, Yewei, et al.
Veröffentlicht: (2025)
von: Song, Yewei, et al.
Veröffentlicht: (2025)
From Calibration to Collaboration: LLM Uncertainty Quantification Should Be More Human-Centered
von: Devic, Siddartha, et al.
Veröffentlicht: (2025)
von: Devic, Siddartha, et al.
Veröffentlicht: (2025)
SimpleTool: Parallel Decoding for Real-Time LLM Function Calling
von: Shi, Xiaoxin, et al.
Veröffentlicht: (2026)
von: Shi, Xiaoxin, et al.
Veröffentlicht: (2026)
CoRet: Improved Retriever for Code Editing
von: Fehr, Fabio, et al.
Veröffentlicht: (2025)
von: Fehr, Fabio, et al.
Veröffentlicht: (2025)
Agentic Uncertainty Quantification
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
Fine-Tuning Large Language Models to Appropriately Abstain with Semantic Entropy
von: Tjandra, Benedict Aaron, et al.
Veröffentlicht: (2024)
von: Tjandra, Benedict Aaron, et al.
Veröffentlicht: (2024)
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025)
Evaluating Uncertainty Quantification Methods in Argumentative Large Language Models
von: Zhou, Kevin, et al.
Veröffentlicht: (2025)
von: Zhou, Kevin, et al.
Veröffentlicht: (2025)
DiscoUQ: Structured Disagreement Analysis for Uncertainty Quantification in LLM Agent Ensembles
von: Jiang, Bo
Veröffentlicht: (2026)
von: Jiang, Bo
Veröffentlicht: (2026)
Beyond Semantic Entropy: Boosting LLM Uncertainty Quantification with Pairwise Semantic Similarity
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Attention Heads: Efficient Unsupervised Uncertainty Quantification for LLMs
von: Vazhentsev, Artem, et al.
Veröffentlicht: (2025)
von: Vazhentsev, Artem, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Revisiting Uncertainty Quantification Evaluation in Language Models: Spurious Interactions with Response Length Bias Results
von: Santilli, Andrea, et al.
Veröffentlicht: (2025) -
SelfReflect: Can LLMs Communicate Their Internal Answer Distribution?
von: Kirchhof, Michael, et al.
Veröffentlicht: (2025) -
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
von: Nakkiran, Preetum, et al.
Veröffentlicht: (2025) -
Uncertainty-Aware Step-wise Verification with Generative Reward Models
von: Ye, Zihuiwen, et al.
Veröffentlicht: (2025) -
BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design
von: Choudhury, Deepro, et al.
Veröffentlicht: (2025)