Monte Carlo Temperature: a robust sampling strategy for LLM's uncertainty quantification methods
Fuente:
arXiv
Saved in:
| Main Authors: | Cecere, Nicola, Bacciu, Andrea, Tobías, Ignacio Fernández, Mantrach, Amin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Select, Label, Evaluate: Active Testing in NLP
by: Purificato, Antonio, et al.
Published: (2026)
by: Purificato, Antonio, et al.
Published: (2026)
Cross-Lingual LLM-Judge Transfer via Evaluation Decomposition
by: Sheth, Ivaxi, et al.
Published: (2026)
by: Sheth, Ivaxi, et al.
Published: (2026)
Multilingual Self-Taught Faithfulness Evaluators
by: Alfano, Carlo, et al.
Published: (2025)
by: Alfano, Carlo, et al.
Published: (2025)
RRAML: Reinforced Retrieval Augmented Machine Learning
by: Bacciu, Andrea, et al.
Published: (2023)
by: Bacciu, Andrea, et al.
Published: (2023)
Semantic uncertainty in advanced decoding methods for LLM generation
by: Foodeei, Darius, et al.
Published: (2025)
by: Foodeei, Darius, et al.
Published: (2025)
Efficient semantic uncertainty quantification in language models via diversity-steered sampling
by: Park, Ji Won, et al.
Published: (2025)
by: Park, Ji Won, et al.
Published: (2025)
Faster LLM Inference via Sequential Monte Carlo
by: Emara, Yahya, et al.
Published: (2026)
by: Emara, Yahya, et al.
Published: (2026)
Handling Ontology Gaps in Semantic Parsing
by: Bacciu, Andrea, et al.
Published: (2024)
by: Bacciu, Andrea, et al.
Published: (2024)
Self-generated Replay Memories for Continual Neural Machine Translation
by: Resta, Michele, et al.
Published: (2024)
by: Resta, Michele, et al.
Published: (2024)
LLM-MC-Affect: LLM-Based Monte Carlo Modeling of Affective Trajectories and Latent Ambiguity for Interpersonal Dynamic Insight
by: Lin, Yu-Zheng, et al.
Published: (2026)
by: Lin, Yu-Zheng, et al.
Published: (2026)
An overview of model uncertainty and variability in LLM-based sentiment analysis. Challenges, mitigation strategies and the role of explainability
by: Herrera-Poyatos, David, et al.
Published: (2025)
by: Herrera-Poyatos, David, et al.
Published: (2025)
Directional Concentration Uncertainty: A representational approach to uncertainty quantification for generative models
by: Chattopadhyay, Souradeep, et al.
Published: (2026)
by: Chattopadhyay, Souradeep, et al.
Published: (2026)
Neural machine translation of clinical procedure codes for medical diagnosis and uncertainty quantification
by: Chung, Pei-Hung, et al.
Published: (2024)
by: Chung, Pei-Hung, et al.
Published: (2024)
Process Supervision for Chain-of-Thought Reasoning via Monte Carlo Net Information Gain
by: Royer, Corentin, et al.
Published: (2026)
by: Royer, Corentin, et al.
Published: (2026)
Monte Carlo Sampling for Analyzing In-Context Examples
by: Schoch, Stephanie, et al.
Published: (2025)
by: Schoch, Stephanie, et al.
Published: (2025)
Surrogate-based multilevel Monte Carlo methods for uncertainty quantification in the Grad-Shafranov free boundary problem
by: Elman, Howard, et al.
Published: (2025)
by: Elman, Howard, et al.
Published: (2025)
Examining the robustness of LLM evaluation to the distributional assumptions of benchmarks
by: Ailem, Melissa, et al.
Published: (2024)
by: Ailem, Melissa, et al.
Published: (2024)
AgentSHAP: Interpreting LLM Agent Tool Importance with Monte Carlo Shapley Value Estimation
by: Horovicz, Miriam
Published: (2025)
by: Horovicz, Miriam
Published: (2025)
GPT as a Monte Carlo Language Tree: A Probabilistic Perspective
by: Ning, Kun-Peng, et al.
Published: (2025)
by: Ning, Kun-Peng, et al.
Published: (2025)
The Majority Vote Paradigm Shift: When Popular Meets Optimal
by: Purificato, Antonio, et al.
Published: (2025)
by: Purificato, Antonio, et al.
Published: (2025)
Quasi-Monte Carlo methods for uncertainty quantification of tumor growth modeled by a parametric semi-linear parabolic reaction-diffusion equation
by: Gilbert, Alexander D., et al.
Published: (2025)
by: Gilbert, Alexander D., et al.
Published: (2025)
Cross-lingual robustness of LLM-brain alignment and its computational roots
by: Yang, Ni, et al.
Published: (2026)
by: Yang, Ni, et al.
Published: (2026)
The Necessity of Setting Temperature in LLM-as-a-Judge
by: Li, Lujun, et al.
Published: (2026)
by: Li, Lujun, et al.
Published: (2026)
Interpretable Contrastive Monte Carlo Tree Search Reasoning
by: Gao, Zitian, et al.
Published: (2024)
by: Gao, Zitian, et al.
Published: (2024)
MCTS-RAG: Enhancing Retrieval-Augmented Generation with Monte Carlo Tree Search
by: Hu, Yunhai, et al.
Published: (2025)
by: Hu, Yunhai, et al.
Published: (2025)
Monte Carlo Planning with Large Language Model for Text-Based Game Agents
by: Shi, Zijing, et al.
Published: (2025)
by: Shi, Zijing, et al.
Published: (2025)
GenDec: A robust generative Question-decomposition method for Multi-hop reasoning
by: Wu, Jian, et al.
Published: (2024)
by: Wu, Jian, et al.
Published: (2024)
Comparing human and LLM politeness strategies in free production
by: Zhao, Haoran, et al.
Published: (2025)
by: Zhao, Haoran, et al.
Published: (2025)
TokenSHAP: Interpreting Large Language Models with Monte Carlo Shapley Value Estimation
by: Goldshmidt, Roni, et al.
Published: (2024)
by: Goldshmidt, Roni, et al.
Published: (2024)
A Monte Carlo Language Model Pipeline for Zero-Shot Sociopolitical Event Extraction
by: Cai, Erica, et al.
Published: (2023)
by: Cai, Erica, et al.
Published: (2023)
Predictions from language models for multiple-choice tasks are not robust under variation of scoring methods
by: Tsvilodub, Polina, et al.
Published: (2024)
by: Tsvilodub, Polina, et al.
Published: (2024)
Diffusion Language Model Inference with Monte Carlo Tree Search
by: Huang, Zheng, et al.
Published: (2025)
by: Huang, Zheng, et al.
Published: (2025)
Ensembling Language Models with Sequential Monte Carlo
by: Chan, Robin Shing Moon, et al.
Published: (2026)
by: Chan, Robin Shing Moon, et al.
Published: (2026)
I-MCTS: Enhancing Agentic AutoML via Introspective Monte Carlo Tree Search
by: Liang, Zujie, et al.
Published: (2025)
by: Liang, Zujie, et al.
Published: (2025)
Iterative Hypothesis Generation for Scientific Discovery with Monte Carlo Nash Equilibrium Self-Refining Trees
by: Rabby, Gollam, et al.
Published: (2025)
by: Rabby, Gollam, et al.
Published: (2025)
Enhancing Logical Reasoning in Language Models via Symbolically-Guided Monte Carlo Process Supervision
by: Tan, Xingwei, et al.
Published: (2025)
by: Tan, Xingwei, et al.
Published: (2025)
SAPIENT: Mastering Multi-turn Conversational Recommendation with Strategic Planning and Monte Carlo Tree Search
by: Du, Hanwen, et al.
Published: (2024)
by: Du, Hanwen, et al.
Published: (2024)
Certainty robustness: Evaluating LLM stability under self-challenging prompts
by: Saadat, Mohammadreza, et al.
Published: (2026)
by: Saadat, Mohammadreza, et al.
Published: (2026)
OptiHive: Ensemble Selection for LLM-Based Optimization via Statistical Modeling
by: Bouscary, Maxime, et al.
Published: (2025)
by: Bouscary, Maxime, et al.
Published: (2025)
Bayesian Orchestration of Multi-LLM Agents for Cost-Aware Sequential Decision-Making
by: Amin, Danial
Published: (2026)
by: Amin, Danial
Published: (2026)
Similar Items
-
Select, Label, Evaluate: Active Testing in NLP
by: Purificato, Antonio, et al.
Published: (2026) -
Cross-Lingual LLM-Judge Transfer via Evaluation Decomposition
by: Sheth, Ivaxi, et al.
Published: (2026) -
Multilingual Self-Taught Faithfulness Evaluators
by: Alfano, Carlo, et al.
Published: (2025) -
RRAML: Reinforced Retrieval Augmented Machine Learning
by: Bacciu, Andrea, et al.
Published: (2023) -
Semantic uncertainty in advanced decoding methods for LLM generation
by: Foodeei, Darius, et al.
Published: (2025)