MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Bakman, Yavuz Faruk, Yaldiz, Duygu Nur, Buyukates, Baturalp, Tao, Chenyang, Dimitriadis, Dimitrios, Avestimehr, Salman |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
by: Yaldiz, Duygu Nur, et al.
Published: (2024)
by: Yaldiz, Duygu Nur, et al.
Published: (2024)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
by: Bakman, Yavuz, et al.
Published: (2025)
by: Bakman, Yavuz, et al.
Published: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
by: Ziashahabi, Amir, et al.
Published: (2025)
by: Ziashahabi, Amir, et al.
Published: (2025)
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
by: Mushtaq, Erum, et al.
Published: (2024)
by: Mushtaq, Erum, et al.
Published: (2024)
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
by: Bakman, Yavuz, et al.
Published: (2026)
by: Bakman, Yavuz, et al.
Published: (2026)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
by: Bakman, Yavuz, et al.
Published: (2025)
by: Bakman, Yavuz, et al.
Published: (2025)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
by: Oguz, Metehan, et al.
Published: (2025)
by: Oguz, Metehan, et al.
Published: (2025)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
by: Wang, Nien-Shao, et al.
Published: (2025)
by: Wang, Nien-Shao, et al.
Published: (2025)
Leveraging Uncertainty Estimation for Efficient LLM Routing
by: Zhang, Tuo, et al.
Published: (2025)
by: Zhang, Tuo, et al.
Published: (2025)
Maverick-Aware Shapley Valuation for Client Selection in Federated Learning
by: Yang, Mengwei, et al.
Published: (2024)
by: Yang, Mengwei, et al.
Published: (2024)
HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
by: Mushtaq, Erum, et al.
Published: (2025)
by: Mushtaq, Erum, et al.
Published: (2025)
FedGrAINS: Personalized SubGraph Federated Learning with Adaptive Neighbor Sampling
by: Ceyani, Emir, et al.
Published: (2025)
by: Ceyani, Emir, et al.
Published: (2025)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
by: Yaldiz, Duygu Nur, et al.
Published: (2025)
by: Yaldiz, Duygu Nur, et al.
Published: (2025)
Balancing Information Accuracy and Response Timeliness in Networked LLMs
by: Turkmen, Yigit, et al.
Published: (2025)
by: Turkmen, Yigit, et al.
Published: (2025)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
by: Oğuz, Metehan, et al.
Published: (2024)
by: Oğuz, Metehan, et al.
Published: (2024)
Renormalization Group flow, Optimal Transport and Diffusion-based Generative Model
by: Sheshmani, Artan, et al.
Published: (2024)
by: Sheshmani, Artan, et al.
Published: (2024)
Don't Always Pick the Highest-Performing Model: An Information Theoretic View of LLM Ensemble Selection
by: Turkmen, Yigit, et al.
Published: (2026)
by: Turkmen, Yigit, et al.
Published: (2026)
Balancing Classification and Calibration Performance in Decision-Making LLMs via Calibration Aware Reinforcement Learning
by: Yaldiz, Duygu Nur, et al.
Published: (2026)
by: Yaldiz, Duygu Nur, et al.
Published: (2026)
Revisiting OPRO: The Limitations of Small-Scale LLMs as Optimizers
by: Zhang, Tuo, et al.
Published: (2024)
by: Zhang, Tuo, et al.
Published: (2024)
GPT-FL: Generative Pre-trained Model-Assisted Federated Learning
by: Zhang, Tuo, et al.
Published: (2023)
by: Zhang, Tuo, et al.
Published: (2023)
Beyond the Score: Uncertainty-Calibrated LLMs for Automated Essay Assessment
by: Karim, Ahmed, et al.
Published: (2025)
by: Karim, Ahmed, et al.
Published: (2025)
The Alignment Tax: Response Homogenization in Aligned LLMs and Its Implications for Uncertainty Estimation
by: Liu, Mingyi
Published: (2026)
by: Liu, Mingyi
Published: (2026)
Kick Bad Guys Out! Conditionally Activated Anomaly Detection in Federated Learning with Zero-Knowledge Proof Verification
by: Han, Shanshan, et al.
Published: (2023)
by: Han, Shanshan, et al.
Published: (2023)
Privacy-preserving Prompt Personalization in Federated Learning for Multimodal Large Language Models
by: Hou, Sizai, et al.
Published: (2025)
by: Hou, Sizai, et al.
Published: (2025)
MEMOIR: Lifelong Model Editing with Minimal Overwrite and Informed Retention for LLMs
by: Wang, Ke, et al.
Published: (2025)
by: Wang, Ke, et al.
Published: (2025)
ALIEN: Aligned Entropy Head for Improving Uncertainty Estimation of LLMs
by: Zabolotnyi, Artem, et al.
Published: (2025)
by: Zabolotnyi, Artem, et al.
Published: (2025)
Estimation of Concept Explanations Should be Uncertainty Aware
by: Piratla, Vihari, et al.
Published: (2023)
by: Piratla, Vihari, et al.
Published: (2023)
CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought
by: Zhang, Boxuan, et al.
Published: (2025)
by: Zhang, Boxuan, et al.
Published: (2025)
Score Before You Speak: Improving Persona Consistency in Dialogue Generation using Response Quality Scores
by: Saggar, Arpita, et al.
Published: (2025)
by: Saggar, Arpita, et al.
Published: (2025)
Multicalibration for Confidence Scoring in LLMs
by: Detommaso, Gianluca, et al.
Published: (2024)
by: Detommaso, Gianluca, et al.
Published: (2024)
Uncertainty-Aware Mean Opinion Score Prediction
by: Wang, Hui, et al.
Published: (2024)
by: Wang, Hui, et al.
Published: (2024)
Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning
by: Ozfatura, Emre, et al.
Published: (2024)
by: Ozfatura, Emre, et al.
Published: (2024)
GEM: A Scale-Aware and Distribution-Sensitive Sparse Fine-Tuning Framework for Effective Downstream Adaptation
by: Kang, Sungmin, et al.
Published: (2025)
by: Kang, Sungmin, et al.
Published: (2025)
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
by: Wang, Yinong Oliver, et al.
Published: (2025)
by: Wang, Yinong Oliver, et al.
Published: (2025)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
by: Tao, Linwei, et al.
Published: (2025)
by: Tao, Linwei, et al.
Published: (2025)
Clustering and Median Aggregation Improve Differentially Private Inference
by: Amin, Kareem, et al.
Published: (2025)
by: Amin, Kareem, et al.
Published: (2025)
Leveraging AI Graders for Missing Score Imputation to Achieve Accurate Ability Estimation in Constructed-Response Tests
by: Uto, Masaki, et al.
Published: (2025)
by: Uto, Masaki, et al.
Published: (2025)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
by: Liu, Liangxin, et al.
Published: (2024)
by: Liu, Liangxin, et al.
Published: (2024)
Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
by: Liu, Linyu, et al.
Published: (2024)
by: Liu, Linyu, et al.
Published: (2024)
Similar Items
-
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
by: Yaldiz, Duygu Nur, et al.
Published: (2024) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
by: Kang, Sungmin, et al.
Published: (2025) -
Reconsidering LLM Uncertainty Estimation Methods in the Wild
by: Bakman, Yavuz, et al.
Published: (2025) -
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
by: Ziashahabi, Amir, et al.
Published: (2025) -
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
by: Mushtaq, Erum, et al.
Published: (2024)