MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bakman, Yavuz Faruk, Yaldiz, Duygu Nur, Buyukates, Baturalp, Tao, Chenyang, Dimitriadis, Dimitrios, Avestimehr, Salman |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2024)
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2024)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
von: Kang, Sungmin, et al.
Veröffentlicht: (2025)
von: Kang, Sungmin, et al.
Veröffentlicht: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
von: Ziashahabi, Amir, et al.
Veröffentlicht: (2025)
von: Ziashahabi, Amir, et al.
Veröffentlicht: (2025)
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
von: Mushtaq, Erum, et al.
Veröffentlicht: (2024)
von: Mushtaq, Erum, et al.
Veröffentlicht: (2024)
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
von: Bakman, Yavuz, et al.
Veröffentlicht: (2026)
von: Bakman, Yavuz, et al.
Veröffentlicht: (2026)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
von: Oguz, Metehan, et al.
Veröffentlicht: (2025)
von: Oguz, Metehan, et al.
Veröffentlicht: (2025)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
von: Wang, Nien-Shao, et al.
Veröffentlicht: (2025)
von: Wang, Nien-Shao, et al.
Veröffentlicht: (2025)
Leveraging Uncertainty Estimation for Efficient LLM Routing
von: Zhang, Tuo, et al.
Veröffentlicht: (2025)
von: Zhang, Tuo, et al.
Veröffentlicht: (2025)
Maverick-Aware Shapley Valuation for Client Selection in Federated Learning
von: Yang, Mengwei, et al.
Veröffentlicht: (2024)
von: Yang, Mengwei, et al.
Veröffentlicht: (2024)
HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
von: Mushtaq, Erum, et al.
Veröffentlicht: (2025)
von: Mushtaq, Erum, et al.
Veröffentlicht: (2025)
FedGrAINS: Personalized SubGraph Federated Learning with Adaptive Neighbor Sampling
von: Ceyani, Emir, et al.
Veröffentlicht: (2025)
von: Ceyani, Emir, et al.
Veröffentlicht: (2025)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2025)
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2025)
Balancing Information Accuracy and Response Timeliness in Networked LLMs
von: Turkmen, Yigit, et al.
Veröffentlicht: (2025)
von: Turkmen, Yigit, et al.
Veröffentlicht: (2025)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
von: Oğuz, Metehan, et al.
Veröffentlicht: (2024)
von: Oğuz, Metehan, et al.
Veröffentlicht: (2024)
Renormalization Group flow, Optimal Transport and Diffusion-based Generative Model
von: Sheshmani, Artan, et al.
Veröffentlicht: (2024)
von: Sheshmani, Artan, et al.
Veröffentlicht: (2024)
Don't Always Pick the Highest-Performing Model: An Information Theoretic View of LLM Ensemble Selection
von: Turkmen, Yigit, et al.
Veröffentlicht: (2026)
von: Turkmen, Yigit, et al.
Veröffentlicht: (2026)
Balancing Classification and Calibration Performance in Decision-Making LLMs via Calibration Aware Reinforcement Learning
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2026)
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2026)
Revisiting OPRO: The Limitations of Small-Scale LLMs as Optimizers
von: Zhang, Tuo, et al.
Veröffentlicht: (2024)
von: Zhang, Tuo, et al.
Veröffentlicht: (2024)
GPT-FL: Generative Pre-trained Model-Assisted Federated Learning
von: Zhang, Tuo, et al.
Veröffentlicht: (2023)
von: Zhang, Tuo, et al.
Veröffentlicht: (2023)
Beyond the Score: Uncertainty-Calibrated LLMs for Automated Essay Assessment
von: Karim, Ahmed, et al.
Veröffentlicht: (2025)
von: Karim, Ahmed, et al.
Veröffentlicht: (2025)
The Alignment Tax: Response Homogenization in Aligned LLMs and Its Implications for Uncertainty Estimation
von: Liu, Mingyi
Veröffentlicht: (2026)
von: Liu, Mingyi
Veröffentlicht: (2026)
Kick Bad Guys Out! Conditionally Activated Anomaly Detection in Federated Learning with Zero-Knowledge Proof Verification
von: Han, Shanshan, et al.
Veröffentlicht: (2023)
von: Han, Shanshan, et al.
Veröffentlicht: (2023)
Privacy-preserving Prompt Personalization in Federated Learning for Multimodal Large Language Models
von: Hou, Sizai, et al.
Veröffentlicht: (2025)
von: Hou, Sizai, et al.
Veröffentlicht: (2025)
MEMOIR: Lifelong Model Editing with Minimal Overwrite and Informed Retention for LLMs
von: Wang, Ke, et al.
Veröffentlicht: (2025)
von: Wang, Ke, et al.
Veröffentlicht: (2025)
ALIEN: Aligned Entropy Head for Improving Uncertainty Estimation of LLMs
von: Zabolotnyi, Artem, et al.
Veröffentlicht: (2025)
von: Zabolotnyi, Artem, et al.
Veröffentlicht: (2025)
Estimation of Concept Explanations Should be Uncertainty Aware
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)
CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought
von: Zhang, Boxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Boxuan, et al.
Veröffentlicht: (2025)
Score Before You Speak: Improving Persona Consistency in Dialogue Generation using Response Quality Scores
von: Saggar, Arpita, et al.
Veröffentlicht: (2025)
von: Saggar, Arpita, et al.
Veröffentlicht: (2025)
Multicalibration for Confidence Scoring in LLMs
von: Detommaso, Gianluca, et al.
Veröffentlicht: (2024)
von: Detommaso, Gianluca, et al.
Veröffentlicht: (2024)
Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning
von: Ozfatura, Emre, et al.
Veröffentlicht: (2024)
von: Ozfatura, Emre, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Mean Opinion Score Prediction
von: Wang, Hui, et al.
Veröffentlicht: (2024)
von: Wang, Hui, et al.
Veröffentlicht: (2024)
GEM: A Scale-Aware and Distribution-Sensitive Sparse Fine-Tuning Framework for Effective Downstream Adaptation
von: Kang, Sungmin, et al.
Veröffentlicht: (2025)
von: Kang, Sungmin, et al.
Veröffentlicht: (2025)
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025)
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
Clustering and Median Aggregation Improve Differentially Private Inference
von: Amin, Kareem, et al.
Veröffentlicht: (2025)
von: Amin, Kareem, et al.
Veröffentlicht: (2025)
Leveraging AI Graders for Missing Score Imputation to Achieve Accurate Ability Estimation in Constructed-Response Tests
von: Uto, Masaki, et al.
Veröffentlicht: (2025)
von: Uto, Masaki, et al.
Veröffentlicht: (2025)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
von: Liu, Liangxin, et al.
Veröffentlicht: (2024)
von: Liu, Liangxin, et al.
Veröffentlicht: (2024)
Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
von: Liu, Linyu, et al.
Veröffentlicht: (2024)
von: Liu, Linyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
von: Yaldiz, Duygu Nur, et al.
Veröffentlicht: (2024) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
von: Kang, Sungmin, et al.
Veröffentlicht: (2025) -
Reconsidering LLM Uncertainty Estimation Methods in the Wild
von: Bakman, Yavuz, et al.
Veröffentlicht: (2025) -
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
von: Ziashahabi, Amir, et al.
Veröffentlicht: (2025) -
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
von: Mushtaq, Erum, et al.
Veröffentlicht: (2024)