Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yaldiz, Duygu Nur, Bakman, Yavuz Faruk, Buyukates, Baturalp, Tao, Chenyang, Ramakrishna, Anil, Dimitriadis, Dimitrios, Zhao, Jieyu, Avestimehr, Salman |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
par: Bakman, Yavuz Faruk, et autres
Publié: (2024)
par: Bakman, Yavuz Faruk, et autres
Publié: (2024)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
par: Kang, Sungmin, et autres
Publié: (2025)
par: Kang, Sungmin, et autres
Publié: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
par: Bakman, Yavuz, et autres
Publié: (2025)
par: Bakman, Yavuz, et autres
Publié: (2025)
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
par: Mushtaq, Erum, et autres
Publié: (2024)
par: Mushtaq, Erum, et autres
Publié: (2024)
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
par: Bakman, Yavuz, et autres
Publié: (2026)
par: Bakman, Yavuz, et autres
Publié: (2026)
HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
par: Mushtaq, Erum, et autres
Publié: (2025)
par: Mushtaq, Erum, et autres
Publié: (2025)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
par: Oguz, Metehan, et autres
Publié: (2025)
par: Oguz, Metehan, et autres
Publié: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
par: Ziashahabi, Amir, et autres
Publié: (2025)
par: Ziashahabi, Amir, et autres
Publié: (2025)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
par: Bakman, Yavuz, et autres
Publié: (2025)
par: Bakman, Yavuz, et autres
Publié: (2025)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
par: Wang, Nien-Shao, et autres
Publié: (2025)
par: Wang, Nien-Shao, et autres
Publié: (2025)
Leveraging Uncertainty Estimation for Efficient LLM Routing
par: Zhang, Tuo, et autres
Publié: (2025)
par: Zhang, Tuo, et autres
Publié: (2025)
Maverick-Aware Shapley Valuation for Client Selection in Federated Learning
par: Yang, Mengwei, et autres
Publié: (2024)
par: Yang, Mengwei, et autres
Publié: (2024)
FedGrAINS: Personalized SubGraph Federated Learning with Adaptive Neighbor Sampling
par: Ceyani, Emir, et autres
Publié: (2025)
par: Ceyani, Emir, et autres
Publié: (2025)
Renormalization Group flow, Optimal Transport and Diffusion-based Generative Model
par: Sheshmani, Artan, et autres
Publié: (2024)
par: Sheshmani, Artan, et autres
Publié: (2024)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
par: Oğuz, Metehan, et autres
Publié: (2024)
par: Oğuz, Metehan, et autres
Publié: (2024)
Balancing Information Accuracy and Response Timeliness in Networked LLMs
par: Turkmen, Yigit, et autres
Publié: (2025)
par: Turkmen, Yigit, et autres
Publié: (2025)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
par: Yaldiz, Duygu Nur, et autres
Publié: (2025)
par: Yaldiz, Duygu Nur, et autres
Publié: (2025)
Kick Bad Guys Out! Conditionally Activated Anomaly Detection in Federated Learning with Zero-Knowledge Proof Verification
par: Han, Shanshan, et autres
Publié: (2023)
par: Han, Shanshan, et autres
Publié: (2023)
Privacy-preserving Prompt Personalization in Federated Learning for Multimodal Large Language Models
par: Hou, Sizai, et autres
Publié: (2025)
par: Hou, Sizai, et autres
Publié: (2025)
Don't Always Pick the Highest-Performing Model: An Information Theoretic View of LLM Ensemble Selection
par: Turkmen, Yigit, et autres
Publié: (2026)
par: Turkmen, Yigit, et autres
Publié: (2026)
Balancing Classification and Calibration Performance in Decision-Making LLMs via Calibration Aware Reinforcement Learning
par: Yaldiz, Duygu Nur, et autres
Publié: (2026)
par: Yaldiz, Duygu Nur, et autres
Publié: (2026)
Revisiting OPRO: The Limitations of Small-Scale LLMs as Optimizers
par: Zhang, Tuo, et autres
Publié: (2024)
par: Zhang, Tuo, et autres
Publié: (2024)
FedSecurity: Benchmarking Attacks and Defenses in Federated Learning and Federated LLMs
par: Han, Shanshan, et autres
Publié: (2023)
par: Han, Shanshan, et autres
Publié: (2023)
GPT-FL: Generative Pre-trained Model-Assisted Federated Learning
par: Zhang, Tuo, et autres
Publié: (2023)
par: Zhang, Tuo, et autres
Publié: (2023)
Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning
par: Ozfatura, Emre, et autres
Publié: (2024)
par: Ozfatura, Emre, et autres
Publié: (2024)
Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation
par: Aravindan, Ashwath Vaithinathan, et autres
Publié: (2025)
par: Aravindan, Ashwath Vaithinathan, et autres
Publié: (2025)
Backdoor Defense in Diffusion Models via Spatial Attention Unlearning
par: Jha, Abha, et autres
Publié: (2025)
par: Jha, Abha, et autres
Publié: (2025)
Reseña de "Class Acts: Service and Inequality in Luxury Hotels" de Rachel Sherman
par: Duygu Salman
Publié: (2012)
par: Duygu Salman
Publié: (2012)
Rethinking of Cities, Culture and Tourism within a Creative Perspective
par: Duygu Salman
Publié: (2010)
par: Duygu Salman
Publié: (2010)
Spatiotemporal variability and trends of droughts in the Mediterranean coastal region of Türkiye
par: Erdal Kesgin, et autres
Publié: (2024)
par: Erdal Kesgin, et autres
Publié: (2024)
An objective criterion for evaluating new physical theories
par: Bakman, Yefim
Publié: (2025)
par: Bakman, Yefim
Publié: (2025)
POR ENTRE AS TRAMAS FAMILIARES: Avós JUDEUS E SEUS NETOS POR ADOÇÃO
par: Gizele Bakman
Publié: (2023)
par: Gizele Bakman
Publié: (2023)
Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences
par: Han, Shanshan, et autres
Publié: (2025)
par: Han, Shanshan, et autres
Publié: (2025)
Understanding Communication Backends in Cross-Silo Federated Learning
par: Ziashahabi, Amir, et autres
Publié: (2026)
par: Ziashahabi, Amir, et autres
Publié: (2026)
Differentially Private Federated Learning without Noise Addition: When is it Possible?
par: Zhang, Jiang, et autres
Publié: (2024)
par: Zhang, Jiang, et autres
Publié: (2024)
ATP: Enabling Fast LLM Serving via Attention on Top Principal Keys
par: Niu, Yue, et autres
Publié: (2024)
par: Niu, Yue, et autres
Publié: (2024)
Metrics for Assessing Inclusivity and Empowerment of People for Supporting the Design of Inclusive Product Lifecycles
par: Yaldiz, Naz, et autres
Publié: (2024)
par: Yaldiz, Naz, et autres
Publié: (2024)
Statistical Inference for Score Decompositions
par: Dimitriadis, Timo, et autres
Publié: (2026)
par: Dimitriadis, Timo, et autres
Publié: (2026)
Green Fiscal Stance and Climate Pressure in Advanced Europe: Evidence From a Multidimensional Climate Index
par: Elif Duygu Kömürcüoğlu, et autres
Publié: (2026)
par: Elif Duygu Kömürcüoğlu, et autres
Publié: (2026)
ModalityMirror: Improving Audio Classification in Modality Heterogeneity Federated Learning with Multimodal Distillation
par: Feng, Tiantian, et autres
Publié: (2024)
par: Feng, Tiantian, et autres
Publié: (2024)
Documents similaires
-
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
par: Bakman, Yavuz Faruk, et autres
Publié: (2024) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
par: Kang, Sungmin, et autres
Publié: (2025) -
Reconsidering LLM Uncertainty Estimation Methods in the Wild
par: Bakman, Yavuz, et autres
Publié: (2025) -
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
par: Mushtaq, Erum, et autres
Publié: (2024) -
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
par: Bakman, Yavuz, et autres
Publié: (2026)