Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Yaldiz, Duygu Nur, Bakman, Yavuz Faruk, Buyukates, Baturalp, Tao, Chenyang, Ramakrishna, Anil, Dimitriadis, Dimitrios, Zhao, Jieyu, Avestimehr, Salman |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
di: Bakman, Yavuz Faruk, et al.
Pubblicazione: (2024)
di: Bakman, Yavuz Faruk, et al.
Pubblicazione: (2024)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
di: Kang, Sungmin, et al.
Pubblicazione: (2025)
di: Kang, Sungmin, et al.
Pubblicazione: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
di: Bakman, Yavuz, et al.
Pubblicazione: (2025)
di: Bakman, Yavuz, et al.
Pubblicazione: (2025)
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
di: Mushtaq, Erum, et al.
Pubblicazione: (2024)
di: Mushtaq, Erum, et al.
Pubblicazione: (2024)
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
di: Bakman, Yavuz, et al.
Pubblicazione: (2026)
di: Bakman, Yavuz, et al.
Pubblicazione: (2026)
HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
di: Mushtaq, Erum, et al.
Pubblicazione: (2025)
di: Mushtaq, Erum, et al.
Pubblicazione: (2025)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
di: Oguz, Metehan, et al.
Pubblicazione: (2025)
di: Oguz, Metehan, et al.
Pubblicazione: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
di: Ziashahabi, Amir, et al.
Pubblicazione: (2025)
di: Ziashahabi, Amir, et al.
Pubblicazione: (2025)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
di: Bakman, Yavuz, et al.
Pubblicazione: (2025)
di: Bakman, Yavuz, et al.
Pubblicazione: (2025)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
di: Wang, Nien-Shao, et al.
Pubblicazione: (2025)
di: Wang, Nien-Shao, et al.
Pubblicazione: (2025)
Leveraging Uncertainty Estimation for Efficient LLM Routing
di: Zhang, Tuo, et al.
Pubblicazione: (2025)
di: Zhang, Tuo, et al.
Pubblicazione: (2025)
Maverick-Aware Shapley Valuation for Client Selection in Federated Learning
di: Yang, Mengwei, et al.
Pubblicazione: (2024)
di: Yang, Mengwei, et al.
Pubblicazione: (2024)
FedGrAINS: Personalized SubGraph Federated Learning with Adaptive Neighbor Sampling
di: Ceyani, Emir, et al.
Pubblicazione: (2025)
di: Ceyani, Emir, et al.
Pubblicazione: (2025)
Renormalization Group flow, Optimal Transport and Diffusion-based Generative Model
di: Sheshmani, Artan, et al.
Pubblicazione: (2024)
di: Sheshmani, Artan, et al.
Pubblicazione: (2024)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
di: Oğuz, Metehan, et al.
Pubblicazione: (2024)
di: Oğuz, Metehan, et al.
Pubblicazione: (2024)
Balancing Information Accuracy and Response Timeliness in Networked LLMs
di: Turkmen, Yigit, et al.
Pubblicazione: (2025)
di: Turkmen, Yigit, et al.
Pubblicazione: (2025)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
di: Yaldiz, Duygu Nur, et al.
Pubblicazione: (2025)
di: Yaldiz, Duygu Nur, et al.
Pubblicazione: (2025)
Kick Bad Guys Out! Conditionally Activated Anomaly Detection in Federated Learning with Zero-Knowledge Proof Verification
di: Han, Shanshan, et al.
Pubblicazione: (2023)
di: Han, Shanshan, et al.
Pubblicazione: (2023)
Privacy-preserving Prompt Personalization in Federated Learning for Multimodal Large Language Models
di: Hou, Sizai, et al.
Pubblicazione: (2025)
di: Hou, Sizai, et al.
Pubblicazione: (2025)
Don't Always Pick the Highest-Performing Model: An Information Theoretic View of LLM Ensemble Selection
di: Turkmen, Yigit, et al.
Pubblicazione: (2026)
di: Turkmen, Yigit, et al.
Pubblicazione: (2026)
Balancing Classification and Calibration Performance in Decision-Making LLMs via Calibration Aware Reinforcement Learning
di: Yaldiz, Duygu Nur, et al.
Pubblicazione: (2026)
di: Yaldiz, Duygu Nur, et al.
Pubblicazione: (2026)
Revisiting OPRO: The Limitations of Small-Scale LLMs as Optimizers
di: Zhang, Tuo, et al.
Pubblicazione: (2024)
di: Zhang, Tuo, et al.
Pubblicazione: (2024)
FedSecurity: Benchmarking Attacks and Defenses in Federated Learning and Federated LLMs
di: Han, Shanshan, et al.
Pubblicazione: (2023)
di: Han, Shanshan, et al.
Pubblicazione: (2023)
GPT-FL: Generative Pre-trained Model-Assisted Federated Learning
di: Zhang, Tuo, et al.
Pubblicazione: (2023)
di: Zhang, Tuo, et al.
Pubblicazione: (2023)
Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning
di: Ozfatura, Emre, et al.
Pubblicazione: (2024)
di: Ozfatura, Emre, et al.
Pubblicazione: (2024)
Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation
di: Aravindan, Ashwath Vaithinathan, et al.
Pubblicazione: (2025)
di: Aravindan, Ashwath Vaithinathan, et al.
Pubblicazione: (2025)
Backdoor Defense in Diffusion Models via Spatial Attention Unlearning
di: Jha, Abha, et al.
Pubblicazione: (2025)
di: Jha, Abha, et al.
Pubblicazione: (2025)
Reseña de "Class Acts: Service and Inequality in Luxury Hotels" de Rachel Sherman
di: Duygu Salman
Pubblicazione: (2012)
di: Duygu Salman
Pubblicazione: (2012)
Rethinking of Cities, Culture and Tourism within a Creative Perspective
di: Duygu Salman
Pubblicazione: (2010)
di: Duygu Salman
Pubblicazione: (2010)
Spatiotemporal variability and trends of droughts in the Mediterranean coastal region of Türkiye
di: Erdal Kesgin, et al.
Pubblicazione: (2024)
di: Erdal Kesgin, et al.
Pubblicazione: (2024)
An objective criterion for evaluating new physical theories
di: Bakman, Yefim
Pubblicazione: (2025)
di: Bakman, Yefim
Pubblicazione: (2025)
POR ENTRE AS TRAMAS FAMILIARES: Avós JUDEUS E SEUS NETOS POR ADOÇÃO
di: Gizele Bakman
Pubblicazione: (2023)
di: Gizele Bakman
Pubblicazione: (2023)
Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences
di: Han, Shanshan, et al.
Pubblicazione: (2025)
di: Han, Shanshan, et al.
Pubblicazione: (2025)
Understanding Communication Backends in Cross-Silo Federated Learning
di: Ziashahabi, Amir, et al.
Pubblicazione: (2026)
di: Ziashahabi, Amir, et al.
Pubblicazione: (2026)
Differentially Private Federated Learning without Noise Addition: When is it Possible?
di: Zhang, Jiang, et al.
Pubblicazione: (2024)
di: Zhang, Jiang, et al.
Pubblicazione: (2024)
ATP: Enabling Fast LLM Serving via Attention on Top Principal Keys
di: Niu, Yue, et al.
Pubblicazione: (2024)
di: Niu, Yue, et al.
Pubblicazione: (2024)
Metrics for Assessing Inclusivity and Empowerment of People for Supporting the Design of Inclusive Product Lifecycles
di: Yaldiz, Naz, et al.
Pubblicazione: (2024)
di: Yaldiz, Naz, et al.
Pubblicazione: (2024)
Statistical Inference for Score Decompositions
di: Dimitriadis, Timo, et al.
Pubblicazione: (2026)
di: Dimitriadis, Timo, et al.
Pubblicazione: (2026)
Green Fiscal Stance and Climate Pressure in Advanced Europe: Evidence From a Multidimensional Climate Index
di: Elif Duygu Kömürcüoğlu, et al.
Pubblicazione: (2026)
di: Elif Duygu Kömürcüoğlu, et al.
Pubblicazione: (2026)
ModalityMirror: Improving Audio Classification in Modality Heterogeneity Federated Learning with Multimodal Distillation
di: Feng, Tiantian, et al.
Pubblicazione: (2024)
di: Feng, Tiantian, et al.
Pubblicazione: (2024)
Documenti analoghi
-
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
di: Bakman, Yavuz Faruk, et al.
Pubblicazione: (2024) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
di: Kang, Sungmin, et al.
Pubblicazione: (2025) -
Reconsidering LLM Uncertainty Estimation Methods in the Wild
di: Bakman, Yavuz, et al.
Pubblicazione: (2025) -
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
di: Mushtaq, Erum, et al.
Pubblicazione: (2024) -
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
di: Bakman, Yavuz, et al.
Pubblicazione: (2026)