Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Yaldiz, Duygu Nur, Bakman, Yavuz Faruk, Buyukates, Baturalp, Tao, Chenyang, Ramakrishna, Anil, Dimitriadis, Dimitrios, Zhao, Jieyu, Avestimehr, Salman |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
por: Bakman, Yavuz Faruk, et al.
Publicado: (2024)
por: Bakman, Yavuz Faruk, et al.
Publicado: (2024)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
por: Kang, Sungmin, et al.
Publicado: (2025)
por: Kang, Sungmin, et al.
Publicado: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
por: Bakman, Yavuz, et al.
Publicado: (2025)
por: Bakman, Yavuz, et al.
Publicado: (2025)
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
por: Mushtaq, Erum, et al.
Publicado: (2024)
por: Mushtaq, Erum, et al.
Publicado: (2024)
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
por: Bakman, Yavuz, et al.
Publicado: (2026)
por: Bakman, Yavuz, et al.
Publicado: (2026)
HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
por: Mushtaq, Erum, et al.
Publicado: (2025)
por: Mushtaq, Erum, et al.
Publicado: (2025)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
por: Oguz, Metehan, et al.
Publicado: (2025)
por: Oguz, Metehan, et al.
Publicado: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
por: Ziashahabi, Amir, et al.
Publicado: (2025)
por: Ziashahabi, Amir, et al.
Publicado: (2025)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
por: Bakman, Yavuz, et al.
Publicado: (2025)
por: Bakman, Yavuz, et al.
Publicado: (2025)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
por: Wang, Nien-Shao, et al.
Publicado: (2025)
por: Wang, Nien-Shao, et al.
Publicado: (2025)
Leveraging Uncertainty Estimation for Efficient LLM Routing
por: Zhang, Tuo, et al.
Publicado: (2025)
por: Zhang, Tuo, et al.
Publicado: (2025)
Maverick-Aware Shapley Valuation for Client Selection in Federated Learning
por: Yang, Mengwei, et al.
Publicado: (2024)
por: Yang, Mengwei, et al.
Publicado: (2024)
FedGrAINS: Personalized SubGraph Federated Learning with Adaptive Neighbor Sampling
por: Ceyani, Emir, et al.
Publicado: (2025)
por: Ceyani, Emir, et al.
Publicado: (2025)
Renormalization Group flow, Optimal Transport and Diffusion-based Generative Model
por: Sheshmani, Artan, et al.
Publicado: (2024)
por: Sheshmani, Artan, et al.
Publicado: (2024)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
por: Oğuz, Metehan, et al.
Publicado: (2024)
por: Oğuz, Metehan, et al.
Publicado: (2024)
Balancing Information Accuracy and Response Timeliness in Networked LLMs
por: Turkmen, Yigit, et al.
Publicado: (2025)
por: Turkmen, Yigit, et al.
Publicado: (2025)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
por: Yaldiz, Duygu Nur, et al.
Publicado: (2025)
por: Yaldiz, Duygu Nur, et al.
Publicado: (2025)
Kick Bad Guys Out! Conditionally Activated Anomaly Detection in Federated Learning with Zero-Knowledge Proof Verification
por: Han, Shanshan, et al.
Publicado: (2023)
por: Han, Shanshan, et al.
Publicado: (2023)
Privacy-preserving Prompt Personalization in Federated Learning for Multimodal Large Language Models
por: Hou, Sizai, et al.
Publicado: (2025)
por: Hou, Sizai, et al.
Publicado: (2025)
Don't Always Pick the Highest-Performing Model: An Information Theoretic View of LLM Ensemble Selection
por: Turkmen, Yigit, et al.
Publicado: (2026)
por: Turkmen, Yigit, et al.
Publicado: (2026)
Balancing Classification and Calibration Performance in Decision-Making LLMs via Calibration Aware Reinforcement Learning
por: Yaldiz, Duygu Nur, et al.
Publicado: (2026)
por: Yaldiz, Duygu Nur, et al.
Publicado: (2026)
Revisiting OPRO: The Limitations of Small-Scale LLMs as Optimizers
por: Zhang, Tuo, et al.
Publicado: (2024)
por: Zhang, Tuo, et al.
Publicado: (2024)
FedSecurity: Benchmarking Attacks and Defenses in Federated Learning and Federated LLMs
por: Han, Shanshan, et al.
Publicado: (2023)
por: Han, Shanshan, et al.
Publicado: (2023)
GPT-FL: Generative Pre-trained Model-Assisted Federated Learning
por: Zhang, Tuo, et al.
Publicado: (2023)
por: Zhang, Tuo, et al.
Publicado: (2023)
Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning
por: Ozfatura, Emre, et al.
Publicado: (2024)
por: Ozfatura, Emre, et al.
Publicado: (2024)
Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation
por: Aravindan, Ashwath Vaithinathan, et al.
Publicado: (2025)
por: Aravindan, Ashwath Vaithinathan, et al.
Publicado: (2025)
Backdoor Defense in Diffusion Models via Spatial Attention Unlearning
por: Jha, Abha, et al.
Publicado: (2025)
por: Jha, Abha, et al.
Publicado: (2025)
Reseña de "Class Acts: Service and Inequality in Luxury Hotels" de Rachel Sherman
por: Duygu Salman
Publicado: (2012)
por: Duygu Salman
Publicado: (2012)
Rethinking of Cities, Culture and Tourism within a Creative Perspective
por: Duygu Salman
Publicado: (2010)
por: Duygu Salman
Publicado: (2010)
Spatiotemporal variability and trends of droughts in the Mediterranean coastal region of Türkiye
por: Erdal Kesgin, et al.
Publicado: (2024)
por: Erdal Kesgin, et al.
Publicado: (2024)
An objective criterion for evaluating new physical theories
por: Bakman, Yefim
Publicado: (2025)
por: Bakman, Yefim
Publicado: (2025)
POR ENTRE AS TRAMAS FAMILIARES: Avós JUDEUS E SEUS NETOS POR ADOÇÃO
por: Gizele Bakman
Publicado: (2023)
por: Gizele Bakman
Publicado: (2023)
Bridging the Safety Gap: A Guardrail Pipeline for Trustworthy LLM Inferences
por: Han, Shanshan, et al.
Publicado: (2025)
por: Han, Shanshan, et al.
Publicado: (2025)
Understanding Communication Backends in Cross-Silo Federated Learning
por: Ziashahabi, Amir, et al.
Publicado: (2026)
por: Ziashahabi, Amir, et al.
Publicado: (2026)
Differentially Private Federated Learning without Noise Addition: When is it Possible?
por: Zhang, Jiang, et al.
Publicado: (2024)
por: Zhang, Jiang, et al.
Publicado: (2024)
ATP: Enabling Fast LLM Serving via Attention on Top Principal Keys
por: Niu, Yue, et al.
Publicado: (2024)
por: Niu, Yue, et al.
Publicado: (2024)
Metrics for Assessing Inclusivity and Empowerment of People for Supporting the Design of Inclusive Product Lifecycles
por: Yaldiz, Naz, et al.
Publicado: (2024)
por: Yaldiz, Naz, et al.
Publicado: (2024)
Statistical Inference for Score Decompositions
por: Dimitriadis, Timo, et al.
Publicado: (2026)
por: Dimitriadis, Timo, et al.
Publicado: (2026)
Green Fiscal Stance and Climate Pressure in Advanced Europe: Evidence From a Multidimensional Climate Index
por: Elif Duygu Kömürcüoğlu, et al.
Publicado: (2026)
por: Elif Duygu Kömürcüoğlu, et al.
Publicado: (2026)
ModalityMirror: Improving Audio Classification in Modality Heterogeneity Federated Learning with Multimodal Distillation
por: Feng, Tiantian, et al.
Publicado: (2024)
por: Feng, Tiantian, et al.
Publicado: (2024)
Ejemplares similares
-
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
por: Bakman, Yavuz Faruk, et al.
Publicado: (2024) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
por: Kang, Sungmin, et al.
Publicado: (2025) -
Reconsidering LLM Uncertainty Estimation Methods in the Wild
por: Bakman, Yavuz, et al.
Publicado: (2025) -
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
por: Mushtaq, Erum, et al.
Publicado: (2024) -
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
por: Bakman, Yavuz, et al.
Publicado: (2026)