Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Linyu, Pan, Yu, Li, Xiaocheng, Chen, Guanting |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Language Model-Driven Semi-Supervised Ensemble Framework for Illicit Market Detection Across Deep/Dark Web and Social Platforms
di: Yazdanjue, Navid, et al.
Pubblicazione: (2025)
di: Yazdanjue, Navid, et al.
Pubblicazione: (2025)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
di: Wang, Youkang, et al.
Pubblicazione: (2025)
di: Wang, Youkang, et al.
Pubblicazione: (2025)
Clustering in pure-attention hardmax transformers and its role in sentiment analysis
di: Alcalde, Albert, et al.
Pubblicazione: (2024)
di: Alcalde, Albert, et al.
Pubblicazione: (2024)
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
di: Fu, Tianyu, et al.
Pubblicazione: (2025)
di: Fu, Tianyu, et al.
Pubblicazione: (2025)
Bridging the Language Gap: Enhancing Multilingual Prompt-Based Code Generation in LLMs via Zero-Shot Cross-Lingual Transfer
di: Li, Mingda, et al.
Pubblicazione: (2024)
di: Li, Mingda, et al.
Pubblicazione: (2024)
Self-Attention as Transport: Limits of Symmetric Spectral Diagnostics
di: Dahlem, Dominik, et al.
Pubblicazione: (2026)
di: Dahlem, Dominik, et al.
Pubblicazione: (2026)
Parameter-Efficient Transformer Embeddings
di: Ndubuaku, Henry, et al.
Pubblicazione: (2025)
di: Ndubuaku, Henry, et al.
Pubblicazione: (2025)
DYNAMAX: Dynamic computing for Transformers and Mamba based architectures
di: Nogales, Miguel, et al.
Pubblicazione: (2025)
di: Nogales, Miguel, et al.
Pubblicazione: (2025)
A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness
di: Wang, Fali, et al.
Pubblicazione: (2024)
di: Wang, Fali, et al.
Pubblicazione: (2024)
Pay Attention to What You Need
di: Gao, Yifei, et al.
Pubblicazione: (2023)
di: Gao, Yifei, et al.
Pubblicazione: (2023)
OPENXRD: A Comprehensive Benchmark Framework for LLM/MLLM XRD Question Answering
di: Vosoughi, Ali, et al.
Pubblicazione: (2025)
di: Vosoughi, Ali, et al.
Pubblicazione: (2025)
Strategic Doctrine Language Models (sdLM): A Learning-System Framework for Doctrinal Consistency and Geopolitical Forecasting
di: Imanov, Olaf Yunus Laitinen, et al.
Pubblicazione: (2026)
di: Imanov, Olaf Yunus Laitinen, et al.
Pubblicazione: (2026)
Inference acceleration for large language models using "stairs" assisted greedy generation
di: Grigaliūnas, Domas, et al.
Pubblicazione: (2024)
di: Grigaliūnas, Domas, et al.
Pubblicazione: (2024)
Mechanistic Analysis of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
di: Imanov, Olaf Yunus Laitinen
Pubblicazione: (2026)
di: Imanov, Olaf Yunus Laitinen
Pubblicazione: (2026)
Latent Object Permanence: Topological Phase Transitions, Free-Energy Principles, and Renormalization Group Flows in Deep Transformer Manifolds
di: Alpay, Faruk, et al.
Pubblicazione: (2026)
di: Alpay, Faruk, et al.
Pubblicazione: (2026)
Smoothed Embeddings for Robust Language Models
di: Hase, Ryo, et al.
Pubblicazione: (2025)
di: Hase, Ryo, et al.
Pubblicazione: (2025)
Retrieval-augmented code completion for local projects using large language models
di: Hostnik, Marko, et al.
Pubblicazione: (2024)
di: Hostnik, Marko, et al.
Pubblicazione: (2024)
Exact Sequence Interpolation with Transformers
di: Alcalde, Albert, et al.
Pubblicazione: (2025)
di: Alcalde, Albert, et al.
Pubblicazione: (2025)
Empirical analysis of binding precedent efficiency in Brazilian Supreme Court via case classification
di: Tinarrage, Raphaël, et al.
Pubblicazione: (2024)
di: Tinarrage, Raphaël, et al.
Pubblicazione: (2024)
PolyTruth: Multilingual Disinformation Detection using Transformer-Based Language Models
di: Gouliev, Zaur, et al.
Pubblicazione: (2025)
di: Gouliev, Zaur, et al.
Pubblicazione: (2025)
Beyond Long Context: When Semantics Matter More than Tokens
di: Chawdhury, Tarun Kumar, et al.
Pubblicazione: (2025)
di: Chawdhury, Tarun Kumar, et al.
Pubblicazione: (2025)
MetaCheckGPT -- A Multi-task Hallucination Detector Using LLM Uncertainty and Meta-models
di: Mehta, Rahul, et al.
Pubblicazione: (2024)
di: Mehta, Rahul, et al.
Pubblicazione: (2024)
Forging GEMs: Advancing Greek NLP through Quality-Based Corpus Curation
di: Apostolopoulou, Alexandra, et al.
Pubblicazione: (2025)
di: Apostolopoulou, Alexandra, et al.
Pubblicazione: (2025)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
di: Chen, Tiejin, et al.
Pubblicazione: (2026)
di: Chen, Tiejin, et al.
Pubblicazione: (2026)
ReFactor GNNs: Revisiting Factorisation-based Models from a Message-Passing Perspective
di: Chen, Yihong, et al.
Pubblicazione: (2022)
di: Chen, Yihong, et al.
Pubblicazione: (2022)
Surfing the modeling of PoS taggers in low-resource scenarios
di: Ferro, Manuel Vilares, et al.
Pubblicazione: (2024)
di: Ferro, Manuel Vilares, et al.
Pubblicazione: (2024)
Atyaephyra at SemEval-2025 Task 4: Low-Rank Negative Preference Optimization
di: Bronec, Jan, et al.
Pubblicazione: (2025)
di: Bronec, Jan, et al.
Pubblicazione: (2025)
Can Out-of-Distribution Evaluations Uncover Reliance on Shortcuts? A Case Study in Question Answering
di: Štefánik, Michal, et al.
Pubblicazione: (2025)
di: Štefánik, Michal, et al.
Pubblicazione: (2025)
Extracting Sentence Embeddings from Pretrained Transformer Models
di: Stankevičius, Lukas, et al.
Pubblicazione: (2024)
di: Stankevičius, Lukas, et al.
Pubblicazione: (2024)
Sentiment Analysis of Lithuanian Online Reviews Using Large Language Models
di: Vileikytė, Brigita, et al.
Pubblicazione: (2024)
di: Vileikytė, Brigita, et al.
Pubblicazione: (2024)
mHC-SSM: Manifold-Constrained Hyper-Connections for State Space Language Models with Stream-Specialized Adapters
di: Mutlu, Abdulvahap, et al.
Pubblicazione: (2026)
di: Mutlu, Abdulvahap, et al.
Pubblicazione: (2026)
A Generalization Bound for a Family of Implicit Networks
di: Fung, Samy Wu, et al.
Pubblicazione: (2024)
di: Fung, Samy Wu, et al.
Pubblicazione: (2024)
Mediator: Memory-efficient LLM Merging with Less Parameter Conflicts and Uncertainty Based Routing
di: Lai, Kunfeng, et al.
Pubblicazione: (2025)
di: Lai, Kunfeng, et al.
Pubblicazione: (2025)
Healthy LLMs? Benchmarking LLM Knowledge of UK Government Public Health Information
di: Harris, Joshua, et al.
Pubblicazione: (2025)
di: Harris, Joshua, et al.
Pubblicazione: (2025)
Evaluating Embedding Generalization: How LLMs, LoRA, and SLERP Shape Representational Geometry
di: Kabane, Siyaxolisa
Pubblicazione: (2025)
di: Kabane, Siyaxolisa
Pubblicazione: (2025)
Linguistic Collapse: Neural Collapse in (Large) Language Models
di: Wu, Robert, et al.
Pubblicazione: (2024)
di: Wu, Robert, et al.
Pubblicazione: (2024)
ProbeScale: Probing Analysis to Optimize Neural Scaling Laws for Efficient Small Language Model Inference
di: Das, Sourav
Pubblicazione: (2026)
di: Das, Sourav
Pubblicazione: (2026)
Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts
di: Garg, Saloni, et al.
Pubblicazione: (2026)
di: Garg, Saloni, et al.
Pubblicazione: (2026)
Research on a hybrid LSTM-CNN-Attention model for text-based web content classification
di: Kuz, Mykola, et al.
Pubblicazione: (2025)
di: Kuz, Mykola, et al.
Pubblicazione: (2025)
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
di: Hill, Brennen, et al.
Pubblicazione: (2025)
di: Hill, Brennen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Language Model-Driven Semi-Supervised Ensemble Framework for Illicit Market Detection Across Deep/Dark Web and Social Platforms
di: Yazdanjue, Navid, et al.
Pubblicazione: (2025) -
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
di: Wang, Youkang, et al.
Pubblicazione: (2025) -
Clustering in pure-attention hardmax transformers and its role in sentiment analysis
di: Alcalde, Albert, et al.
Pubblicazione: (2024) -
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
di: Fu, Tianyu, et al.
Pubblicazione: (2025) -
Bridging the Language Gap: Enhancing Multilingual Prompt-Based Code Generation in LLMs via Zero-Shot Cross-Lingual Transfer
di: Li, Mingda, et al.
Pubblicazione: (2024)