Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aichberger, Lukas, Schweighofer, Kajetan, Hochreiter, Sepp |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On Information-Theoretic Measures of Predictive Uncertainty
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2024)
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2024)
Improving Uncertainty Estimation through Semantically Diverse Language Generation
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024)
Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language Generation
von: Ielanskyi, Mykyta, et al.
Veröffentlicht: (2025)
von: Ielanskyi, Mykyta, et al.
Veröffentlicht: (2025)
The Disparate Benefits of Deep Ensembles
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2024)
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2024)
xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
Unlocking the Working Memory of Large Language Models for Latent Reasoning
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2026)
Rethinking Losses for Diffusion Bridge Samplers
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2025)
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2025)
Uncertainty Quantification for Regression using Proper Scoring Rules
von: Fishkov, Alexander, et al.
Veröffentlicht: (2025)
von: Fishkov, Alexander, et al.
Veröffentlicht: (2025)
Overcoming Saturation in Density Ratio Estimation by Iterated Regularization
von: Gruber, Lukas, et al.
Veröffentlicht: (2024)
von: Gruber, Lukas, et al.
Veröffentlicht: (2024)
Efficient Pre-Training of LLMs through Truncated SVD Layers
von: Kamali, Kaivan, et al.
Veröffentlicht: (2026)
von: Kamali, Kaivan, et al.
Veröffentlicht: (2026)
MIM-Refiner: A Contrastive Learning Boost from Intermediate Pre-Trained Representations
von: Alkin, Benedikt, et al.
Veröffentlicht: (2024)
von: Alkin, Benedikt, et al.
Veröffentlicht: (2024)
Adaptive Retrieval helps Reasoning in LLMs -- but mostly if it's not used
von: Shakya, Srijan, et al.
Veröffentlicht: (2026)
von: Shakya, Srijan, et al.
Veröffentlicht: (2026)
Overcoming Forgetting in LLM Fine-Tuning with Evolution Strategies
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2026)
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2026)
Safe and Certifiable AI Systems: Concepts, Challenges, and Lessons Learned
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2025)
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2025)
FlashRNN: I/O-Aware Optimization of Traditional RNNs on modern hardware
von: Pöppel, Korbinian, et al.
Veröffentlicht: (2024)
von: Pöppel, Korbinian, et al.
Veröffentlicht: (2024)
A Diffusion Model Framework for Unsupervised Neural Combinatorial Optimization
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2024)
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2024)
Pre-trained Forecasting Models: Strong Zero-Shot Feature Extractors for Time Series Classification
von: Auer, Andreas, et al.
Veröffentlicht: (2025)
von: Auer, Andreas, et al.
Veröffentlicht: (2025)
Contrastive Abstraction for Reinforcement Learning
von: Patil, Vihang, et al.
Veröffentlicht: (2024)
von: Patil, Vihang, et al.
Veröffentlicht: (2024)
Tiled Flash Linear Attention: More Efficient Linear RNN and xLSTM Kernels
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)
pLSTM: parallelizable Linear Source Transition Mark networks
von: Pöppel, Korbinian, et al.
Veröffentlicht: (2025)
von: Pöppel, Korbinian, et al.
Veröffentlicht: (2025)
Simplified priors for Object-Centric Learning
von: Patil, Vihang, et al.
Veröffentlicht: (2024)
von: Patil, Vihang, et al.
Veröffentlicht: (2024)
The Offline-Frontier Shift: Diagnosing Distributional Limits in Generative Multi-Objective Optimization
von: Holly, Stephanie, et al.
Veröffentlicht: (2026)
von: Holly, Stephanie, et al.
Veröffentlicht: (2026)
Linear Alignment of Vision-language Models for Image Captioning
von: Paischer, Fabian, et al.
Veröffentlicht: (2023)
von: Paischer, Fabian, et al.
Veröffentlicht: (2023)
Parameter Efficient Fine-tuning via Explained Variance Adaptation
von: Paischer, Fabian, et al.
Veröffentlicht: (2024)
von: Paischer, Fabian, et al.
Veröffentlicht: (2024)
TiRex: Zero-Shot Forecasting Across Long and Short Horizons with Enhanced In-Context Learning
von: Auer, Andreas, et al.
Veröffentlicht: (2025)
von: Auer, Andreas, et al.
Veröffentlicht: (2025)
AP-OOD: Attention Pooling for Out-of-Distribution Detection
von: Hofmann, Claus, et al.
Veröffentlicht: (2026)
von: Hofmann, Claus, et al.
Veröffentlicht: (2026)
Energy-based Hopfield Boosting for Out-of-Distribution Detection
von: Hofmann, Claus, et al.
Veröffentlicht: (2024)
von: Hofmann, Claus, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Decision Transformer: External Memory for In-context RL
von: Schmied, Thomas, et al.
Veröffentlicht: (2024)
von: Schmied, Thomas, et al.
Veröffentlicht: (2024)
Vision-LSTM: xLSTM as Generic Vision Backbone
von: Alkin, Benedikt, et al.
Veröffentlicht: (2024)
von: Alkin, Benedikt, et al.
Veröffentlicht: (2024)
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025)
VN-EGNN: E(3)-Equivariant Graph Neural Networks with Virtual Nodes Enhance Protein Binding Site Identification
von: Sestak, Florian, et al.
Veröffentlicht: (2024)
von: Sestak, Florian, et al.
Veröffentlicht: (2024)
Geometry-Informed Neural Networks
von: Berzins, Arturs, et al.
Veröffentlicht: (2024)
von: Berzins, Arturs, et al.
Veröffentlicht: (2024)
SymbolicAI: A framework for logic-based approaches combining generative models and solvers
von: Dinu, Marius-Constantin, et al.
Veröffentlicht: (2024)
von: Dinu, Marius-Constantin, et al.
Veröffentlicht: (2024)
Effective Distillation to Hybrid xLSTM Architectures
von: Hauzenberger, Lukas, et al.
Veröffentlicht: (2026)
von: Hauzenberger, Lukas, et al.
Veröffentlicht: (2026)
From Risk to Uncertainty: Generating Predictive Uncertainty Measures via Bayesian Estimation
von: Kotelevskii, Nikita, et al.
Veröffentlicht: (2024)
von: Kotelevskii, Nikita, et al.
Veröffentlicht: (2024)
Rethinking Aleatoric and Epistemic Uncertainty
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)
A Monte Carlo Framework for Calibrated Uncertainty Estimation in Sequence Prediction
von: Yang, Qidong, et al.
Veröffentlicht: (2024)
von: Yang, Qidong, et al.
Veröffentlicht: (2024)
On Context-Content Uncertainty Principle
von: Li, Xin
Veröffentlicht: (2025)
von: Li, Xin
Veröffentlicht: (2025)
Scalable Discrete Diffusion Samplers: Combinatorial Optimization and Statistical Physics
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2025)
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2025)
A Large Recurrent Action Model: xLSTM enables Fast Inference for Robotics Tasks
von: Schmied, Thomas, et al.
Veröffentlicht: (2024)
von: Schmied, Thomas, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
On Information-Theoretic Measures of Predictive Uncertainty
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2024) -
Improving Uncertainty Estimation through Semantically Diverse Language Generation
von: Aichberger, Lukas, et al.
Veröffentlicht: (2024) -
Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language Generation
von: Ielanskyi, Mykyta, et al.
Veröffentlicht: (2025) -
The Disparate Benefits of Deep Ensembles
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2024) -
xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity
von: Beck, Maximilian, et al.
Veröffentlicht: (2025)