Addressing Uncertainty in LLMs to Enhance Reliability in Generative AI
Fuente:
arXiv
Saved in:
| Main Authors: | Kaur, Ramneet, Samplawski, Colin, Cobb, Adam D., Roy, Anirban, Matejek, Brian, Acharya, Manoj, Elenius, Daniel, Berenbeim, Alexander M., Pavlik, John A., Bastian, Nathaniel D., Jha, Susmit |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
by: Padhi, Trilok, et al.
Published: (2025)
by: Padhi, Trilok, et al.
Published: (2025)
Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs
by: Agarwal, Krishiv, et al.
Published: (2026)
by: Agarwal, Krishiv, et al.
Published: (2026)
Do Diffusion Models Dream of Electric Planes? Discrete and Continuous Simulation-Based Inference for Aircraft Design
by: Ghiglino, Aurelien, et al.
Published: (2026)
by: Ghiglino, Aurelien, et al.
Published: (2026)
From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents
by: Padhi, Trilok, et al.
Published: (2026)
by: Padhi, Trilok, et al.
Published: (2026)
Scalable Bayesian Low-Rank Adaptation of Large Language Models via Stochastic Variational Subspace Inference
by: Samplawski, Colin, et al.
Published: (2025)
by: Samplawski, Colin, et al.
Published: (2025)
Privacy Preserving In-Context-Learning Framework for Large Language Models
by: Bhusal, Bishnu, et al.
Published: (2025)
by: Bhusal, Bishnu, et al.
Published: (2025)
AGENT: An Aerial Vehicle Generation and Design Tool Using Large Language Models
by: Samplawski, Colin, et al.
Published: (2025)
by: Samplawski, Colin, et al.
Published: (2025)
Polysemantic Dropout: Conformal OOD Detection for Specialized LLMs
by: Gupta, Ayush, et al.
Published: (2025)
by: Gupta, Ayush, et al.
Published: (2025)
TeleLoRA: Teleporting Model-Specific Alignment Across LLMs
by: Lin, Xiao, et al.
Published: (2025)
by: Lin, Xiao, et al.
Published: (2025)
Closed-Loop Neural Activation Control in Vision-Language-Action Models
by: Babu, Abhijith, et al.
Published: (2026)
by: Babu, Abhijith, et al.
Published: (2026)
Resource-Constrained Heuristic for Max-SAT
by: Matejek, Brian, et al.
Published: (2024)
by: Matejek, Brian, et al.
Published: (2024)
SAFE-NID: Self-Attention with Normalizing-Flow Encodings for Network Intrusion Detection Dataset
by: Matejek, Brian, et al.
Published: (2025)
by: Matejek, Brian, et al.
Published: (2025)
Spatio-Temporal Pruning for Compressed Spiking Large Language Models
by: Jiang, Yi, et al.
Published: (2025)
by: Jiang, Yi, et al.
Published: (2025)
Backpropagation-Free Metropolis-Adjusted Langevin Algorithm
by: Cobb, Adam D., et al.
Published: (2025)
by: Cobb, Adam D., et al.
Published: (2025)
TOGA: Temporally Grounded Open-Ended Video QA with Weak Supervision
by: Gupta, Ayush, et al.
Published: (2025)
by: Gupta, Ayush, et al.
Published: (2025)
Taylor-Model Physics-Informed Neural Networks (PINNs) for Ordinary Differential Equations
by: Nagesh, Chandra Kanth, et al.
Published: (2025)
by: Nagesh, Chandra Kanth, et al.
Published: (2025)
Safety Monitoring for Learning-Enabled Cyber-Physical Systems in Out-of-Distribution Scenarios
by: Lin, Vivian, et al.
Published: (2025)
by: Lin, Vivian, et al.
Published: (2025)
Optimal Abstractions for Verifying Properties of Kolmogorov-Arnold Networks (KANs)
by: Schwartz, Noah, et al.
Published: (2026)
by: Schwartz, Noah, et al.
Published: (2026)
Second-Order Forward-Mode Automatic Differentiation for Optimization
by: Cobb, Adam D., et al.
Published: (2024)
by: Cobb, Adam D., et al.
Published: (2024)
Models Can Model, But Can't Bind: Structured Grounding in Text-to-Optimization
by: Gao, Zhiqi, et al.
Published: (2026)
by: Gao, Zhiqi, et al.
Published: (2026)
Shrinking POMCP: A Framework for Real-Time UAV Search and Rescue
by: Zhang, Yunuo, et al.
Published: (2024)
by: Zhang, Yunuo, et al.
Published: (2024)
Hessian-Enhanced Token Attribution (HETA): Interpreting Autoregressive LLMs
by: Pramanik, Vishal, et al.
Published: (2026)
by: Pramanik, Vishal, et al.
Published: (2026)
Melanoma Detection with Uncertainty Quantification
by: Kim, SangHyuk, et al.
Published: (2024)
by: Kim, SangHyuk, et al.
Published: (2024)
SCoOP: Semantic Consistent Opinion Pooling for Uncertainty Quantification in Multiple Vision-Language Model Systems
by: Yu, Chung-En Johnny, et al.
Published: (2026)
by: Yu, Chung-En Johnny, et al.
Published: (2026)
Concept-based Analysis of Neural Networks via Vision-Language Models
by: Mangal, Ravi, et al.
Published: (2024)
by: Mangal, Ravi, et al.
Published: (2024)
Characterizations and properties of solutions to parabolic problems of linear growth
by: Elenius, Theo
Published: (2025)
by: Elenius, Theo
Published: (2025)
Non-Markovian Quantum Control via Model Maximum Likelihood Estimation and Reinforcement Learning
by: Neema, Tanmay, et al.
Published: (2024)
by: Neema, Tanmay, et al.
Published: (2024)
Neurosymbolic AI Transfer Learning Improves Network Intrusion Detection
by: Tran, Huynh T. T., et al.
Published: (2025)
by: Tran, Huynh T. T., et al.
Published: (2025)
Probabilistic Foundations for Metacognition via Hybrid-AI
by: Shakarian, Paulo, et al.
Published: (2025)
by: Shakarian, Paulo, et al.
Published: (2025)
Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion
by: Pramanik, Vishal, et al.
Published: (2026)
by: Pramanik, Vishal, et al.
Published: (2026)
A Synergistic Approach In Network Intrusion Detection By Neurosymbolic AI
by: Bizzarri, Alice, et al.
Published: (2024)
by: Bizzarri, Alice, et al.
Published: (2024)
On the Dataless Training of Neural Networks
by: Velasquez, Alvaro, et al.
Published: (2025)
by: Velasquez, Alvaro, et al.
Published: (2025)
Interactive Critique-Revision Training for Reliable Structured LLM Generation
by: Yu, Fei Xu, et al.
Published: (2026)
by: Yu, Fei Xu, et al.
Published: (2026)
A Red Teaming Framework for Evaluating Robustness of AI-enabled Security Orchestration, Automation, and Response Systems
by: Shaikh, Ayan Javeed, et al.
Published: (2026)
by: Shaikh, Ayan Javeed, et al.
Published: (2026)
Generative AI in the Construction Industry: Opportunities & Challenges
by: Ghimire, Prashnna, et al.
Published: (2023)
by: Ghimire, Prashnna, et al.
Published: (2023)
ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models
by: Yu, Chung-En Johnny, et al.
Published: (2025)
by: Yu, Chung-En Johnny, et al.
Published: (2025)
Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling
by: Yu, Chung-En Johnny, et al.
Published: (2025)
by: Yu, Chung-En Johnny, et al.
Published: (2025)
Inside the Jaynes-Cummings sum
by: Pavlik, S. I.
Published: (2023)
by: Pavlik, S. I.
Published: (2023)
When self-similarity meets mass spectrum and anisotropy
by: Pavlík, Václav
Published: (2026)
by: Pavlík, Václav
Published: (2026)
Study of argon/oxygen plasma used for creation of aluminium oxide thin films
by: Jaroslav Pavlík
Published: (1999)
by: Jaroslav Pavlík
Published: (1999)
Similar Items
-
Calibrating Uncertainty Quantification of Multi-Modal LLMs using Grounding
by: Padhi, Trilok, et al.
Published: (2025) -
Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs
by: Agarwal, Krishiv, et al.
Published: (2026) -
Do Diffusion Models Dream of Electric Planes? Discrete and Continuous Simulation-Based Inference for Aircraft Design
by: Ghiglino, Aurelien, et al.
Published: (2026) -
From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents
by: Padhi, Trilok, et al.
Published: (2026) -
Scalable Bayesian Low-Rank Adaptation of Large Language Models via Stochastic Variational Subspace Inference
by: Samplawski, Colin, et al.
Published: (2025)