Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Duan, Jinhao, Cheng, Hao, Wang, Shiqi, Zavalny, Alex, Wang, Chenan, Xu, Renjing, Kailkhura, Bhavya, Xu, Kaidi |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
par: Fan, Haozhi, et autres
Publié: (2026)
par: Fan, Haozhi, et autres
Publié: (2026)
Mixture of Robust Experts (MoRE):A Robust Denoising Method towards multiple perturbations
par: Cheng, Hao, et autres
Publié: (2021)
par: Cheng, Hao, et autres
Publié: (2021)
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
par: Duan, Jinhao, et autres
Publié: (2025)
par: Duan, Jinhao, et autres
Publié: (2025)
TruthPrInt: Mitigating Large Vision-Language Models Object Hallucination Via Latent Truthful-Guided Pre-Intervention
par: Duan, Jinhao, et autres
Publié: (2025)
par: Duan, Jinhao, et autres
Publié: (2025)
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
par: Wang, Zhiyuan, et autres
Publié: (2024)
par: Wang, Zhiyuan, et autres
Publié: (2024)
Unveiling Typographic Deceptions: Insights of the Typographic Vulnerability in Large Vision-Language Model
par: Cheng, Hao, et autres
Publié: (2024)
par: Cheng, Hao, et autres
Publié: (2024)
ConU: Conformal Uncertainty in Large Language Models with Correctness Coverage Guarantees
par: Wang, Zhiyuan, et autres
Publié: (2024)
par: Wang, Zhiyuan, et autres
Publié: (2024)
Transfer Attack for Bad and Good: Explain and Boost Adversarial Transferability across Multimodal Large Language Models
par: Cheng, Hao, et autres
Publié: (2024)
par: Cheng, Hao, et autres
Publié: (2024)
GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluations
par: Duan, Jinhao, et autres
Publié: (2024)
par: Duan, Jinhao, et autres
Publié: (2024)
Can Protective Perturbation Safeguard Personal Data from Being Exploited by Stable Diffusion?
par: Zhao, Zhengyue, et autres
Publié: (2023)
par: Zhao, Zhengyue, et autres
Publié: (2023)
ACT-Diffusion: Efficient Adversarial Consistency Training for One-step Diffusion Models
par: Kong, Fei, et autres
Publié: (2023)
par: Kong, Fei, et autres
Publié: (2023)
Unlearnable Examples for Diffusion Models: Protect Data from Unauthorized Exploitation
par: Zhao, Zhengyue, et autres
Publié: (2023)
par: Zhao, Zhengyue, et autres
Publié: (2023)
COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees
par: Wang, Zhiyuan, et autres
Publié: (2025)
par: Wang, Zhiyuan, et autres
Publié: (2025)
Double Visual Defense: Adversarial Pre-training and Instruction Tuning for Improving Vision-Language Model Robustness
par: Wang, Zeyu, et autres
Publié: (2025)
par: Wang, Zeyu, et autres
Publié: (2025)
A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly
par: Yao, Yifan, et autres
Publié: (2023)
par: Yao, Yifan, et autres
Publié: (2023)
Dynamic Adversarial Attacks on Autonomous Driving Systems
par: Chahe, Amirhosein, et autres
Publié: (2023)
par: Chahe, Amirhosein, et autres
Publié: (2023)
Virtual Traffic Police: Large Language Model-Augmented Traffic Signal Control for Unforeseen Incidents
par: Wei, Shiqi, et autres
Publié: (2026)
par: Wei, Shiqi, et autres
Publié: (2026)
DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation
par: Hu, Wenhao, et autres
Publié: (2025)
par: Hu, Wenhao, et autres
Publié: (2025)
Sparse Neurons Carry Strong Signals of Question Ambiguity in LLMs
par: Zhang, Zhuoxuan, et autres
Publié: (2025)
par: Zhang, Zhuoxuan, et autres
Publié: (2025)
Forecasting Fails: Unveiling Evasion Attacks in Weather Prediction Models
par: Arif, Huzaifa, et autres
Publié: (2025)
par: Arif, Huzaifa, et autres
Publié: (2025)
Tackling Incomplete Data in Air Quality Prediction: A Bayesian Deep Learning Framework for Uncertainty Quantification
par: Pian, Yuzhuang, et autres
Publié: (2025)
par: Pian, Yuzhuang, et autres
Publié: (2025)
The Imitation Game: Using Large Language Models as Chatbots to Combat Chat-Based Cybercrimes
par: Yao, Yifan, et autres
Publié: (2025)
par: Yao, Yifan, et autres
Publié: (2025)
FedCluster: Boosting the Convergence of Federated Learning via Cluster-Cycling
par: Chen, Cheng, et autres
Publié: (2020)
par: Chen, Cheng, et autres
Publié: (2020)
Certifiably-Robust Federated Adversarial Learning via Randomized Smoothing
par: Chen, Cheng, et autres
Publié: (2021)
par: Chen, Cheng, et autres
Publié: (2021)
Training Dynamics of Transformers to Recognize Word Co-occurrence via Gradient Flow Analysis
par: Yang, Hongru, et autres
Publié: (2024)
par: Yang, Hongru, et autres
Publié: (2024)
Shifting Uncertainty to Critical Moments: Towards Reliable Uncertainty Quantification for VLA Model
par: Tang, Yanchuan, et autres
Publié: (2026)
par: Tang, Yanchuan, et autres
Publié: (2026)
KV Shifting Attention Enhances Language Modeling
par: Xu, Mingyu, et autres
Publié: (2024)
par: Xu, Mingyu, et autres
Publié: (2024)
SConU: Selective Conformal Uncertainty in Large Language Models
par: Wang, Zhiyuan, et autres
Publié: (2025)
par: Wang, Zhiyuan, et autres
Publié: (2025)
LRR-Bench: Left, Right or Rotate? Vision-Language models Still Struggle With Spatial Understanding Tasks
par: Kong, Fei, et autres
Publié: (2025)
par: Kong, Fei, et autres
Publié: (2025)
Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression
par: Hong, Junyuan, et autres
Publié: (2024)
par: Hong, Junyuan, et autres
Publié: (2024)
COPU: Conformal Prediction for Uncertainty Quantification in Natural Language Generation
par: Wang, Sean, et autres
Publié: (2025)
par: Wang, Sean, et autres
Publié: (2025)
Detecting Misbehaviors of Large Vision-Language Models by Evidential Uncertainty Quantification
par: Huang, Tao, et autres
Publié: (2026)
par: Huang, Tao, et autres
Publié: (2026)
The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages
par: Onyame, Eric, et autres
Publié: (2026)
par: Onyame, Eric, et autres
Publié: (2026)
LLM Unlearning Reveals a Stronger-Than-Expected Coreset Effect in Current Benchmarks
par: Pal, Soumyadeep, et autres
Publié: (2025)
par: Pal, Soumyadeep, et autres
Publié: (2025)
Improving Robustness In Sparse Autoencoders via Masked Regularization
par: Narayanaswamy, Vivek, et autres
Publié: (2026)
par: Narayanaswamy, Vivek, et autres
Publié: (2026)
End-to-End Mesh Optimization of a Hybrid Deep Learning Black-Box PDE Solver
par: Ma, Shaocong, et autres
Publié: (2024)
par: Ma, Shaocong, et autres
Publié: (2024)
Language Model Uncertainty Quantification with Attention Chain
par: Li, Yinghao, et autres
Publié: (2025)
par: Li, Yinghao, et autres
Publié: (2025)
Uncertainty Quantification for Clinical Outcome Predictions with (Large) Language Models
par: Chen, Zizhang, et autres
Publié: (2024)
par: Chen, Zizhang, et autres
Publié: (2024)
Data-Driven Prediction and Uncertainty Quantification of PWR Crud-Induced Power Shift Using Convolutional Neural Networks
par: Furlong, Aidan, et autres
Publié: (2024)
par: Furlong, Aidan, et autres
Publié: (2024)
CUQDS: Conformal Uncertainty Quantification under Distribution Shift for Trajectory Prediction
par: Huang, Huiqun, et autres
Publié: (2024)
par: Huang, Huiqun, et autres
Publié: (2024)
Documents similaires
-
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
par: Fan, Haozhi, et autres
Publié: (2026) -
Mixture of Robust Experts (MoRE):A Robust Denoising Method towards multiple perturbations
par: Cheng, Hao, et autres
Publié: (2021) -
UProp: Investigating the Uncertainty Propagation of LLMs in Multi-Step Agentic Decision-Making
par: Duan, Jinhao, et autres
Publié: (2025) -
TruthPrInt: Mitigating Large Vision-Language Models Object Hallucination Via Latent Truthful-Guided Pre-Intervention
par: Duan, Jinhao, et autres
Publié: (2025) -
Word-Sequence Entropy: Towards Uncertainty Estimation in Free-Form Medical Question Answering Applications and Beyond
par: Wang, Zhiyuan, et autres
Publié: (2024)