Mitigating LLM Hallucinations via Conformal Abstention
Fuente:
arXiv
Guardado en:
| Autores principales: | Yadkori, Yasin Abbasi, Kuzborskij, Ilja, Stutz, David, György, András, Fisch, Adam, Doucet, Arnaud, Beloshapka, Iuliya, Weng, Wei-Hung, Yang, Yao-Yuan, Szepesvári, Csaba, Cemgil, Ali Taylan, Tomasev, Nenad |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
To Believe or Not to Believe Your LLM
por: Yadkori, Yasin Abbasi, et al.
Publicado: (2024)
por: Yadkori, Yasin Abbasi, et al.
Publicado: (2024)
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
por: Kuzborskij, Ilja, et al.
Publicado: (2025)
por: Kuzborskij, Ilja, et al.
Publicado: (2025)
Low-rank bias, weight decay, and model merging in neural networks
por: Kuzborskij, Ilja, et al.
Publicado: (2025)
por: Kuzborskij, Ilja, et al.
Publicado: (2025)
Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence
por: György, András, et al.
Publicado: (2025)
por: György, András, et al.
Publicado: (2025)
To See the Unseen: on the Generalization Ability of Transformers in Symbolic Reasoning
por: Lazić, Nevena, et al.
Publicado: (2026)
por: Lazić, Nevena, et al.
Publicado: (2026)
Conformalized Credal Regions for Classification with Ambiguous Ground Truth
por: Caprio, Michele, et al.
Publicado: (2024)
por: Caprio, Michele, et al.
Publicado: (2024)
Frontier LLMs Still Struggle with Simple Reasoning Tasks
por: Malek, Alan, et al.
Publicado: (2025)
por: Malek, Alan, et al.
Publicado: (2025)
The signature and cusp geometry of hyperbolic knots
por: Davies, Alex, et al.
Publicado: (2021)
por: Davies, Alex, et al.
Publicado: (2021)
Sharper Guarantees for Misspecified Kernelized Bandit Optimization
por: Maran, Davide, et al.
Publicado: (2026)
por: Maran, Davide, et al.
Publicado: (2026)
Associative algebras in $CR$-geometry
por: Beloshapka, Valerii
Publicado: (2026)
por: Beloshapka, Valerii
Publicado: (2026)
On-Average Stability of Multipass Preconditioned SGD and Effective Dimension
por: Vary, Simon, et al.
Publicado: (2026)
por: Vary, Simon, et al.
Publicado: (2026)
Best of both worlds: Stochastic & adversarial best-arm identification
por: Abbasi-Yadkori, Yasin, et al.
Publicado: (2026)
por: Abbasi-Yadkori, Yasin, et al.
Publicado: (2026)
Intelligent AI Delegation
por: Tomašev, Nenad, et al.
Publicado: (2026)
por: Tomašev, Nenad, et al.
Publicado: (2026)
I-CALM: Incentivizing Confidence-Aware Abstention for LLM Hallucination Mitigation
por: Zong, Haotian, et al.
Publicado: (2026)
por: Zong, Haotian, et al.
Publicado: (2026)
Rectifying Regression in Reinforcement Learning
por: Ayoub, Alex, et al.
Publicado: (2025)
por: Ayoub, Alex, et al.
Publicado: (2025)
Trajectory Data Suffices for Statistically Efficient Learning in Offline RL with Linear $q^π$-Realizability and Concentrability
por: Tkachuk, Volodymyr, et al.
Publicado: (2024)
por: Tkachuk, Volodymyr, et al.
Publicado: (2024)
Balancing optimism and pessimism in offline-to-online learning
por: Sentenac, Flore, et al.
Publicado: (2025)
por: Sentenac, Flore, et al.
Publicado: (2025)
Sharp analysis of linear ensemble sampling
por: Akhavan, Arya, et al.
Publicado: (2026)
por: Akhavan, Arya, et al.
Publicado: (2026)
Ensemble sampling for linear bandits: small ensembles suffice
por: Janz, David, et al.
Publicado: (2023)
por: Janz, David, et al.
Publicado: (2023)
Confident Natural Policy Gradient for Local Planning in $q_π$-realizable Constrained MDPs
por: Tian, Tian, et al.
Publicado: (2024)
por: Tian, Tian, et al.
Publicado: (2024)
Better-than-KL PAC-Bayes Bounds
por: Kuzborskij, Ilja, et al.
Publicado: (2024)
por: Kuzborskij, Ilja, et al.
Publicado: (2024)
DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs
por: Zhang, Yuanhe, et al.
Publicado: (2025)
por: Zhang, Yuanhe, et al.
Publicado: (2025)
Evaluating Model Bias Requires Characterizing its Mistakes
por: Albuquerque, Isabela, et al.
Publicado: (2024)
por: Albuquerque, Isabela, et al.
Publicado: (2024)
Exploration via linearly perturbed loss minimisation
por: Janz, David, et al.
Publicado: (2023)
por: Janz, David, et al.
Publicado: (2023)
Geometry-Calibrated Conformal Abstention for Language Models
por: Xu, Rui, et al.
Publicado: (2026)
por: Xu, Rui, et al.
Publicado: (2026)
The unknotting number, hard unknot diagrams, and reinforcement learning
por: Applebaum, Taylor, et al.
Publicado: (2024)
por: Applebaum, Taylor, et al.
Publicado: (2024)
Distributional AGI Safety
por: Tomašev, Nenad, et al.
Publicado: (2025)
por: Tomašev, Nenad, et al.
Publicado: (2025)
On the Advice Complexity of Online Matching on the Line
por: Csaba, Béla, et al.
Publicado: (2024)
por: Csaba, Béla, et al.
Publicado: (2024)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
por: Harzli, Ouns El, et al.
Publicado: (2026)
por: Harzli, Ouns El, et al.
Publicado: (2026)
Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations
por: Tomani, Christian, et al.
Publicado: (2024)
por: Tomani, Christian, et al.
Publicado: (2024)
Calibrated Counterfactual Conformal Fairness ($C^3F$): Post-hoc, Shift-Aware Coverage Parity via Conformal Prediction and Counterfactual Regularization
por: Alpay, Faruk, et al.
Publicado: (2025)
por: Alpay, Faruk, et al.
Publicado: (2025)
Effective Skill Unlearning through Intervention and Abstention
por: Li, Yongce, et al.
Publicado: (2025)
por: Li, Yongce, et al.
Publicado: (2025)
HRM-Text: Efficient Pretraining Beyond Scaling
por: Wang, Guan, et al.
Publicado: (2026)
por: Wang, Guan, et al.
Publicado: (2026)
Hierarchical Reasoning Model
por: Wang, Guan, et al.
Publicado: (2025)
por: Wang, Guan, et al.
Publicado: (2025)
Learning to Reason Efficiently with Discounted Reinforcement Learning
por: Ayoub, Alex, et al.
Publicado: (2025)
por: Ayoub, Alex, et al.
Publicado: (2025)
Eluder dimension: localise it!
por: Bakhtiari, Alireza, et al.
Publicado: (2026)
por: Bakhtiari, Alireza, et al.
Publicado: (2026)
Optimistic Policy Optimization is Provably Efficient in Non-stationary MDPs
por: Zhong, Han, et al.
Publicado: (2021)
por: Zhong, Han, et al.
Publicado: (2021)
Learning What to Recommend: Minimax Optimal Simple Regret in Logistic Bandits
por: Liu, Shuai, et al.
Publicado: (2026)
por: Liu, Shuai, et al.
Publicado: (2026)
Almost Free: Self-concordance in Natural Exponential Families and an Application to Bandits
por: Liu, Shuai, et al.
Publicado: (2024)
por: Liu, Shuai, et al.
Publicado: (2024)
Conformalized Credal Set Predictors
por: Javanmardi, Alireza, et al.
Publicado: (2024)
por: Javanmardi, Alireza, et al.
Publicado: (2024)
Ejemplares similares
-
To Believe or Not to Believe Your LLM
por: Yadkori, Yasin Abbasi, et al.
Publicado: (2024) -
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
por: Kuzborskij, Ilja, et al.
Publicado: (2025) -
Low-rank bias, weight decay, and model merging in neural networks
por: Kuzborskij, Ilja, et al.
Publicado: (2025) -
Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence
por: György, András, et al.
Publicado: (2025) -
To See the Unseen: on the Generalization Ability of Transformers in Symbolic Reasoning
por: Lazić, Nevena, et al.
Publicado: (2026)