To Believe or Not to Believe Your LLM
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yadkori, Yasin Abbasi, Kuzborskij, Ilja, György, András, Szepesvári, Csaba |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mitigating LLM Hallucinations via Conformal Abstention
von: Yadkori, Yasin Abbasi, et al.
Veröffentlicht: (2024)
von: Yadkori, Yasin Abbasi, et al.
Veröffentlicht: (2024)
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
von: Kuzborskij, Ilja, et al.
Veröffentlicht: (2025)
von: Kuzborskij, Ilja, et al.
Veröffentlicht: (2025)
Low-rank bias, weight decay, and model merging in neural networks
von: Kuzborskij, Ilja, et al.
Veröffentlicht: (2025)
von: Kuzborskij, Ilja, et al.
Veröffentlicht: (2025)
Frontier LLMs Still Struggle with Simple Reasoning Tasks
von: Malek, Alan, et al.
Veröffentlicht: (2025)
von: Malek, Alan, et al.
Veröffentlicht: (2025)
Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence
von: György, András, et al.
Veröffentlicht: (2025)
von: György, András, et al.
Veröffentlicht: (2025)
Believe It or Not: How Deeply do LLMs Believe Implanted Facts?
von: Slocum, Stewart, et al.
Veröffentlicht: (2025)
von: Slocum, Stewart, et al.
Veröffentlicht: (2025)
Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding
von: Deng, Ailin, et al.
Veröffentlicht: (2024)
von: Deng, Ailin, et al.
Veröffentlicht: (2024)
Balancing optimism and pessimism in offline-to-online learning
von: Sentenac, Flore, et al.
Veröffentlicht: (2025)
von: Sentenac, Flore, et al.
Veröffentlicht: (2025)
See Me and Believe Me: Causality and Intersectionality in Testimonial Injustice in Healthcare
von: Andrews, Kenya S., et al.
Veröffentlicht: (2024)
von: Andrews, Kenya S., et al.
Veröffentlicht: (2024)
To See the Unseen: on the Generalization Ability of Transformers in Symbolic Reasoning
von: Lazić, Nevena, et al.
Veröffentlicht: (2026)
von: Lazić, Nevena, et al.
Veröffentlicht: (2026)
Check Your LLM's Secret Dictionary! Five Lines of Code Reveal What Your LLM Learned (Including What It Shouldn't Have)
von: Miyashita, Hisashi
Veröffentlicht: (2026)
von: Miyashita, Hisashi
Veröffentlicht: (2026)
Learning from Relevant Subgoals in Successful Dialogs using Iterative Training for Task-oriented Dialog Systems
von: Kaiser, Magdalena, et al.
Veröffentlicht: (2024)
von: Kaiser, Magdalena, et al.
Veröffentlicht: (2024)
DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs
von: Zhang, Yuanhe, et al.
Veröffentlicht: (2025)
von: Zhang, Yuanhe, et al.
Veröffentlicht: (2025)
Maximize Your Data's Potential: Enhancing LLM Accuracy with Two-Phase Pretraining
von: Feng, Steven, et al.
Veröffentlicht: (2024)
von: Feng, Steven, et al.
Veröffentlicht: (2024)
Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents
von: Shao, Shuai, et al.
Veröffentlicht: (2025)
von: Shao, Shuai, et al.
Veröffentlicht: (2025)
Adversarial Preference Optimization: Enhancing Your Alignment via RM-LLM Game
von: Cheng, Pengyu, et al.
Veröffentlicht: (2023)
von: Cheng, Pengyu, et al.
Veröffentlicht: (2023)
What is Your Data Worth to GPT? LLM-Scale Data Valuation with Influence Functions
von: Choe, Sang Keun, et al.
Veröffentlicht: (2024)
von: Choe, Sang Keun, et al.
Veröffentlicht: (2024)
Seeing Isn't Believing: Mitigating Belief Inertia via Active Intervention in Embodied Agents
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)
von: Wang, Hanlin, et al.
Veröffentlicht: (2026)
LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
von: Zhang, Xuan, et al.
Veröffentlicht: (2024)
I Can't Believe It's Not Robust: Catastrophic Collapse of Safety Classifiers under Embedding Drift
von: Sahoo, Subramanyam, et al.
Veröffentlicht: (2026)
von: Sahoo, Subramanyam, et al.
Veröffentlicht: (2026)
How Learning Rate Decay Wastes Your Best Data in Curriculum-Based LLM Pretraining
von: Luo, Kairong, et al.
Veröffentlicht: (2025)
von: Luo, Kairong, et al.
Veröffentlicht: (2025)
MIR-Bench: Can Your LLM Recognize Complicated Patterns via Many-Shot In-Context Reasoning?
von: Yan, Kai, et al.
Veröffentlicht: (2025)
von: Yan, Kai, et al.
Veröffentlicht: (2025)
I Can't Believe It's Not Real: CV-MuSeNet: Complex-Valued Multi-Signal Segmentation
von: Shin, Sangwon, et al.
Veröffentlicht: (2025)
von: Shin, Sangwon, et al.
Veröffentlicht: (2025)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
von: Harzli, Ouns El, et al.
Veröffentlicht: (2026)
von: Harzli, Ouns El, et al.
Veröffentlicht: (2026)
Your Transformer is Secretly Linear
von: Razzhigaev, Anton, et al.
Veröffentlicht: (2024)
von: Razzhigaev, Anton, et al.
Veröffentlicht: (2024)
Aligning Human and Machine Attention for Enhanced Supervised Learning
von: Chriqui, Avihay, et al.
Veröffentlicht: (2025)
von: Chriqui, Avihay, et al.
Veröffentlicht: (2025)
CHILL at SemEval-2025 Task 2: You Can't Just Throw Entities and Hope -- Make Your LLM to Get Them Right
von: Lee, Jaebok, et al.
Veröffentlicht: (2025)
von: Lee, Jaebok, et al.
Veröffentlicht: (2025)
Don't Believe Everything You Read: Enhancing Summarization Interpretability through Automatic Identification of Hallucinations in Large Language Models
von: Vakharia, Priyesh, et al.
Veröffentlicht: (2023)
von: Vakharia, Priyesh, et al.
Veröffentlicht: (2023)
Hierarchical Reasoning Model
von: Wang, Guan, et al.
Veröffentlicht: (2025)
von: Wang, Guan, et al.
Veröffentlicht: (2025)
Investigating the performance of Retrieval-Augmented Generation and fine-tuning for the development of AI-driven knowledge-based systems
von: Lakatos, Robert, et al.
Veröffentlicht: (2024)
von: Lakatos, Robert, et al.
Veröffentlicht: (2024)
Not All Synthetic Data Is Yours to Learn From
von: Alemohammad, Sina, et al.
Veröffentlicht: (2026)
von: Alemohammad, Sina, et al.
Veröffentlicht: (2026)
Watch Your Steps: Observable and Modular Chains of Thought
von: Cohen, Cassandra A., et al.
Veröffentlicht: (2024)
von: Cohen, Cassandra A., et al.
Veröffentlicht: (2024)
Intent Factored Generation: Unleashing the Diversity in Your Language Model
von: Ahmed, Eltayeb, et al.
Veröffentlicht: (2025)
von: Ahmed, Eltayeb, et al.
Veröffentlicht: (2025)
Your Absorbing Discrete Diffusion Secretly Models the Bayesian Posterior
von: Doyle, Cooper
Veröffentlicht: (2025)
von: Doyle, Cooper
Veröffentlicht: (2025)
TalkWithMachines: Enhancing Human-Robot Interaction for Interpretable Industrial Robotics Through Large/Vision Language Models
von: Abbas, Ammar N., et al.
Veröffentlicht: (2024)
von: Abbas, Ammar N., et al.
Veröffentlicht: (2024)
Specialised or Generic? Tokenization Choices for Radiology Language Models
von: Warr, Hermione, et al.
Veröffentlicht: (2025)
von: Warr, Hermione, et al.
Veröffentlicht: (2025)
Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025)
von: Wang, Yinong Oliver, et al.
Veröffentlicht: (2025)
Reasoning with Sampling: Your Base Model is Smarter Than You Think
von: Karan, Aayush, et al.
Veröffentlicht: (2025)
von: Karan, Aayush, et al.
Veröffentlicht: (2025)
Should You Use Your Large Language Model to Explore or Exploit?
von: Harris, Keegan, et al.
Veröffentlicht: (2025)
von: Harris, Keegan, et al.
Veröffentlicht: (2025)
Your thoughts tell who you are: Characterize the reasoning patterns of LRMs
von: Chen, Yida, et al.
Veröffentlicht: (2025)
von: Chen, Yida, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mitigating LLM Hallucinations via Conformal Abstention
von: Yadkori, Yasin Abbasi, et al.
Veröffentlicht: (2024) -
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
von: Kuzborskij, Ilja, et al.
Veröffentlicht: (2025) -
Low-rank bias, weight decay, and model merging in neural networks
von: Kuzborskij, Ilja, et al.
Veröffentlicht: (2025) -
Frontier LLMs Still Struggle with Simple Reasoning Tasks
von: Malek, Alan, et al.
Veröffentlicht: (2025) -
Beyond Statistical Learning: Exact Learning Is Essential for General Intelligence
von: György, András, et al.
Veröffentlicht: (2025)