Do Large Language Models Know How Much They Know?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Prato, Gabriele, Huang, Jerry, Parthasarathi, Prasanna, Sodhani, Shagun, Chandar, Sarath |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Effect of Document Packing on the Latent Multi-Hop Reasoning Capabilities of Large Language Models
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
von: Prato, Gabriele, et al.
Veröffentlicht: (2025)
EpiK-Eval: Evaluation for Language Models as Epistemic Models
von: Prato, Gabriele, et al.
Veröffentlicht: (2023)
von: Prato, Gabriele, et al.
Veröffentlicht: (2023)
Towards Practical Tool Usage for Continually Learning LLMs
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
Context-Aware Assistant Selection for Improved Inference Acceleration with Large Language Models
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
von: Huang, Jerry, et al.
Veröffentlicht: (2024)
BertaQA: How Much Do Language Models Know About Local Culture?
von: Etxaniz, Julen, et al.
Veröffentlicht: (2024)
von: Etxaniz, Julen, et al.
Veröffentlicht: (2024)
Are self-explanations from Large Language Models faithful?
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
von: Madsen, Andreas, et al.
Veröffentlicht: (2024)
NanoKnow: How to Know What Your Language Model Knows
von: Gu, Lingwei, et al.
Veröffentlicht: (2026)
von: Gu, Lingwei, et al.
Veröffentlicht: (2026)
Large Language Models Must Be Taught to Know What They Don't Know
von: Kapoor, Sanyam, et al.
Veröffentlicht: (2024)
von: Kapoor, Sanyam, et al.
Veröffentlicht: (2024)
Small Encoders Can Rival Large Decoders in Detecting Groundedness
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
What Large Language Models Know and What People Think They Know
von: Steyvers, Mark, et al.
Veröffentlicht: (2024)
von: Steyvers, Mark, et al.
Veröffentlicht: (2024)
Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models
von: Ferrando, Javier, et al.
Veröffentlicht: (2024)
von: Ferrando, Javier, et al.
Veröffentlicht: (2024)
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
von: Heuillet, Maxime, et al.
Veröffentlicht: (2025)
von: Heuillet, Maxime, et al.
Veröffentlicht: (2025)
GRPO-$λ$: Credit Assignment improves LLM Reasoning
von: Parthasarathi, Prasanna, et al.
Veröffentlicht: (2025)
von: Parthasarathi, Prasanna, et al.
Veröffentlicht: (2025)
What Do Large Language Models Know? Tacit Knowledge as a Potential Causal-Explanatory Structure
von: Budding, Céline
Veröffentlicht: (2025)
von: Budding, Céline
Veröffentlicht: (2025)
CaRT: Teaching LLM Agents to Know When They Know Enough
von: Liu, Grace, et al.
Veröffentlicht: (2025)
von: Liu, Grace, et al.
Veröffentlicht: (2025)
Revisiting Replay and Gradient Alignment for Continual Pre-Training of Large Language Models
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
von: Abbes, Istabrak, et al.
Veröffentlicht: (2025)
Too Big to Fool: Resisting Deception in Language Models
von: Samsami, Mohammad Reza, et al.
Veröffentlicht: (2024)
von: Samsami, Mohammad Reza, et al.
Veröffentlicht: (2024)
Do Androids Know They're Only Dreaming of Electric Sheep?
von: CH-Wang, Sky, et al.
Veröffentlicht: (2023)
von: CH-Wang, Sky, et al.
Veröffentlicht: (2023)
Do Retrieval Augmented Language Models Know When They Don't Know?
von: Zhou, Youchao, et al.
Veröffentlicht: (2025)
von: Zhou, Youchao, et al.
Veröffentlicht: (2025)
Why Don't Prompt-Based Fairness Metrics Correlate?
von: Zayed, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Zayed, Abdelrahman, et al.
Veröffentlicht: (2024)
Should We Attend More or Less? Modulating Attention for Fairness
von: Zayed, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Zayed, Abdelrahman, et al.
Veröffentlicht: (2023)
Text Knows What, Tables Know When: Clinical Timeline Reconstruction via Retrieval-Augmented Multimodal Alignment
von: Kumar, Sayantan, et al.
Veröffentlicht: (2026)
von: Kumar, Sayantan, et al.
Veröffentlicht: (2026)
Do Large Language Models Know Conflict? Investigating Parametric vs. Non-Parametric Knowledge of LLMs for Conflict Forecasting
von: Nemkova, Apollinaire Poli, et al.
Veröffentlicht: (2025)
von: Nemkova, Apollinaire Poli, et al.
Veröffentlicht: (2025)
Characterizing Large Language Model Geometry Helps Solve Toxicity Detection and Generation
von: Balestriero, Randall, et al.
Veröffentlicht: (2023)
von: Balestriero, Randall, et al.
Veröffentlicht: (2023)
Do Large Language Models Know What They Are Capable Of?
von: Barkan, Casey O., et al.
Veröffentlicht: (2025)
von: Barkan, Casey O., et al.
Veröffentlicht: (2025)
How Well Can a Long Sequence Model Model Long Sequences? Comparing Architechtural Inductive Biases on Long-Context Abilities
von: Huang, Jerry
Veröffentlicht: (2024)
von: Huang, Jerry
Veröffentlicht: (2024)
KnowPhish: Large Language Models Meet Multimodal Knowledge Graphs for Enhancing Reference-Based Phishing Detection
von: Li, Yuexin, et al.
Veröffentlicht: (2024)
von: Li, Yuexin, et al.
Veröffentlicht: (2024)
TIDE: Every Layer Knows the Token Beneath the Context
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2026)
von: Jaiswal, Ajay, et al.
Veröffentlicht: (2026)
When Models Know More Than They Say: Probing Analogical Reasoning in LLMs
von: McGovern, Hope, et al.
Veröffentlicht: (2026)
von: McGovern, Hope, et al.
Veröffentlicht: (2026)
KnowRL: Teaching Language Models to Know What They Know
von: Kale, Sahil, et al.
Veröffentlicht: (2025)
von: Kale, Sahil, et al.
Veröffentlicht: (2025)
FaKnow: A Unified Library for Fake News Detection
von: Zhu, Yiyuan, et al.
Veröffentlicht: (2024)
von: Zhu, Yiyuan, et al.
Veröffentlicht: (2024)
How Much Do Large Language Models Know about Human Motion? A Case Study in 3D Avatar Control
von: Li, Kunhang, et al.
Veröffentlicht: (2025)
von: Li, Kunhang, et al.
Veröffentlicht: (2025)
The Markovian Thinker: Architecture-Agnostic Linear Scaling of Reasoning
von: Aghajohari, Milad, et al.
Veröffentlicht: (2025)
von: Aghajohari, Milad, et al.
Veröffentlicht: (2025)
How Much is Too Much? Exploring LoRA Rank Trade-offs for Retaining Knowledge and Domain Robustness
von: Rathore, Darshita, et al.
Veröffentlicht: (2025)
von: Rathore, Darshita, et al.
Veröffentlicht: (2025)
KnowCoder-X: Boosting Multilingual Information Extraction via Code
von: Zuo, Yuxin, et al.
Veröffentlicht: (2024)
von: Zuo, Yuxin, et al.
Veröffentlicht: (2024)
Know Thyself? On the Incapability and Implications of AI Self-Recognition
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2025)
von: Bai, Xiaoyan, et al.
Veröffentlicht: (2025)
MaestroMotif: Skill Design from Artificial Intelligence Feedback
von: Klissarov, Martin, et al.
Veröffentlicht: (2024)
von: Klissarov, Martin, et al.
Veröffentlicht: (2024)
Know Your Limits: Entropy Estimation Modeling for Compression and Generalization
von: Badger, Benjamin L., et al.
Veröffentlicht: (2025)
von: Badger, Benjamin L., et al.
Veröffentlicht: (2025)
(How) Do Language Models Track State?
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Effect of Document Packing on the Latent Multi-Hop Reasoning Capabilities of Large Language Models
von: Prato, Gabriele, et al.
Veröffentlicht: (2025) -
EpiK-Eval: Evaluation for Language Models as Epistemic Models
von: Prato, Gabriele, et al.
Veröffentlicht: (2023) -
Towards Practical Tool Usage for Continually Learning LLMs
von: Huang, Jerry, et al.
Veröffentlicht: (2024) -
Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination
von: Huang, Jerry, et al.
Veröffentlicht: (2024) -
Context-Aware Assistant Selection for Improved Inference Acceleration with Large Language Models
von: Huang, Jerry, et al.
Veröffentlicht: (2024)