Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory
Fuente:
arXiv
Saved in:
| Main Author: | Cacioli, Jon-Paul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Lightweight Approach to Detection of AI-Generated Texts Using Stylometric Features
by: Aityan, Sergey K., et al.
Published: (2025)
by: Aityan, Sergey K., et al.
Published: (2025)
LLMs as Signal Detectors: Sensitivity, Bias, and the Temperature-Criterion Analogy
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
When is dataset cartography ineffective? Using training dynamics does not improve robustness against Adversarial SQuAD
by: Mandal, Paul K.
Published: (2025)
by: Mandal, Paul K.
Published: (2025)
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
by: Banthia, Saumya, et al.
Published: (2020)
by: Banthia, Saumya, et al.
Published: (2020)
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
by: Okpala, Izunna, et al.
Published: (2023)
by: Okpala, Izunna, et al.
Published: (2023)
DROID: Dual Representation for Out-of-Scope Intent Detection
by: Rashwan, Wael, et al.
Published: (2025)
by: Rashwan, Wael, et al.
Published: (2025)
Below-Chance Blindness: Prompted Underperformance in Small LLMs Produces Positional Bias Rather than Answer Avoidance
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Beyond the Mean: Within-Model Reliable Change Detection for LLM Evaluation
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Distilling Self-Consistency into Verbal Confidence: A Pre-Registered Negative Result and Post-Hoc Rescue on Gemma 3 4B
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Exemplar Retrieval Without Overhypothesis Induction: Limits of Distributional Sequence Learning in Early Word Learning
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
A Multi-Encoder Frozen-Decoder Approach for Fine-Tuning Large Language Models
by: Dhole, Kaustubh D.
Published: (2025)
by: Dhole, Kaustubh D.
Published: (2025)
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs
by: Asanuma, Haruka, et al.
Published: (2025)
by: Asanuma, Haruka, et al.
Published: (2025)
MIMIC-SR-ICD11: A Dataset for Narrative-Based Diagnosis
by: Wu, Yuexin, et al.
Published: (2025)
by: Wu, Yuexin, et al.
Published: (2025)
Applied Explainability for Large Language Models: A Comparative Study
by: Kancharla, Venkata Abhinandan
Published: (2026)
by: Kancharla, Venkata Abhinandan
Published: (2026)
Instruction Complexity Induces Positional Collapse in Adversarial LLM Evaluation
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Weber's Law in Transformer Magnitude Representations: Efficient Coding, Representational Geometry, and Psychophysical Laws in Language Models
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Option-Order Randomisation Reveals a Distributional Position Attractor in Prompted Sandbagging
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Efficient Hate Speech Detection: Evaluating 38 Models from Traditional Methods to Transformers
by: Abusaqer, Mahmoud, et al.
Published: (2025)
by: Abusaqer, Mahmoud, et al.
Published: (2025)
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
by: Sharma, Anika, et al.
Published: (2025)
by: Sharma, Anika, et al.
Published: (2025)
K-Way Energy Probes for Metacognition Reduce to Softmax in Discriminative Predictive Coding Networks
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Measuring Intent Comprehension in LLMs
by: Kunievsky, Nadav, et al.
Published: (2025)
by: Kunievsky, Nadav, et al.
Published: (2025)
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
by: Yocam, Eric, et al.
Published: (2026)
by: Yocam, Eric, et al.
Published: (2026)
Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
by: Reddy, Sandeep, et al.
Published: (2025)
by: Reddy, Sandeep, et al.
Published: (2025)
Transparent but Powerful: Explainability, Accuracy, and Generalizability in ADHD Detection from Social Media Data
by: Wiechmann, D., et al.
Published: (2024)
by: Wiechmann, D., et al.
Published: (2024)
Less is More: Learning Graph Tasks with Just LLMs
by: Shirai, Sola, et al.
Published: (2025)
by: Shirai, Sola, et al.
Published: (2025)
Emotional Voice Messages (EMOVOME) database: emotion recognition in spontaneous voice messages
by: Zaragozá, Lucía Gómez, et al.
Published: (2024)
by: Zaragozá, Lucía Gómez, et al.
Published: (2024)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
by: Wen, Yuqiao, et al.
Published: (2025)
by: Wen, Yuqiao, et al.
Published: (2025)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
by: Wen, Yuqiao, et al.
Published: (2024)
by: Wen, Yuqiao, et al.
Published: (2024)
Accelerating Language Model Workflows with Prompt Choreography
by: Bai, TJ, et al.
Published: (2025)
by: Bai, TJ, et al.
Published: (2025)
Semantic Reconstruction of Adversarial Plagiarism: A Context-Aware Framework for Detecting and Restoring "Tortured Phrases" in Scientific Literature
by: Maiti, Agniva, et al.
Published: (2025)
by: Maiti, Agniva, et al.
Published: (2025)
Let Your Graph Do the Talking: Encoding Structured Data for LLMs
by: Perozzi, Bryan, et al.
Published: (2024)
by: Perozzi, Bryan, et al.
Published: (2024)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
by: Hua, Wenjie, et al.
Published: (2025)
by: Hua, Wenjie, et al.
Published: (2025)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
by: Bachar, Or, et al.
Published: (2026)
by: Bachar, Or, et al.
Published: (2026)
ADALog: Adaptive Unsupervised Anomaly detection in Logs with Self-attention Masked Language Model
by: Pospieszny, Przemek, et al.
Published: (2025)
by: Pospieszny, Przemek, et al.
Published: (2025)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
by: Schneider, Felix, et al.
Published: (2026)
by: Schneider, Felix, et al.
Published: (2026)
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
by: Gutiérrez, Julián Camilo Velandia
Published: (2025)
by: Gutiérrez, Julián Camilo Velandia
Published: (2025)
Predicting When to Trust Vision-Language Models for Spatial Reasoning
by: Imran, Muhammad, et al.
Published: (2026)
by: Imran, Muhammad, et al.
Published: (2026)
AI-generated podcasts: Synthetic Intimacy and Cultural Mistranslation in NotebookLM's Audio Overviews
by: Rettberg, Jill Walker
Published: (2025)
by: Rettberg, Jill Walker
Published: (2025)
Similar Items
-
A Lightweight Approach to Detection of AI-Generated Texts Using Stylometric Features
by: Aityan, Sergey K., et al.
Published: (2025) -
LLMs as Signal Detectors: Sensitivity, Bias, and the Temperature-Criterion Analogy
by: Cacioli, Jon-Paul
Published: (2026) -
When is dataset cartography ineffective? Using training dynamics does not improve robustness against Adversarial SQuAD
by: Mandal, Paul K.
Published: (2025) -
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
by: Banthia, Saumya, et al.
Published: (2020) -
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
by: Okpala, Izunna, et al.
Published: (2023)