Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Taghanaki, Saeid Asgari, Monteiro, Joao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MMLU-Pro+: Evaluating Higher-Order Reasoning and Shortcut Learning in LLMs
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2024)
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2024)
Detecting Generative Parroting through Overfitting Masked Autoencoders
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2024)
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2024)
LLMs for XAI: Future Directions for Explaining Explanations
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024)
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024)
TExplain: Explaining Learned Visual Features via Pre-trained (Frozen) Language Models
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2023)
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2023)
SLiMe: Segment Like Me
von: Khani, Aliasghar, et al.
Veröffentlicht: (2023)
von: Khani, Aliasghar, et al.
Veröffentlicht: (2023)
Query-Focused Extractive Summarization for Sentiment Explanation
von: Moubtahij, Ahmed, et al.
Veröffentlicht: (2025)
von: Moubtahij, Ahmed, et al.
Veröffentlicht: (2025)
Selective Explanations
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2024)
von: Paes, Lucas Monteiro, et al.
Veröffentlicht: (2024)
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
von: Wang, Qianli, et al.
Veröffentlicht: (2026)
Disentangled PET Lesion Segmentation
von: Gatsak, Tanya, et al.
Veröffentlicht: (2024)
von: Gatsak, Tanya, et al.
Veröffentlicht: (2024)
A Comprehensive Evaluation of Neural SPARQL Query Generation from Natural Language Questions
von: Diallo, Papa Abdou Karim Karou, et al.
Veröffentlicht: (2023)
von: Diallo, Papa Abdou Karim Karou, et al.
Veröffentlicht: (2023)
Deep Semantic Segmentation of Natural and Medical Images: A Review
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2019)
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2019)
Query Performance Explanation through Large Language Model for HTAP Systems
von: Xiu, Haibo, et al.
Veröffentlicht: (2024)
von: Xiu, Haibo, et al.
Veröffentlicht: (2024)
Query-Conditioned Test-Time Self-Training for Large Language Models
von: Song, Chaehee, et al.
Veröffentlicht: (2026)
von: Song, Chaehee, et al.
Veröffentlicht: (2026)
LLMs Explain't: A Post-Mortem on Semantic Interpretability in Transformer Models
von: Abdelhalim, Alhassan, et al.
Veröffentlicht: (2026)
von: Abdelhalim, Alhassan, et al.
Veröffentlicht: (2026)
LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
von: Mayne, Harry, et al.
Veröffentlicht: (2025)
von: Mayne, Harry, et al.
Veröffentlicht: (2025)
Explaining Length Bias in LLM-Based Preference Evaluations
von: Hu, Zhengyu, et al.
Veröffentlicht: (2024)
von: Hu, Zhengyu, et al.
Veröffentlicht: (2024)
Self-AMPLIFY: Improving Small Language Models with Self Post Hoc Explanations
von: Bhan, Milan, et al.
Veröffentlicht: (2024)
von: Bhan, Milan, et al.
Veröffentlicht: (2024)
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
von: Eilertsen, Brage, et al.
Veröffentlicht: (2025)
von: Eilertsen, Brage, et al.
Veröffentlicht: (2025)
Harnessing LLMs Explanations to Boost Surrogate Models in Tabular Data Classification
von: Shi, Ruxue, et al.
Veröffentlicht: (2025)
von: Shi, Ruxue, et al.
Veröffentlicht: (2025)
Latent Logic Tree Extraction for Event Sequence Explanation from LLMs
von: Song, Zitao, et al.
Veröffentlicht: (2024)
von: Song, Zitao, et al.
Veröffentlicht: (2024)
A Comprehensive Evaluation framework of Alignment Techniques for LLMs
von: Azmat, Muneeza, et al.
Veröffentlicht: (2025)
von: Azmat, Muneeza, et al.
Veröffentlicht: (2025)
GRILE: A Benchmark for Grammar Reasoning and Explanation in Romanian LLMs
von: Dumitran, Adrian-Marius, et al.
Veröffentlicht: (2025)
von: Dumitran, Adrian-Marius, et al.
Veröffentlicht: (2025)
Predicting the Performance of Black-box LLMs through Follow-up Queries
von: Sam, Dylan, et al.
Veröffentlicht: (2025)
von: Sam, Dylan, et al.
Veröffentlicht: (2025)
LLMs Are Not Intelligent Thinkers: Introducing Mathematical Topic Tree Benchmark for Comprehensive Evaluation of LLMs
von: Davoodi, Arash Gholami, et al.
Veröffentlicht: (2024)
von: Davoodi, Arash Gholami, et al.
Veröffentlicht: (2024)
Evaluating LLMs for Hardware Design and Test
von: Blocklove, Jason, et al.
Veröffentlicht: (2024)
von: Blocklove, Jason, et al.
Veröffentlicht: (2024)
Evaluating LLMs for Demographic-Targeted Social Bias Detection: A Comprehensive Benchmark Study
von: Majumdar, Ayan, et al.
Veröffentlicht: (2025)
von: Majumdar, Ayan, et al.
Veröffentlicht: (2025)
Towards Universal and Black-Box Query-Response Only Attack on LLMs with QROA
von: Jawad, Hussein, et al.
Veröffentlicht: (2024)
von: Jawad, Hussein, et al.
Veröffentlicht: (2024)
Shared Lexical Task Representations Explain Behavioral Variability In LLMs
von: Yang, Zhuonan, et al.
Veröffentlicht: (2026)
von: Yang, Zhuonan, et al.
Veröffentlicht: (2026)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
von: Kroeger, Nicholas, et al.
Veröffentlicht: (2023)
von: Kroeger, Nicholas, et al.
Veröffentlicht: (2023)
Word Sense Detection Leveraging Maximum Mean Discrepancy
von: Mitsuzawa, Kensuke
Veröffentlicht: (2025)
von: Mitsuzawa, Kensuke
Veröffentlicht: (2025)
ExPO: Unlocking Hard Reasoning with Self-Explanation-Guided Reinforcement Learning
von: Zhou, Ruiyang, et al.
Veröffentlicht: (2025)
von: Zhou, Ruiyang, et al.
Veröffentlicht: (2025)
Resolving Discrepancies in Compute-Optimal Scaling of Language Models
von: Porian, Tomer, et al.
Veröffentlicht: (2024)
von: Porian, Tomer, et al.
Veröffentlicht: (2024)
Chain of Simulation: A Dual-Mode Reasoning Framework for Large Language Models with Dynamic Problem Routing
von: Sheikhi, Saeid
Veröffentlicht: (2026)
von: Sheikhi, Saeid
Veröffentlicht: (2026)
CFSafety: Comprehensive Fine-grained Safety Assessment for LLMs
von: Liu, Zhihao, et al.
Veröffentlicht: (2024)
von: Liu, Zhihao, et al.
Veröffentlicht: (2024)
An Evaluation of Explanation Methods for Black-Box Detectors of Machine-Generated Text
von: Schoenegger, Loris, et al.
Veröffentlicht: (2024)
von: Schoenegger, Loris, et al.
Veröffentlicht: (2024)
ADAM: A Diverse Archive of Mankind for Evaluating and Enhancing LLMs in Biographical Reasoning
von: Cekinmez, Jasin, et al.
Veröffentlicht: (2025)
von: Cekinmez, Jasin, et al.
Veröffentlicht: (2025)
MMD-Flagger: Leveraging Maximum Mean Discrepancy to Detect Hallucinations
von: Mitsuzawa, Kensuke, et al.
Veröffentlicht: (2025)
von: Mitsuzawa, Kensuke, et al.
Veröffentlicht: (2025)
Sample-Efficient Human Evaluation of Large Language Models via Maximum Discrepancy Competition
von: Feng, Kehua, et al.
Veröffentlicht: (2024)
von: Feng, Kehua, et al.
Veröffentlicht: (2024)
Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees
von: Montreuil, Yannis, et al.
Veröffentlicht: (2024)
von: Montreuil, Yannis, et al.
Veröffentlicht: (2024)
SoftQE: Learned Representations of Queries Expanded by LLMs
von: Pimpalkhute, Varad, et al.
Veröffentlicht: (2024)
von: Pimpalkhute, Varad, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MMLU-Pro+: Evaluating Higher-Order Reasoning and Shortcut Learning in LLMs
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2024) -
Detecting Generative Parroting through Overfitting Masked Autoencoders
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2024) -
LLMs for XAI: Future Directions for Explaining Explanations
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024) -
TExplain: Explaining Learned Visual Features via Pre-trained (Frozen) Language Models
von: Taghanaki, Saeid Asgari, et al.
Veröffentlicht: (2023) -
SLiMe: Segment Like Me
von: Khani, Aliasghar, et al.
Veröffentlicht: (2023)