Evaluating the Reliability of Self-Explanations in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Randl, Korbinian, Pavlopoulos, John, Henriksson, Aron, Lindgren, Tony |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CICLe: Conformal In-Context Learning for Largescale Multi-Class Food Risk Classification
by: Randl, Korbinian, et al.
Published: (2024)
by: Randl, Korbinian, et al.
Published: (2024)
SemEval-2025 Task 9: The Food Hazard Detection Challenge
by: Randl, Korbinian, et al.
Published: (2025)
by: Randl, Korbinian, et al.
Published: (2025)
RAG-E: Quantifying Retriever-Generator Alignment and Failure Modes
by: Randl, Korbinian, et al.
Published: (2026)
by: Randl, Korbinian, et al.
Published: (2026)
CICLe: Conformal In-Context Learning for Largescale Multi-Class Food Risk Classification
by: Randl, Korbinian, et al.
Published: (2024)
by: Randl, Korbinian, et al.
Published: (2024)
Efficient Text Classification with Conformal In-Context Learning
by: Pantelidis, Ippokratis, et al.
Published: (2025)
by: Pantelidis, Ippokratis, et al.
Published: (2025)
The Impact of Steering Large Language Models with Persona Vectors in Educational Applications
by: Wu, Yongchao, et al.
Published: (2026)
by: Wu, Yongchao, et al.
Published: (2026)
Dialect Normalization using Large Language Models and Morphological Rules
by: Dimakis, Antonios, et al.
Published: (2025)
by: Dimakis, Antonios, et al.
Published: (2025)
CHiL(L)Grader: Calibrated Human-in-the-Loop Short-Answer Grading
by: Raikote, Pranav, et al.
Published: (2026)
by: Raikote, Pranav, et al.
Published: (2026)
Data-Constrained Synthesis of Training Data for De-Identification
by: Vakili, Thomas, et al.
Published: (2025)
by: Vakili, Thomas, et al.
Published: (2025)
Proverbs or Pythian Oracles? Sentiments and Emotions in Greek Sayings
by: Korre, Katerina, et al.
Published: (2025)
by: Korre, Katerina, et al.
Published: (2025)
Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models
by: Agarwal, Chirag, et al.
Published: (2024)
by: Agarwal, Chirag, et al.
Published: (2024)
Large Language Models as Evaluators for Recommendation Explanations
by: Zhang, Xiaoyu, et al.
Published: (2024)
by: Zhang, Xiaoyu, et al.
Published: (2024)
A Systematic Survey of Natural Language Processing for the Greek Language
by: Bakagianni, Juli, et al.
Published: (2024)
by: Bakagianni, Juli, et al.
Published: (2024)
Investigating Training and Generalization in Faithful Self-Explanations of Large Language Models
by: Doi, Tomoki, et al.
Published: (2025)
by: Doi, Tomoki, et al.
Published: (2025)
TopClustRAG at SIGIR 2025 LiveRAG Challenge
by: Bakagianni, Juli, et al.
Published: (2025)
by: Bakagianni, Juli, et al.
Published: (2025)
xFinder: Large Language Models as Automated Evaluators for Reliable Evaluation
by: Yu, Qingchen, et al.
Published: (2024)
by: Yu, Qingchen, et al.
Published: (2024)
Are we chasing ghosts? Quantifying unattributable polarization, and attributing the rest to annotator groups
by: Tsirmpas, Dimitris, et al.
Published: (2026)
by: Tsirmpas, Dimitris, et al.
Published: (2026)
Leveraging LLMs for Translating and Classifying Mental Health Data
by: Skianis, Konstantinos, et al.
Published: (2024)
by: Skianis, Konstantinos, et al.
Published: (2024)
PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking
by: Zhu, Wang Bill, et al.
Published: (2026)
by: Zhu, Wang Bill, et al.
Published: (2026)
Cross-Examiner: Evaluating Consistency of Large Language Model-Generated Explanations
by: Villa, Danielle, et al.
Published: (2025)
by: Villa, Danielle, et al.
Published: (2025)
Revisiting the Reliability of Psychological Scales on Large Language Models
by: Huang, Jen-tse, et al.
Published: (2023)
by: Huang, Jen-tse, et al.
Published: (2023)
Open or Closed LLM for Lesser-Resourced Languages? Lessons from Greek
by: Pavlopoulos, John, et al.
Published: (2025)
by: Pavlopoulos, John, et al.
Published: (2025)
Automating Dataset Updates Towards Reliable and Timely Evaluation of Large Language Models
by: Ying, Jiahao, et al.
Published: (2024)
by: Ying, Jiahao, et al.
Published: (2024)
Exploring the Reliability of Large Language Models as Customized Evaluators for Diverse NLP Tasks
by: Li, Qintong, et al.
Published: (2023)
by: Li, Qintong, et al.
Published: (2023)
No Reliable Evidence of Self-Reported Sentience in Small Large Language Models
by: Kaiser, Caspar, et al.
Published: (2026)
by: Kaiser, Caspar, et al.
Published: (2026)
ReliableMath: Benchmark of Reliable Mathematical Reasoning on Large Language Models
by: Xue, Boyang, et al.
Published: (2025)
by: Xue, Boyang, et al.
Published: (2025)
Building Multilingual Datasets for Predicting Mental Health Severity through LLMs: Prospects and Challenges
by: Skianis, Konstantinos, et al.
Published: (2024)
by: Skianis, Konstantinos, et al.
Published: (2024)
On the Reliability of Large Language Models for Causal Discovery
by: Feng, Tao, et al.
Published: (2024)
by: Feng, Tao, et al.
Published: (2024)
Evaluation Framework for Highlight Explanations of Context Utilisation in Language Models
by: Sun, Jingyi, et al.
Published: (2025)
by: Sun, Jingyi, et al.
Published: (2025)
On the Reliability of Watermarks for Large Language Models
by: Kirchenbauer, John, et al.
Published: (2023)
by: Kirchenbauer, John, et al.
Published: (2023)
Hate Speech According to the Law: An Analysis for Effective Detection
by: Korre, Katerina, et al.
Published: (2024)
by: Korre, Katerina, et al.
Published: (2024)
Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations
by: Matton, Katie, et al.
Published: (2025)
by: Matton, Katie, et al.
Published: (2025)
Iterative Self-Improvement of Vision Language Models for Image Scoring and Self-Explanation
by: Tanji, Naoto, et al.
Published: (2025)
by: Tanji, Naoto, et al.
Published: (2025)
Assessing the Reliability of LLMs Annotations in the Context of Demographic Bias and Model Explanation
by: Mohammadi, Hadi, et al.
Published: (2025)
by: Mohammadi, Hadi, et al.
Published: (2025)
AA-Omniscience: Evaluating Cross-Domain Knowledge Reliability in Large Language Models
by: Jackson, Declan, et al.
Published: (2025)
by: Jackson, Declan, et al.
Published: (2025)
Enhancing Large Language Models with Reliable Knowledge Graphs
by: Zhang, Qinggang
Published: (2025)
by: Zhang, Qinggang
Published: (2025)
Register Always Matters: Analysis of LLM Pretraining Data Through the Lens of Language Variation
by: Myntti, Amanda, et al.
Published: (2025)
by: Myntti, Amanda, et al.
Published: (2025)
ELI-Why: Evaluating the Pedagogical Utility of Language Model Explanations
by: Joshi, Brihi, et al.
Published: (2025)
by: Joshi, Brihi, et al.
Published: (2025)
Self-AMPLIFY: Improving Small Language Models with Self Post Hoc Explanations
by: Bhan, Milan, et al.
Published: (2024)
by: Bhan, Milan, et al.
Published: (2024)
How Interpretable are Reasoning Explanations from Prompting Large Language Models?
by: Yeo, Wei Jie, et al.
Published: (2024)
by: Yeo, Wei Jie, et al.
Published: (2024)
Similar Items
-
CICLe: Conformal In-Context Learning for Largescale Multi-Class Food Risk Classification
by: Randl, Korbinian, et al.
Published: (2024) -
SemEval-2025 Task 9: The Food Hazard Detection Challenge
by: Randl, Korbinian, et al.
Published: (2025) -
RAG-E: Quantifying Retriever-Generator Alignment and Failure Modes
by: Randl, Korbinian, et al.
Published: (2026) -
CICLe: Conformal In-Context Learning for Largescale Multi-Class Food Risk Classification
by: Randl, Korbinian, et al.
Published: (2024) -
Efficient Text Classification with Conformal In-Context Learning
by: Pantelidis, Ippokratis, et al.
Published: (2025)