Digital Socrates: Evaluating LLMs through Explanation Critiques
Fuente:
arXiv
Salvato in:
| Autori principali: | Gu, Yuling, Tafjord, Oyvind, Clark, Peter |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SimpleToM: Exposing the Gap between Explicit ToM Inference and Implicit ToM Application in LLMs
di: Gu, Yuling, et al.
Pubblicazione: (2024)
di: Gu, Yuling, et al.
Pubblicazione: (2024)
BaRDa: A Belief and Reasoning Dataset that Separates Factual Accuracy and Reasoning Ability
di: Clark, Peter, et al.
Pubblicazione: (2023)
di: Clark, Peter, et al.
Pubblicazione: (2023)
OLMES: A Standard for Language Model Evaluations
di: Gu, Yuling, et al.
Pubblicazione: (2024)
di: Gu, Yuling, et al.
Pubblicazione: (2024)
DISCOVERYWORLD: A Virtual Environment for Developing and Evaluating Automated Scientific Discovery Agents
di: Jansen, Peter, et al.
Pubblicazione: (2024)
di: Jansen, Peter, et al.
Pubblicazione: (2024)
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
di: Wiegreffe, Sarah, et al.
Pubblicazione: (2024)
di: Wiegreffe, Sarah, et al.
Pubblicazione: (2024)
CodeScientist: End-to-End Semi-Automated Scientific Discovery with Code-based Experimentation
di: Jansen, Peter, et al.
Pubblicazione: (2025)
di: Jansen, Peter, et al.
Pubblicazione: (2025)
Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
di: Weir, Nathaniel, et al.
Pubblicazione: (2024)
di: Weir, Nathaniel, et al.
Pubblicazione: (2024)
Discerning minds or generic tutors? Evaluating instructional guidance capabilities in Socratic LLMs
di: Liu, Ying, et al.
Pubblicazione: (2025)
di: Liu, Ying, et al.
Pubblicazione: (2025)
The Critique of Critique
di: Sun, Shichao, et al.
Pubblicazione: (2024)
di: Sun, Shichao, et al.
Pubblicazione: (2024)
Critique-RL: Training Language Models for Critiquing through Two-Stage Reinforcement Learning
di: Xi, Zhiheng, et al.
Pubblicazione: (2025)
di: Xi, Zhiheng, et al.
Pubblicazione: (2025)
CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
di: Ke, Pei, et al.
Pubblicazione: (2023)
di: Ke, Pei, et al.
Pubblicazione: (2023)
PreScience: A Benchmark for Forecasting Scientific Contributions
di: Ajith, Anirudh, et al.
Pubblicazione: (2026)
di: Ajith, Anirudh, et al.
Pubblicazione: (2026)
Local Explanations and Self-Explanations for Assessing Faithfulness in black-box LLMs
di: Fragkathoulas, Christos, et al.
Pubblicazione: (2024)
di: Fragkathoulas, Christos, et al.
Pubblicazione: (2024)
SocREval: Large Language Models with the Socratic Method for Reference-Free Reasoning Evaluation
di: He, Hangfeng, et al.
Pubblicazione: (2023)
di: He, Hangfeng, et al.
Pubblicazione: (2023)
Boundless Socratic Learning with Language Games
di: Schaul, Tom
Pubblicazione: (2024)
di: Schaul, Tom
Pubblicazione: (2024)
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
di: Lin, Zicheng, et al.
Pubblicazione: (2024)
di: Lin, Zicheng, et al.
Pubblicazione: (2024)
Retrieval is Not Enough: Enhancing RAG Reasoning through Test-Time Critique and Optimization
di: Wei, Jiaqi, et al.
Pubblicazione: (2025)
di: Wei, Jiaqi, et al.
Pubblicazione: (2025)
From Critique to Clarity: A Pathway to Faithful and Personalized Code Explanations with Large Language Models
di: Xu, Zexing, et al.
Pubblicazione: (2024)
di: Xu, Zexing, et al.
Pubblicazione: (2024)
MM-CRITIC: A Holistic Evaluation of Large Multimodal Models as Multimodal Critique
di: Zeng, Gailun, et al.
Pubblicazione: (2025)
di: Zeng, Gailun, et al.
Pubblicazione: (2025)
Establishing Task Scaling Laws via Compute-Efficient Model Ladders
di: Bhagia, Akshita, et al.
Pubblicazione: (2024)
di: Bhagia, Akshita, et al.
Pubblicazione: (2024)
AttentionRAG: Attention-Guided Context Pruning in Retrieval-Augmented Generation
di: Fang, Yixiong, et al.
Pubblicazione: (2025)
di: Fang, Yixiong, et al.
Pubblicazione: (2025)
Distilling Text Style Transfer With Self-Explanation From LLMs
di: Zhang, Chiyu, et al.
Pubblicazione: (2024)
di: Zhang, Chiyu, et al.
Pubblicazione: (2024)
Because we have LLMs, we Can and Should Pursue Agentic Interpretability
di: Kim, Been, et al.
Pubblicazione: (2025)
di: Kim, Been, et al.
Pubblicazione: (2025)
No Need for Explanations: LLMs can implicitly learn from mistakes in-context
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)
di: Alazraki, Lisa, et al.
Pubblicazione: (2025)
Paloma: A Benchmark for Evaluating Language Model Fit
di: Magnusson, Ian, et al.
Pubblicazione: (2023)
di: Magnusson, Ian, et al.
Pubblicazione: (2023)
MCQG-SRefine: Multiple Choice Question Generation and Evaluation with Iterative Self-Critique, Correction, and Comparison Feedback
di: Yao, Zonghai, et al.
Pubblicazione: (2024)
di: Yao, Zonghai, et al.
Pubblicazione: (2024)
Socratic-PRMBench: Benchmarking Process Reward Models with Systematic Reasoning Patterns
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
Evaluating Moral Beliefs across LLMs through a Pluralistic Framework
di: Liu, Xuelin, et al.
Pubblicazione: (2024)
di: Liu, Xuelin, et al.
Pubblicazione: (2024)
Halu-J: Critique-Based Hallucination Judge
di: Wang, Binjie, et al.
Pubblicazione: (2024)
di: Wang, Binjie, et al.
Pubblicazione: (2024)
Evaluating Evidence Attribution in Generated Fact Checking Explanations
di: Xing, Rui, et al.
Pubblicazione: (2024)
di: Xing, Rui, et al.
Pubblicazione: (2024)
On the Importance and Evaluation of Narrativity in Natural Language AI Explanations
di: Cedro, Mateusz, et al.
Pubblicazione: (2026)
di: Cedro, Mateusz, et al.
Pubblicazione: (2026)
LastingBench: Defend Benchmarks Against Knowledge Leakage
di: Fang, Yixiong, et al.
Pubblicazione: (2025)
di: Fang, Yixiong, et al.
Pubblicazione: (2025)
XplainLLM: A Knowledge-Augmented Dataset for Reliable Grounded Explanations in LLMs
di: Chen, Zichen, et al.
Pubblicazione: (2023)
di: Chen, Zichen, et al.
Pubblicazione: (2023)
RefineCoder: Iterative Improving of Large Language Models via Adaptive Critique Refinement for Code Generation
di: Zhou, Changzhi, et al.
Pubblicazione: (2025)
di: Zhou, Changzhi, et al.
Pubblicazione: (2025)
MARS: Multi-Agent Adaptive Reasoning with Socratic Guidance for Automated Prompt Optimization
di: Zhang, Jian, et al.
Pubblicazione: (2025)
di: Zhang, Jian, et al.
Pubblicazione: (2025)
Conversation for Non-verifiable Learning: Self-Evolving LLMs through Meta-Evaluation
di: Sui, Yuan, et al.
Pubblicazione: (2026)
di: Sui, Yuan, et al.
Pubblicazione: (2026)
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
di: Toker, Gilat, et al.
Pubblicazione: (2026)
di: Toker, Gilat, et al.
Pubblicazione: (2026)
Generation, Evaluation, and Explanation of Novelists' Styles with Single-Token Prompts
di: Rezaei, Mosab, et al.
Pubblicazione: (2025)
di: Rezaei, Mosab, et al.
Pubblicazione: (2025)
Med-CoDE: Medical Critique based Disagreement Evaluation Framework
di: Gupta, Mohit, et al.
Pubblicazione: (2025)
di: Gupta, Mohit, et al.
Pubblicazione: (2025)
RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques
di: Tang, Zhengyang, et al.
Pubblicazione: (2025)
di: Tang, Zhengyang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
SimpleToM: Exposing the Gap between Explicit ToM Inference and Implicit ToM Application in LLMs
di: Gu, Yuling, et al.
Pubblicazione: (2024) -
BaRDa: A Belief and Reasoning Dataset that Separates Factual Accuracy and Reasoning Ability
di: Clark, Peter, et al.
Pubblicazione: (2023) -
OLMES: A Standard for Language Model Evaluations
di: Gu, Yuling, et al.
Pubblicazione: (2024) -
DISCOVERYWORLD: A Virtual Environment for Developing and Evaluating Automated Scientific Discovery Agents
di: Jansen, Peter, et al.
Pubblicazione: (2024) -
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
di: Wiegreffe, Sarah, et al.
Pubblicazione: (2024)