SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Tianyang, Wu, Shujin, Diao, Shizhe, Liu, Xiaoze, Wang, Xingyao, Chen, Yangyi, Gao, Jing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MedReflect: Teaching Medical LLMs to Self-Improve via Reflective Correction
di: Huang, Yue, et al.
Pubblicazione: (2025)
di: Huang, Yue, et al.
Pubblicazione: (2025)
SUV: Scalable Large Language Model Copyright Compliance with Regularized Selective Unlearning
di: Xu, Tianyang, et al.
Pubblicazione: (2025)
di: Xu, Tianyang, et al.
Pubblicazione: (2025)
Examining LLMs' Uncertainty Expression Towards Questions Outside Parametric Knowledge
di: Liu, Genglin, et al.
Pubblicazione: (2023)
di: Liu, Genglin, et al.
Pubblicazione: (2023)
Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
Confidence Improves Self-Consistency in LLMs
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)
SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
Towards Efficient CoT Distillation: Self-Guided Rationale Selector for Better Performance with Fewer Rationales
di: Yan, Jianzhi, et al.
Pubblicazione: (2025)
di: Yan, Jianzhi, et al.
Pubblicazione: (2025)
MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback
di: Wang, Xingyao, et al.
Pubblicazione: (2023)
di: Wang, Xingyao, et al.
Pubblicazione: (2023)
Teach-to-Reason with Scoring: Self-Explainable Rationale-Driven Multi-Trait Essay Scoring
di: Do, Heejin, et al.
Pubblicazione: (2025)
di: Do, Heejin, et al.
Pubblicazione: (2025)
CRAFT: Customizing LLMs by Creating and Retrieving from Specialized Toolsets
di: Yuan, Lifan, et al.
Pubblicazione: (2023)
di: Yuan, Lifan, et al.
Pubblicazione: (2023)
CodeGraph: Enhancing Graph Reasoning of LLMs with Code
di: Cai, Qiaolong, et al.
Pubblicazione: (2024)
di: Cai, Qiaolong, et al.
Pubblicazione: (2024)
Reasoning Models Better Express Their Confidence
di: Yoon, Dongkeun, et al.
Pubblicazione: (2025)
di: Yoon, Dongkeun, et al.
Pubblicazione: (2025)
Teaching LLMs to Abstain via Fine-Grained Semantic Confidence Reward
di: An, Hao, et al.
Pubblicazione: (2025)
di: An, Hao, et al.
Pubblicazione: (2025)
Can We Verify Step by Step for Incorrect Answer Detection?
di: Xu, Xin, et al.
Pubblicazione: (2024)
di: Xu, Xin, et al.
Pubblicazione: (2024)
Filtered Reasoning Score: Evaluating Reasoning Quality on a Model's Most-Confident Traces
di: Pathak, Manas, et al.
Pubblicazione: (2026)
di: Pathak, Manas, et al.
Pubblicazione: (2026)
ConfTuner: Training Large Language Models to Express Their Confidence Verbally
di: Li, Yibo, et al.
Pubblicazione: (2025)
di: Li, Yibo, et al.
Pubblicazione: (2025)
R-Tuning: Instructing Large Language Models to Say `I Don't Know'
di: Zhang, Hanning, et al.
Pubblicazione: (2023)
di: Zhang, Hanning, et al.
Pubblicazione: (2023)
Beyond Confidence: Rethinking Self-Assessments for Performance Prediction in LLMs
di: Bhattacharyya, Sree, et al.
Pubblicazione: (2026)
di: Bhattacharyya, Sree, et al.
Pubblicazione: (2026)
TasTe: Teaching Large Language Models to Translate through Self-Reflection
di: Wang, Yutong, et al.
Pubblicazione: (2024)
di: Wang, Yutong, et al.
Pubblicazione: (2024)
Self-Training Meets Consistency: Improving LLMs' Reasoning with Consistency-Driven Rationale Evaluation
di: Lee, Jaehyeok, et al.
Pubblicazione: (2024)
di: Lee, Jaehyeok, et al.
Pubblicazione: (2024)
CAMEL: Confidence-Gated Reflection for Reward Modeling
di: Zhu, Zirui, et al.
Pubblicazione: (2026)
di: Zhu, Zirui, et al.
Pubblicazione: (2026)
Fact-Level Confidence Calibration and Self-Correction
di: Yuan, Yige, et al.
Pubblicazione: (2024)
di: Yuan, Yige, et al.
Pubblicazione: (2024)
Executable Code Actions Elicit Better LLM Agents
di: Wang, Xingyao, et al.
Pubblicazione: (2024)
di: Wang, Xingyao, et al.
Pubblicazione: (2024)
Learning to Plan Before Answering: Self-Teaching LLMs to Learn Abstract Plans for Problem Solving
di: Zhang, Jin, et al.
Pubblicazione: (2025)
di: Zhang, Jin, et al.
Pubblicazione: (2025)
Calibrating Verbalized Confidence with Self-Generated Distractors
di: Wang, Victor, et al.
Pubblicazione: (2025)
di: Wang, Victor, et al.
Pubblicazione: (2025)
LLMs on a Budget? Say HOLA
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)
di: Siddiqui, Zohaib Hasan, et al.
Pubblicazione: (2025)
Reinforcement Learning from Reflective Feedback (RLRF): Aligning and Improving LLMs via Fine-Grained Self-Reflection
di: Lee, Kyungjae, et al.
Pubblicazione: (2024)
di: Lee, Kyungjae, et al.
Pubblicazione: (2024)
SelectIT: Selective Instruction Tuning for LLMs via Uncertainty-Aware Self-Reflection
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
di: Liu, Liangxin, et al.
Pubblicazione: (2024)
Confidence-aware Self-Semantic Distillation on Knowledge Graph Embedding
di: Liu, Yichen, et al.
Pubblicazione: (2022)
di: Liu, Yichen, et al.
Pubblicazione: (2022)
When to Trust Context: Self-Reflective Debates for Context Reliability
di: Zhou, Zeqi, et al.
Pubblicazione: (2025)
di: Zhou, Zeqi, et al.
Pubblicazione: (2025)
Rationale Behind Essay Scores: Enhancing S-LLM's Multi-Trait Essay Scoring with Rationale Generated by LLMs
di: Chu, SeongYeub, et al.
Pubblicazione: (2024)
di: Chu, SeongYeub, et al.
Pubblicazione: (2024)
SelfReflect: Can LLMs Communicate Their Internal Answer Distribution?
di: Kirchhof, Michael, et al.
Pubblicazione: (2025)
di: Kirchhof, Michael, et al.
Pubblicazione: (2025)
The Riddle of Reflection: Evaluating Reasoning and Self-Awareness in Multilingual LLMs using Indian Riddles
di: M, Abhinav P, et al.
Pubblicazione: (2025)
di: M, Abhinav P, et al.
Pubblicazione: (2025)
SyncMind: Measuring Agent Out-of-Sync Recovery in Collaborative Software Engineering
di: Guo, Xuehang, et al.
Pubblicazione: (2025)
di: Guo, Xuehang, et al.
Pubblicazione: (2025)
TheoremLlama: Transforming General-Purpose LLMs into Lean4 Experts
di: Wang, Ruida, et al.
Pubblicazione: (2024)
di: Wang, Ruida, et al.
Pubblicazione: (2024)
Confidence Matters: Revisiting Intrinsic Self-Correction Capabilities of Large Language Models
di: Li, Loka, et al.
Pubblicazione: (2024)
di: Li, Loka, et al.
Pubblicazione: (2024)
Scaling Laws for Predicting Downstream Performance in LLMs
di: Chen, Yangyi, et al.
Pubblicazione: (2024)
di: Chen, Yangyi, et al.
Pubblicazione: (2024)
Self-Evaluating LLMs for Multi-Step Tasks: Stepwise Confidence Estimation for Failure Detection
di: Mavi, Vaibhav, et al.
Pubblicazione: (2025)
di: Mavi, Vaibhav, et al.
Pubblicazione: (2025)
CoRefine: Confidence-Guided Self-Refinement for Adaptive Test-Time Compute
di: Jin, Chen, et al.
Pubblicazione: (2026)
di: Jin, Chen, et al.
Pubblicazione: (2026)
Efficient Reasoning Through Suppression of Self-Affirmation Reflections in Large Reasoning Models
di: Liu, Kaiyuan, et al.
Pubblicazione: (2025)
di: Liu, Kaiyuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MedReflect: Teaching Medical LLMs to Self-Improve via Reflective Correction
di: Huang, Yue, et al.
Pubblicazione: (2025) -
SUV: Scalable Large Language Model Copyright Compliance with Regularized Selective Unlearning
di: Xu, Tianyang, et al.
Pubblicazione: (2025) -
Examining LLMs' Uncertainty Expression Towards Questions Outside Parametric Knowledge
di: Liu, Genglin, et al.
Pubblicazione: (2023) -
Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs
di: Liu, Xiaoze, et al.
Pubblicazione: (2024) -
Confidence Improves Self-Consistency in LLMs
di: Taubenfeld, Amir, et al.
Pubblicazione: (2025)