Self-controller: Controlling LLMs with Multi-round Step-by-step Self-awareness
Fuente:
arXiv
Guardado en:
| Autores principales: | Peng, Xiao, Geng, Xufan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Do LLMs Really Think Step-by-step In Implicit Reasoning?
por: Yu, Yijiong
Publicado: (2024)
por: Yu, Yijiong
Publicado: (2024)
Step-On-Feet Tuning: Scaling Self-Alignment of LLMs via Bootstrapping
por: Wang, Haoyu, et al.
Publicado: (2024)
por: Wang, Haoyu, et al.
Publicado: (2024)
Self-Evaluating LLMs for Multi-Step Tasks: Stepwise Confidence Estimation for Failure Detection
por: Mavi, Vaibhav, et al.
Publicado: (2025)
por: Mavi, Vaibhav, et al.
Publicado: (2025)
From Long to Lean: Performance-aware and Adaptive Chain-of-Thought Compression via Multi-round Refinement
por: Yan, Jianzhi, et al.
Publicado: (2025)
por: Yan, Jianzhi, et al.
Publicado: (2025)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
por: Zhang, Xiaoying, et al.
Publicado: (2024)
por: Zhang, Xiaoying, et al.
Publicado: (2024)
LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked
por: Phute, Mansi, et al.
Publicado: (2023)
por: Phute, Mansi, et al.
Publicado: (2023)
Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning
por: Li, Yiwei, et al.
Publicado: (2024)
por: Li, Yiwei, et al.
Publicado: (2024)
Step-Tagging: Toward controlling the generation of Language Reasoning Models through step monitoring
por: Belkhiter, Yannis, et al.
Publicado: (2025)
por: Belkhiter, Yannis, et al.
Publicado: (2025)
Knowing You Don't Know: Learning When to Continue Search in Multi-round RAG through Self-Practicing
por: Yang, Diji, et al.
Publicado: (2025)
por: Yang, Diji, et al.
Publicado: (2025)
Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
por: Wang, Peiyi, et al.
Publicado: (2023)
por: Wang, Peiyi, et al.
Publicado: (2023)
Step Back to Leap Forward: Self-Backtracking for Boosting Reasoning of Language Models
por: Yang, Xiao-Wen, et al.
Publicado: (2025)
por: Yang, Xiao-Wen, et al.
Publicado: (2025)
Multi-round jailbreak attack on large language models
por: Zhou, Yihua, et al.
Publicado: (2024)
por: Zhou, Yihua, et al.
Publicado: (2024)
A Lightweight Framework for Trigger-Guided LoRA-Based Self-Adaptation in LLMs
por: Wei, Jiacheng, et al.
Publicado: (2025)
por: Wei, Jiacheng, et al.
Publicado: (2025)
Self-Evolved Reward Learning for LLMs
por: Huang, Chenghua, et al.
Publicado: (2024)
por: Huang, Chenghua, et al.
Publicado: (2024)
Confidence Improves Self-Consistency in LLMs
por: Taubenfeld, Amir, et al.
Publicado: (2025)
por: Taubenfeld, Amir, et al.
Publicado: (2025)
Self-supervised Attribute-aware Dynamic Preference Ranking Alignment
por: Yang, Hongyu, et al.
Publicado: (2025)
por: Yang, Hongyu, et al.
Publicado: (2025)
The Self-Execution Benchmark: Measuring LLMs' Attempts to Overcome Their Lack of Self-Execution
por: Ezra, Elon, et al.
Publicado: (2025)
por: Ezra, Elon, et al.
Publicado: (2025)
Improving the Reliability of LLMs: Combining CoT, RAG, Self-Consistency, and Self-Verification
por: Kumar, Adarsh, et al.
Publicado: (2025)
por: Kumar, Adarsh, et al.
Publicado: (2025)
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
por: Tie, Guiyao, et al.
Publicado: (2025)
por: Tie, Guiyao, et al.
Publicado: (2025)
AskToAct: Enhancing LLMs Tool Use via Self-Correcting Clarification
por: Zhang, Xuan, et al.
Publicado: (2025)
por: Zhang, Xuan, et al.
Publicado: (2025)
Regression-aware Inference with LLMs
por: Lukasik, Michal, et al.
Publicado: (2024)
por: Lukasik, Michal, et al.
Publicado: (2024)
Theory of Mind and Self-Attributions of Mentality are Dissociable in LLMs
por: Kim, Junsol, et al.
Publicado: (2026)
por: Kim, Junsol, et al.
Publicado: (2026)
PoTPTQ: A Two-step Power-of-Two Post-training for LLMs
por: Wang, Xinyu, et al.
Publicado: (2025)
por: Wang, Xinyu, et al.
Publicado: (2025)
Mitigating Attention Localization in Small Scale: Self-Attention Refinement via One-step Belief Propagation
por: Lee, Nakyung, et al.
Publicado: (2025)
por: Lee, Nakyung, et al.
Publicado: (2025)
What Defines Good Reasoning in LLMs? Dissecting Reasoning Steps with Multi-Aspect Evaluation
por: Do, Heejin, et al.
Publicado: (2025)
por: Do, Heejin, et al.
Publicado: (2025)
From Building Blocks to Planning: Multi-Step Spatial Reasoning in LLMs with Reinforcement Learning
por: Tahmasbi, Amir, et al.
Publicado: (2025)
por: Tahmasbi, Amir, et al.
Publicado: (2025)
Self-Improving Customer Review Response Generation Based on LLMs
por: Azov, Guy, et al.
Publicado: (2024)
por: Azov, Guy, et al.
Publicado: (2024)
Distilling Text Style Transfer With Self-Explanation From LLMs
por: Zhang, Chiyu, et al.
Publicado: (2024)
por: Zhang, Chiyu, et al.
Publicado: (2024)
SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition
por: Wu, Mengsong, et al.
Publicado: (2025)
por: Wu, Mengsong, et al.
Publicado: (2025)
Cascaded Self-Evaluation Augmented Training for Lightweight Multimodal LLMs
por: Lv, Zheqi, et al.
Publicado: (2025)
por: Lv, Zheqi, et al.
Publicado: (2025)
Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
por: Lai, Xin, et al.
Publicado: (2024)
por: Lai, Xin, et al.
Publicado: (2024)
Step-by-Step Reasoning to Solve Grid Puzzles: Where do LLMs Falter?
por: Tyagi, Nemika, et al.
Publicado: (2024)
por: Tyagi, Nemika, et al.
Publicado: (2024)
SmartThinker: Learning to Compress and Preserve Reasoning by Step-Level Length Control
por: He, Xingyang, et al.
Publicado: (2025)
por: He, Xingyang, et al.
Publicado: (2025)
SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese
por: Xu, Liang, et al.
Publicado: (2024)
por: Xu, Liang, et al.
Publicado: (2024)
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
por: Xu, Tianyang, et al.
Publicado: (2024)
por: Xu, Tianyang, et al.
Publicado: (2024)
Tracking the Limits of Knowledge Propagation: How LLMs Fail at Multi-Step Reasoning with Conflicting Knowledge
por: Feng, Yiyang, et al.
Publicado: (2026)
por: Feng, Yiyang, et al.
Publicado: (2026)
Local Explanations and Self-Explanations for Assessing Faithfulness in black-box LLMs
por: Fragkathoulas, Christos, et al.
Publicado: (2024)
por: Fragkathoulas, Christos, et al.
Publicado: (2024)
From Chains to Graphs: Self-Structured Reasoning for General-Domain LLMs
por: Chen, Yingjian, et al.
Publicado: (2026)
por: Chen, Yingjian, et al.
Publicado: (2026)
Lie to Me: Knowledge Graphs for Robust Hallucination Self-Detection in LLMs
por: Kale, Sahil, et al.
Publicado: (2025)
por: Kale, Sahil, et al.
Publicado: (2025)
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
por: Xu, Wenda, et al.
Publicado: (2025)
por: Xu, Wenda, et al.
Publicado: (2025)
Ejemplares similares
-
Do LLMs Really Think Step-by-step In Implicit Reasoning?
por: Yu, Yijiong
Publicado: (2024) -
Step-On-Feet Tuning: Scaling Self-Alignment of LLMs via Bootstrapping
por: Wang, Haoyu, et al.
Publicado: (2024) -
Self-Evaluating LLMs for Multi-Step Tasks: Stepwise Confidence Estimation for Failure Detection
por: Mavi, Vaibhav, et al.
Publicado: (2025) -
From Long to Lean: Performance-aware and Adaptive Chain-of-Thought Compression via Multi-round Refinement
por: Yan, Jianzhi, et al.
Publicado: (2025) -
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
por: Zhang, Xiaoying, et al.
Publicado: (2024)