Self-Tuning: Instructing LLMs to Effectively Acquire New Knowledge through Self-Teaching
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Xiaoying, Peng, Baolin, Tian, Ye, Zhou, Jingyan, Zhang, Yipeng, Mi, Haitao, Meng, Helen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024)
Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
von: Wang, Xiyao, et al.
Veröffentlicht: (2024)
Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing
von: Tian, Ye, et al.
Veröffentlicht: (2024)
von: Tian, Ye, et al.
Veröffentlicht: (2024)
SIaM: Self-Improving Code-Assisted Mathematical Reasoning of Large Language Models
von: Yu, Dian, et al.
Veröffentlicht: (2024)
von: Yu, Dian, et al.
Veröffentlicht: (2024)
Teaching Language Models to Self-Improve through Interactive Demonstrations
von: Yu, Xiao, et al.
Veröffentlicht: (2023)
von: Yu, Xiao, et al.
Veröffentlicht: (2023)
Rethinking Machine Ethics -- Can LLMs Perform Moral Reasoning through the Lens of Moral Theories?
von: Zhou, Jingyan, et al.
Veröffentlicht: (2023)
von: Zhou, Jingyan, et al.
Veröffentlicht: (2023)
Self-Consistency Boosts Calibration for Math Reasoning
von: Wang, Ante, et al.
Veröffentlicht: (2024)
von: Wang, Ante, et al.
Veröffentlicht: (2024)
Fine-Grained Self-Endorsement Improves Factuality and Reasoning
von: Wang, Ante, et al.
Veröffentlicht: (2024)
von: Wang, Ante, et al.
Veröffentlicht: (2024)
Collaborative decoding of critical tokens for boosting factuality of large language models
von: Jin, Lifeng, et al.
Veröffentlicht: (2024)
von: Jin, Lifeng, et al.
Veröffentlicht: (2024)
InverseCoder: Self-improving Instruction-Tuned Code LLMs with Inverse-Instruct
von: Wu, Yutong, et al.
Veröffentlicht: (2024)
von: Wu, Yutong, et al.
Veröffentlicht: (2024)
Teaching LLMs to Refine with Tools
von: Yu, Dian, et al.
Veröffentlicht: (2024)
von: Yu, Dian, et al.
Veröffentlicht: (2024)
Semi-Instruct: Bridging Natural-Instruct and Self-Instruct for Code Large Language Models
von: Luo, Xianzhen, et al.
Veröffentlicht: (2024)
von: Luo, Xianzhen, et al.
Veröffentlicht: (2024)
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2024)
Seal-Tools: Self-Instruct Tool Learning Dataset for Agent Tuning and Detailed Benchmark
von: Wu, Mengsong, et al.
Veröffentlicht: (2024)
von: Wu, Mengsong, et al.
Veröffentlicht: (2024)
Self-Error-Instruct: Generalizing from Errors for LLMs Mathematical Reasoning
von: Yu, Erxin, et al.
Veröffentlicht: (2025)
von: Yu, Erxin, et al.
Veröffentlicht: (2025)
InstructRAG: Instructing Retrieval-Augmented Generation via Self-Synthesized Rationales
von: Wei, Zhepei, et al.
Veröffentlicht: (2024)
von: Wei, Zhepei, et al.
Veröffentlicht: (2024)
AbsInstruct: Eliciting Abstraction Ability from LLMs through Explanation Tuning with Plausibility Estimation
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
von: Wang, Zhaowei, et al.
Veröffentlicht: (2024)
Building Task Bots with Self-learning for Enhanced Adaptability, Extensibility, and Factuality
von: Zhang, Xiaoying
Veröffentlicht: (2025)
von: Zhang, Xiaoying
Veröffentlicht: (2025)
TasTe: Teaching Large Language Models to Translate through Self-Reflection
von: Wang, Yutong, et al.
Veröffentlicht: (2024)
von: Wang, Yutong, et al.
Veröffentlicht: (2024)
Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models
von: Li, Miaoran, et al.
Veröffentlicht: (2023)
von: Li, Miaoran, et al.
Veröffentlicht: (2023)
Entropy Guided Extrapolative Decoding to Improve Factuality in Large Language Models
von: Das, Souvik, et al.
Veröffentlicht: (2024)
von: Das, Souvik, et al.
Veröffentlicht: (2024)
Triplets Better Than Pairs: Towards Stable and Effective Self-Play Fine-Tuning for LLMs
von: Wang, Yibo, et al.
Veröffentlicht: (2026)
von: Wang, Yibo, et al.
Veröffentlicht: (2026)
Learning to Retrieve and Reason on Knowledge Graph through Active Self-Reflection
von: Zhang, Han, et al.
Veröffentlicht: (2025)
von: Zhang, Han, et al.
Veröffentlicht: (2025)
S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
von: Ma, Ruotian, et al.
Veröffentlicht: (2025)
von: Ma, Ruotian, et al.
Veröffentlicht: (2025)
LiteSearch: Efficacious Tree Search for LLM
von: Wang, Ante, et al.
Veröffentlicht: (2024)
von: Wang, Ante, et al.
Veröffentlicht: (2024)
Self-Instructed Derived Prompt Generation Meets In-Context Learning: Unlocking New Potential of Black-Box LLMs
von: Li, Zhuo, et al.
Veröffentlicht: (2024)
von: Li, Zhuo, et al.
Veröffentlicht: (2024)
Purple-teaming LLMs with Adversarial Defender Training
von: Zhou, Jingyan, et al.
Veröffentlicht: (2024)
von: Zhou, Jingyan, et al.
Veröffentlicht: (2024)
Guided Self-Evolving LLMs with Minimal Human Supervision
von: Yu, Wenhao, et al.
Veröffentlicht: (2025)
von: Yu, Wenhao, et al.
Veröffentlicht: (2025)
Self-Prompt Tuning: Enable Autonomous Role-Playing in LLMs
von: Kong, Aobo, et al.
Veröffentlicht: (2024)
von: Kong, Aobo, et al.
Veröffentlicht: (2024)
Critique-GRPO: Advancing LLM Reasoning with Natural Language and Numerical Feedback
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2025)
WebEvolver: Enhancing Web Agent Self-Improvement with Coevolving World Model
von: Fang, Tianqing, et al.
Veröffentlicht: (2025)
von: Fang, Tianqing, et al.
Veröffentlicht: (2025)
Learning to Instruct for Visual Instruction Tuning
von: Zhou, Zhihan, et al.
Veröffentlicht: (2025)
von: Zhou, Zhihan, et al.
Veröffentlicht: (2025)
SeDi-Instruct: Enhancing Alignment of Language Models through Self-Directed Instruction Generation
von: Kim, Jungwoo, et al.
Veröffentlicht: (2025)
von: Kim, Jungwoo, et al.
Veröffentlicht: (2025)
How Do LLMs Acquire New Knowledge? A Knowledge Circuits Perspective on Continual Pre-Training
von: Ou, Yixin, et al.
Veröffentlicht: (2025)
von: Ou, Yixin, et al.
Veröffentlicht: (2025)
Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs
von: Li, Wu, et al.
Veröffentlicht: (2026)
von: Li, Wu, et al.
Veröffentlicht: (2026)
Step-On-Feet Tuning: Scaling Self-Alignment of LLMs via Bootstrapping
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
von: Wang, Haoyu, et al.
Veröffentlicht: (2024)
On the Intrinsic Self-Correction Capability of LLMs: Uncertainty and Latent Concept
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
von: Liu, Guangliang, et al.
Veröffentlicht: (2024)
InstructDiff: Domain-Adaptive Data Selection via Differential Entropy for Efficient LLM Fine-Tuning
von: Su, Junyou, et al.
Veröffentlicht: (2026)
von: Su, Junyou, et al.
Veröffentlicht: (2026)
WebCoT: Enhancing Web Agent Reasoning by Reconstructing Chain-of-Thought in Reflection, Branching, and Rollback
von: Hu, Minda, et al.
Veröffentlicht: (2025)
von: Hu, Minda, et al.
Veröffentlicht: (2025)
Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
von: Gekhman, Zorik, et al.
Veröffentlicht: (2024)
von: Gekhman, Zorik, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2024) -
Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning
von: Wang, Xiyao, et al.
Veröffentlicht: (2024) -
Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing
von: Tian, Ye, et al.
Veröffentlicht: (2024) -
SIaM: Self-Improving Code-Assisted Mathematical Reasoning of Large Language Models
von: Yu, Dian, et al.
Veröffentlicht: (2024) -
Teaching Language Models to Self-Improve through Interactive Demonstrations
von: Yu, Xiao, et al.
Veröffentlicht: (2023)