Small Language Model Can Self-correct
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Haixia, Liang, Jiaqing, Shi, Jie, He, Qianyu, Xiao, Yanghua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CEM: A Data-Efficient Method for Large Language Models to Continue Evolving From Mistakes
von: Zhao, Haokun, et al.
Veröffentlicht: (2024)
von: Zhao, Haokun, et al.
Veröffentlicht: (2024)
Enhancing Quantitative Reasoning Skills of Large Language Models through Dimension Perception
von: Huang, Yuncheng, et al.
Veröffentlicht: (2023)
von: Huang, Yuncheng, et al.
Veröffentlicht: (2023)
Step-by-Step Mastery: Enhancing Soft Constraint Following Ability of Large Language Models
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
SED: Self-Evaluation Decoding Enhances Large Language Models for Better Generation
von: Luo, Ziqin, et al.
Veröffentlicht: (2024)
von: Luo, Ziqin, et al.
Veröffentlicht: (2024)
Reason from Fallacy: Enhancing Large Language Models' Logical Reasoning through Logical Fallacy Understanding
von: Li, Yanda, et al.
Veröffentlicht: (2024)
von: Li, Yanda, et al.
Veröffentlicht: (2024)
Instructions are all you need: Self-supervised Reinforcement Learning for Instruction Following
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
Can Pre-trained Language Models Understand Chinese Humor?
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
QUILL: Quotation Generation Enhancement of Large Language Models
von: Xiao, Jin, et al.
Veröffentlicht: (2024)
von: Xiao, Jin, et al.
Veröffentlicht: (2024)
Laying the Foundation First? Investigating the Generalization from Atomic Skills to Complex Reasoning Tasks
von: Huang, Yuncheng, et al.
Veröffentlicht: (2024)
von: Huang, Yuncheng, et al.
Veröffentlicht: (2024)
Order Doesn't Matter, But Reasoning Does: Training LLMs with Order-Centric Augmentation
von: He, Qianxi, et al.
Veröffentlicht: (2025)
von: He, Qianxi, et al.
Veröffentlicht: (2025)
Can Large Language Models Understand Real-World Complex Instructions?
von: He, Qianyu, et al.
Veröffentlicht: (2023)
von: He, Qianyu, et al.
Veröffentlicht: (2023)
Order Matters: Investigate the Position Bias in Multi-constraint Instruction Following
von: Zeng, Jie, et al.
Veröffentlicht: (2025)
von: Zeng, Jie, et al.
Veröffentlicht: (2025)
What Makes an Ideal Quote? Recommending "Unexpected yet Rational" Quotations via Novelty
von: Zhang, Bowei, et al.
Veröffentlicht: (2025)
von: Zhang, Bowei, et al.
Veröffentlicht: (2025)
Think Thrice Before You Act: Progressive Thought Refinement in Large Language Models
von: Du, Chengyu, et al.
Veröffentlicht: (2024)
von: Du, Chengyu, et al.
Veröffentlicht: (2024)
ANALOGYKB: Unlocking Analogical Reasoning of Language Models with A Million-scale Knowledge Base
von: Yuan, Siyu, et al.
Veröffentlicht: (2023)
von: Yuan, Siyu, et al.
Veröffentlicht: (2023)
Chain-of-Knowledge: Integrating Knowledge Reasoning into Large Language Models by Learning from Knowledge Graphs
von: Zhang, Yifei, et al.
Veröffentlicht: (2024)
von: Zhang, Yifei, et al.
Veröffentlicht: (2024)
From Complex to Simple: Enhancing Multi-Constraint Complex Instruction Following Ability of Large Language Models
von: He, Qianyu, et al.
Veröffentlicht: (2024)
von: He, Qianyu, et al.
Veröffentlicht: (2024)
Past Meets Present: Creating Historical Analogy with Large Language Models
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
Do Large Language Models have Problem-Solving Capability under Incomplete Information Scenarios?
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
A Stitch in Time Saves Nine: Proactive Self-Refinement for Language Models
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
Recent Advancement of Emotion Cognition in Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Skeletons Matter: Dynamic Data Augmentation for Text-to-Query
von: Ji, Yuchen, et al.
Veröffentlicht: (2025)
von: Ji, Yuchen, et al.
Veröffentlicht: (2025)
Mind the Generation Process: Fine-Grained Confidence Estimation During LLM Generation
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
Is There a One-Model-Fits-All Approach to Information Extraction? Revisiting Task Definition Biases
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
Enhancing Confidence Expression in Large Language Models Through Learning from Past Experience
von: Han, Haixia, et al.
Veröffentlicht: (2024)
von: Han, Haixia, et al.
Veröffentlicht: (2024)
AutoScraper: A Progressive Understanding Web Agent for Web Scraper Generation
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
von: Huang, Wenhao, et al.
Veröffentlicht: (2024)
Selective Expert Guidance for Effective and Diverse Exploration in Reinforcement Learning of LLMs
von: Jiang, Zishang, et al.
Veröffentlicht: (2025)
von: Jiang, Zishang, et al.
Veröffentlicht: (2025)
Light Up the Shadows: Enhance Long-Tailed Entity Grounding with Concept-Guided Vision-Language Models
von: Zhang, Yikai, et al.
Veröffentlicht: (2024)
von: Zhang, Yikai, et al.
Veröffentlicht: (2024)
ToNER: Type-oriented Named Entity Recognition with Generative Language Model
von: Jiang, Guochao, et al.
Veröffentlicht: (2024)
von: Jiang, Guochao, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models for Entity Matching
von: Huang, Qianyu, et al.
Veröffentlicht: (2024)
von: Huang, Qianyu, et al.
Veröffentlicht: (2024)
Your Models Have Thought Enough: Training Large Reasoning Models to Stop Overthinking
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
von: Han, Jinyi, et al.
Veröffentlicht: (2025)
Beyond the Trade-off: Self-Supervised Reinforcement Learning for Reasoning Models' Instruction Following
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
von: Ren, Qingyu, et al.
Veröffentlicht: (2025)
Can Language Models Solve Graph Problems in Natural Language?
von: Wang, Heng, et al.
Veröffentlicht: (2023)
von: Wang, Heng, et al.
Veröffentlicht: (2023)
CultureScope: A Dimensional Lens for Probing Cultural Understanding in LLMs
von: Zhang, Jinghao, et al.
Veröffentlicht: (2025)
von: Zhang, Jinghao, et al.
Veröffentlicht: (2025)
LightReasoner: Can Small Language Models Teach Large Language Models Reasoning?
von: Wang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Wang, Jingyuan, et al.
Veröffentlicht: (2025)
AdaptiveLog: An Adaptive Log Analysis Framework with the Collaboration of Large and Small Language Model
von: Ma, Lipeng, et al.
Veröffentlicht: (2025)
von: Ma, Lipeng, et al.
Veröffentlicht: (2025)
SelfGoal: Your Language Agents Already Know How to Achieve High-level Goals
von: Yang, Ruihan, et al.
Veröffentlicht: (2024)
von: Yang, Ruihan, et al.
Veröffentlicht: (2024)
StrucText-Eval: Evaluating Large Language Model's Reasoning Ability in Structure-Rich Text
von: Gu, Zhouhong, et al.
Veröffentlicht: (2024)
von: Gu, Zhouhong, et al.
Veröffentlicht: (2024)
ChemAmp: Amplified Chemistry Tools via Composable Agents
von: Li, Zhucong, et al.
Veröffentlicht: (2025)
von: Li, Zhucong, et al.
Veröffentlicht: (2025)
EmotionQueen: A Benchmark for Evaluating Empathy of Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CEM: A Data-Efficient Method for Large Language Models to Continue Evolving From Mistakes
von: Zhao, Haokun, et al.
Veröffentlicht: (2024) -
Enhancing Quantitative Reasoning Skills of Large Language Models through Dimension Perception
von: Huang, Yuncheng, et al.
Veröffentlicht: (2023) -
Step-by-Step Mastery: Enhancing Soft Constraint Following Ability of Large Language Models
von: Ren, Qingyu, et al.
Veröffentlicht: (2025) -
SED: Self-Evaluation Decoding Enhances Large Language Models for Better Generation
von: Luo, Ziqin, et al.
Veröffentlicht: (2024) -
Reason from Fallacy: Enhancing Large Language Models' Logical Reasoning through Logical Fallacy Understanding
von: Li, Yanda, et al.
Veröffentlicht: (2024)