Beyond Sparse Rewards: Enhancing Reinforcement Learning with Language Model Critique in Text Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cao, Meng, Shu, Lei, Yu, Lei, Zhu, Yun, Wichers, Nevan, Liu, Yinxiao, Meng, Lei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Fusion-Eval: Integrating Assistant Evaluators with LLMs
von: Shu, Lei, et al.
Veröffentlicht: (2023)
von: Shu, Lei, et al.
Veröffentlicht: (2023)
Rewarding Creativity: A Human-Aligned Generative Reward Model for Reinforcement Learning in Storytelling
von: Li, Zhaoyan, et al.
Veröffentlicht: (2026)
von: Li, Zhaoyan, et al.
Veröffentlicht: (2026)
CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
von: Ke, Pei, et al.
Veröffentlicht: (2023)
von: Ke, Pei, et al.
Veröffentlicht: (2023)
Self-Generated Critiques Boost Reward Modeling for Language Models
von: Yu, Yue, et al.
Veröffentlicht: (2024)
von: Yu, Yue, et al.
Veröffentlicht: (2024)
Gradient-Based Language Model Red Teaming
von: Wichers, Nevan, et al.
Veröffentlicht: (2024)
von: Wichers, Nevan, et al.
Veröffentlicht: (2024)
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations
von: Yu, Lei, et al.
Veröffentlicht: (2024)
von: Yu, Lei, et al.
Veröffentlicht: (2024)
The Reasoning-Memorization Interplay in Language Models Is Mediated by a Single Direction
von: Hong, Yihuai, et al.
Veröffentlicht: (2025)
von: Hong, Yihuai, et al.
Veröffentlicht: (2025)
Model Spec Midtraining: Improving How Alignment Training Generalizes
von: Li, Chloe, et al.
Veröffentlicht: (2026)
von: Li, Chloe, et al.
Veröffentlicht: (2026)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
von: Xie, Tianbao, et al.
Veröffentlicht: (2023)
Critique-RL: Training Language Models for Critiquing through Two-Stage Reinforcement Learning
von: Xi, Zhiheng, et al.
Veröffentlicht: (2025)
von: Xi, Zhiheng, et al.
Veröffentlicht: (2025)
DynaWeb: Model-Based Reinforcement Learning of Web Agents
von: Ding, Hang, et al.
Veröffentlicht: (2026)
von: Ding, Hang, et al.
Veröffentlicht: (2026)
LLaMAX: Scaling Linguistic Horizons of LLM by Enhancing Translation Capabilities Beyond 100 Languages
von: Lu, Yinquan, et al.
Veröffentlicht: (2024)
von: Lu, Yinquan, et al.
Veröffentlicht: (2024)
Dr Genre: Reinforcement Learning from Decoupled LLM Feedback for Generic Text Rewriting
von: Li, Yufei, et al.
Veröffentlicht: (2025)
von: Li, Yufei, et al.
Veröffentlicht: (2025)
A Practical Examination of AI-Generated Text Detectors for Large Language Models
von: Tufts, Brian, et al.
Veröffentlicht: (2024)
von: Tufts, Brian, et al.
Veröffentlicht: (2024)
CriSPO: Multi-Aspect Critique-Suggestion-guided Automatic Prompt Optimization for Text Generation
von: He, Han, et al.
Veröffentlicht: (2024)
von: He, Han, et al.
Veröffentlicht: (2024)
An LLM Maturity Model for Reliable and Transparent Text-to-Query
von: Yu, Lei, et al.
Veröffentlicht: (2024)
von: Yu, Lei, et al.
Veröffentlicht: (2024)
Visualizing Neural Network Imagination
von: Wichers, Nevan, et al.
Veröffentlicht: (2024)
von: Wichers, Nevan, et al.
Veröffentlicht: (2024)
Towards Better Text-to-Image Generation Alignment via Attention Modulation
von: Wu, Yihang, et al.
Veröffentlicht: (2024)
von: Wu, Yihang, et al.
Veröffentlicht: (2024)
Teaching Language Models to Critique via Reinforcement Learning
von: Xie, Zhihui, et al.
Veröffentlicht: (2025)
von: Xie, Zhihui, et al.
Veröffentlicht: (2025)
WildReward: Learning Reward Models from In-the-Wild Human Interactions
von: Peng, Hao, et al.
Veröffentlicht: (2026)
von: Peng, Hao, et al.
Veröffentlicht: (2026)
Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals
von: Chen, Sirui, et al.
Veröffentlicht: (2026)
von: Chen, Sirui, et al.
Veröffentlicht: (2026)
LGM: Enhancing Large Language Models with Conceptual Meta-Relations and Iterative Retrieval
von: Lei, Wenchang, et al.
Veröffentlicht: (2025)
von: Lei, Wenchang, et al.
Veröffentlicht: (2025)
CTRL-RAG: Contrastive Likelihood Reward Based Reinforcement Learning for Context-Faithful RAG Models
von: Tan, Zhehao, et al.
Veröffentlicht: (2026)
von: Tan, Zhehao, et al.
Veröffentlicht: (2026)
Balancing Rewards in Text Summarization: Multi-Objective Reinforcement Learning via HyperVolume Optimization
von: Song, Junjie, et al.
Veröffentlicht: (2025)
von: Song, Junjie, et al.
Veröffentlicht: (2025)
Survey on Large Language Model-Enhanced Reinforcement Learning: Concept, Taxonomy, and Methods
von: Cao, Yuji, et al.
Veröffentlicht: (2024)
von: Cao, Yuji, et al.
Veröffentlicht: (2024)
Segmenting Text and Learning Their Rewards for Improved RLHF in Language Model
von: Yin, Yueqin, et al.
Veröffentlicht: (2025)
von: Yin, Yueqin, et al.
Veröffentlicht: (2025)
Beyond Gradient and Priors in Privacy Attacks: Leveraging Pooler Layer Inputs of Language Models in Federated Learning
von: Li, Jianwei, et al.
Veröffentlicht: (2023)
von: Li, Jianwei, et al.
Veröffentlicht: (2023)
Instructing LLMs to Negotiate using Reinforcement Learning with Verifiable Rewards
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2026)
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2026)
Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains
von: Gunjal, Anisha, et al.
Veröffentlicht: (2025)
von: Gunjal, Anisha, et al.
Veröffentlicht: (2025)
StoryAlign: Evaluating and Training Reward Models for Story Generation
von: Xia, Haotian, et al.
Veröffentlicht: (2026)
von: Xia, Haotian, et al.
Veröffentlicht: (2026)
Companion Agents: A Table-Information Mining Paradigm for Text-to-SQL
von: Chen, Jiahui, et al.
Veröffentlicht: (2025)
von: Chen, Jiahui, et al.
Veröffentlicht: (2025)
GDGB: A Benchmark for Generative Dynamic Text-Attributed Graph Learning
von: Peng, Jie, et al.
Veröffentlicht: (2025)
von: Peng, Jie, et al.
Veröffentlicht: (2025)
Accelerating Inference of Retrieval-Augmented Generation via Sparse Context Selection
von: Zhu, Yun, et al.
Veröffentlicht: (2024)
von: Zhu, Yun, et al.
Veröffentlicht: (2024)
ProtSAE: Disentangling and Interpreting Protein Language Models via Semantically-Guided Sparse Autoencoders
von: Liu, Xiangyu, et al.
Veröffentlicht: (2025)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2025)
MRScore: Evaluating Radiology Report Generation with LLM-based Reward System
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
Critique-GRPO: Advancing LLM Reasoning with Natural Language and Numerical Feedback
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoying, et al.
Veröffentlicht: (2025)
Towards Faithful and Controllable Personalization via Critique-Post-Edit Reinforcement Learning
von: Zhu, Chenghao, et al.
Veröffentlicht: (2025)
von: Zhu, Chenghao, et al.
Veröffentlicht: (2025)
Self-Rewarding Rubric-Based Reinforcement Learning for Open-Ended Reasoning
von: Ye, Zhiling, et al.
Veröffentlicht: (2025)
von: Ye, Zhiling, et al.
Veröffentlicht: (2025)
Reward Is Enough: LLMs Are In-Context Reinforcement Learners
von: Song, Kefan, et al.
Veröffentlicht: (2025)
von: Song, Kefan, et al.
Veröffentlicht: (2025)
The Critique of Critique
von: Sun, Shichao, et al.
Veröffentlicht: (2024)
von: Sun, Shichao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Fusion-Eval: Integrating Assistant Evaluators with LLMs
von: Shu, Lei, et al.
Veröffentlicht: (2023) -
Rewarding Creativity: A Human-Aligned Generative Reward Model for Reinforcement Learning in Storytelling
von: Li, Zhaoyan, et al.
Veröffentlicht: (2026) -
CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
von: Ke, Pei, et al.
Veröffentlicht: (2023) -
Self-Generated Critiques Boost Reward Modeling for Language Models
von: Yu, Yue, et al.
Veröffentlicht: (2024) -
Gradient-Based Language Model Red Teaming
von: Wichers, Nevan, et al.
Veröffentlicht: (2024)