Towards Socially and Morally Aware RL agent: Reward Design With LLM
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Wang, Zhaoyue |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
von: Muslimani, Calarina, et al.
Veröffentlicht: (2025)
von: Muslimani, Calarina, et al.
Veröffentlicht: (2025)
On Designing Effective RL Reward at Training Time for LLM Reasoning
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2024)
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2024)
When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL
von: Wang, Youting, et al.
Veröffentlicht: (2026)
von: Wang, Youting, et al.
Veröffentlicht: (2026)
Boosting Universal LLM Reward Design through Heuristic Reward Observation Space Evolution
von: Heng, Zen Kit, et al.
Veröffentlicht: (2025)
von: Heng, Zen Kit, et al.
Veröffentlicht: (2025)
FlowRL: Matching Reward Distributions for LLM Reasoning
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)
ProgAgent:A Continual RL Agent with Progress-Aware Rewards
von: Tan, Jinzhou, et al.
Veröffentlicht: (2026)
von: Tan, Jinzhou, et al.
Veröffentlicht: (2026)
RHyVE: Competence-Aware Verification and Phase-Aware Deployment for LLM-Generated Reward Hypotheses
von: Wu, Feiyu, et al.
Veröffentlicht: (2026)
von: Wu, Feiyu, et al.
Veröffentlicht: (2026)
Accelerating Diffusion Planners in Offline RL via Reward-Aware Consistency Trajectory Distillation
von: Duan, Xintong, et al.
Veröffentlicht: (2025)
von: Duan, Xintong, et al.
Veröffentlicht: (2025)
RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models
von: Feng, Xiao, et al.
Veröffentlicht: (2026)
von: Feng, Xiao, et al.
Veröffentlicht: (2026)
STO-RL: Offline RL under Sparse Rewards via LLM-Guided Subgoal Temporal Order
von: Gu, Chengyang, et al.
Veröffentlicht: (2026)
von: Gu, Chengyang, et al.
Veröffentlicht: (2026)
Exploring RL-based LLM Training for Formal Language Tasks with Programmed Rewards
von: Padula, Alexander G., et al.
Veröffentlicht: (2024)
von: Padula, Alexander G., et al.
Veröffentlicht: (2024)
Towards Simulating Social Influence Dynamics with LLM-based Multi-agents
von: Lin, Hsien-Tsung, et al.
Veröffentlicht: (2025)
von: Lin, Hsien-Tsung, et al.
Veröffentlicht: (2025)
Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026)
PROTEUS: SLA-Aware Routing via Lagrangian RL for Multi-LLM Serving Systems
von: Bhatti, Amit Singh, et al.
Veröffentlicht: (2026)
von: Bhatti, Amit Singh, et al.
Veröffentlicht: (2026)
Evaluating and Improving Cultural Awareness of Reward Models for LLM Alignment
von: Zhang, Hongbin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongbin, et al.
Veröffentlicht: (2025)
On the Generalization Gap in LLM Planning: Tests and Verifier-Reward RL
von: Belcamino, Valerio, et al.
Veröffentlicht: (2026)
von: Belcamino, Valerio, et al.
Veröffentlicht: (2026)
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
von: Gaven, Loris, et al.
Veröffentlicht: (2024)
von: Gaven, Loris, et al.
Veröffentlicht: (2024)
Verifier-Free RL for LLMs via Intrinsic Gradient-Norm Reward
von: Wen, Xuexiang, et al.
Veröffentlicht: (2026)
von: Wen, Xuexiang, et al.
Veröffentlicht: (2026)
Planner-R1: Reward Shaping Enables Efficient Agentic RL with Smaller LLMs
von: Zhu, Siyu, et al.
Veröffentlicht: (2025)
von: Zhu, Siyu, et al.
Veröffentlicht: (2025)
Multi-agent transformer-accelerated RL for satisfaction of STL specifications
von: Forsberg, Albin Larsson, et al.
Veröffentlicht: (2024)
von: Forsberg, Albin Larsson, et al.
Veröffentlicht: (2024)
DVM: Towards Controllable LLM Agents in Social Deduction Games
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
PickLLM: Context-Aware RL-Assisted Large Language Model Routing
von: Sikeridis, Dimitrios, et al.
Veröffentlicht: (2024)
von: Sikeridis, Dimitrios, et al.
Veröffentlicht: (2024)
Cooperative Multi-agent RL with Communication Constraints
von: Xiong, Nuoya, et al.
Veröffentlicht: (2026)
von: Xiong, Nuoya, et al.
Veröffentlicht: (2026)
ToolRL: Reward is All Tool Learning Needs
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL
von: Wu, Junyi, et al.
Veröffentlicht: (2026)
von: Wu, Junyi, et al.
Veröffentlicht: (2026)
Distribution-Aware Algorithm Design with LLM Agents
von: Koganti, Saharsh, et al.
Veröffentlicht: (2026)
von: Koganti, Saharsh, et al.
Veröffentlicht: (2026)
Agent^2 RL-Bench: Can LLM Agents Engineer Agentic RL Post-Training?
von: Chen, Wanyi, et al.
Veröffentlicht: (2026)
von: Chen, Wanyi, et al.
Veröffentlicht: (2026)
SkyRL-Agent: Efficient RL Training for Multi-turn LLM Agent
von: Cao, Shiyi, et al.
Veröffentlicht: (2025)
von: Cao, Shiyi, et al.
Veröffentlicht: (2025)
Transformable Gaussian Reward Function for Socially-Aware Navigation with Deep Reinforcement Learning
von: Kim, Jinyeob, et al.
Veröffentlicht: (2024)
von: Kim, Jinyeob, et al.
Veröffentlicht: (2024)
Less is more? Rewards in RL for Cyber Defence
von: Bates, Elizabeth, et al.
Veröffentlicht: (2025)
von: Bates, Elizabeth, et al.
Veröffentlicht: (2025)
Your Reward Function for RL is Your Best PRM for Search: Unifying RL and Search-Based TTS
von: Jin, Can, et al.
Veröffentlicht: (2025)
von: Jin, Can, et al.
Veröffentlicht: (2025)
SPARD: Self-Paced Curriculum for RL Alignment via Integrating Reward Dynamics and Data Utility
von: Zhi, Xuyang, et al.
Veröffentlicht: (2026)
von: Zhi, Xuyang, et al.
Veröffentlicht: (2026)
Natural Emergent Misalignment from Reward Hacking in Production RL
von: MacDiarmid, Monte, et al.
Veröffentlicht: (2025)
von: MacDiarmid, Monte, et al.
Veröffentlicht: (2025)
AgentV-RL: Scaling Reward Modeling with Agentic Verifier
von: Zhang, Jiazheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jiazheng, et al.
Veröffentlicht: (2026)
Text2Touch: Tactile In-Hand Manipulation with LLM-Designed Reward Functions
von: Field, Harrison, et al.
Veröffentlicht: (2025)
von: Field, Harrison, et al.
Veröffentlicht: (2025)
Debate as Reward: A Multi-Agent Reward System for Scientific Ideation via RL Post-Training
von: Salimi, Moein, et al.
Veröffentlicht: (2026)
von: Salimi, Moein, et al.
Veröffentlicht: (2026)
ReviewRL: Towards Automated Scientific Review with RL
von: Zeng, Sihang, et al.
Veröffentlicht: (2025)
von: Zeng, Sihang, et al.
Veröffentlicht: (2025)
Realistic pedestrian-driver interaction modelling using multi-agent RL with human perceptual-motor constraints
von: Wang, Yueyang, et al.
Veröffentlicht: (2025)
von: Wang, Yueyang, et al.
Veröffentlicht: (2025)
MUA-RL: Multi-turn User-interacting Agent Reinforcement Learning for agentic tool use
von: Zhao, Weikang, et al.
Veröffentlicht: (2025)
von: Zhao, Weikang, et al.
Veröffentlicht: (2025)
ProRL Agent: Rollout-as-a-Service for RL Training of Multi-Turn LLM Agents
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
von: Muslimani, Calarina, et al.
Veröffentlicht: (2025) -
On Designing Effective RL Reward at Training Time for LLM Reasoning
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2024) -
When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL
von: Wang, Youting, et al.
Veröffentlicht: (2026) -
Boosting Universal LLM Reward Design through Heuristic Reward Observation Space Evolution
von: Heng, Zen Kit, et al.
Veröffentlicht: (2025) -
FlowRL: Matching Reward Distributions for LLM Reasoning
von: Zhu, Xuekai, et al.
Veröffentlicht: (2025)