Average-Reward Maximum Entropy Reinforcement Learning for Global Policy in Double Pendulum Tasks
Fuente:
arXiv
Guardado en:
| Autores principales: | Choe, Jean Seong Bjorn, Choi, Bumkyu, Kim, Jong-kook |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Average-Reward Maximum Entropy Reinforcement Learning for Underactuated Double Pendulum Tasks
por: Choe, Jean Seong Bjorn, et al.
Publicado: (2024)
por: Choe, Jean Seong Bjorn, et al.
Publicado: (2024)
Maximum Entropy On-Policy Actor-Critic via Entropy Advantage Estimation
por: Choe, Jean Seong Bjorn, et al.
Publicado: (2024)
por: Choe, Jean Seong Bjorn, et al.
Publicado: (2024)
Reinforcement Learning for Robust Athletic Intelligence: Lessons from the 2nd 'AI Olympics with RealAIGym' Competition
por: Wiebe, Felix, et al.
Publicado: (2025)
por: Wiebe, Felix, et al.
Publicado: (2025)
Reward-Punishment Reinforcement Learning with Maximum Entropy
por: Wang, Jiexin, et al.
Publicado: (2024)
por: Wang, Jiexin, et al.
Publicado: (2024)
Learning Impact-Rich Rotational Maneuvers via Centroidal Velocity Rewards and Sim-to-Real Techniques: A One-Leg Hopper Flip Case Study
por: Kang, Dongyun, et al.
Publicado: (2025)
por: Kang, Dongyun, et al.
Publicado: (2025)
3D Operation of Autonomous Excavator based on Reinforcement Learning through Independent Reward for Individual Joints
por: Yoo, Yoonkyu, et al.
Publicado: (2024)
por: Yoo, Yoonkyu, et al.
Publicado: (2024)
Beyond Inverted Pendulums: Task-optimal Simple Models of Legged Locomotion
por: Chen, Yu-Ming, et al.
Publicado: (2023)
por: Chen, Yu-Ming, et al.
Publicado: (2023)
Task-Oriented Grasping Using Reinforcement Learning with a Contextual Reward Machine
por: Li, Hui, et al.
Publicado: (2025)
por: Li, Hui, et al.
Publicado: (2025)
Logic-based Task Representation and Reward Shaping in Multiagent Reinforcement Learning
por: Doshi, Nishant
Publicado: (2025)
por: Doshi, Nishant
Publicado: (2025)
Real-Time Model Predictive Control for the Swing-Up Problem of an Underactuated Double Pendulum
por: Burchard, Blanka, et al.
Publicado: (2025)
por: Burchard, Blanka, et al.
Publicado: (2025)
Swing-Up of a Weakly Actuated Double Pendulum via Nonlinear Normal Modes
por: Sachtler, Arne, et al.
Publicado: (2024)
por: Sachtler, Arne, et al.
Publicado: (2024)
Onboard MuJoCo-based Model Predictive Control for Shipboard Crane with Double-Pendulum Sway Suppression
por: Pang, Oscar, et al.
Publicado: (2026)
por: Pang, Oscar, et al.
Publicado: (2026)
Generating Realistic Arm Movements in Reinforcement Learning: A Quantitative Comparison of Reward Terms and Task Requirements
por: Charaja, Jhon P. F., et al.
Publicado: (2024)
por: Charaja, Jhon P. F., et al.
Publicado: (2024)
Gen-Drive: Enhancing Diffusion Generative Driving Policies with Reward Modeling and Reinforcement Learning Fine-tuning
por: Huang, Zhiyu, et al.
Publicado: (2024)
por: Huang, Zhiyu, et al.
Publicado: (2024)
Decoupling Task and Behavior: A Two-Stage Reward Curriculum in Reinforcement Learning for Robotics
por: Freitag, Kilian, et al.
Publicado: (2026)
por: Freitag, Kilian, et al.
Publicado: (2026)
Transformable Gaussian Reward Function for Socially-Aware Navigation with Deep Reinforcement Learning
por: Kim, Jinyeob, et al.
Publicado: (2024)
por: Kim, Jinyeob, et al.
Publicado: (2024)
LaMOuR: Leveraging Language Models for Out-of-Distribution Recovery in Reinforcement Learning
por: Kim, Chan, et al.
Publicado: (2025)
por: Kim, Chan, et al.
Publicado: (2025)
A Visual Reinforcement Learning-Based Separate Primitive Policy for Peg-in-Hole Tasks
por: Xu, Zichun, et al.
Publicado: (2025)
por: Xu, Zichun, et al.
Publicado: (2025)
Design of a 3-DOF Hopping Robot with an Optimized Gearbox: An Intermediate Platform Toward Bipedal Robots
por: Choe, JongHun, et al.
Publicado: (2025)
por: Choe, JongHun, et al.
Publicado: (2025)
CRAFT: Coaching Reinforcement Learning Autonomously using Foundation Models for Multi-Robot Coordination Tasks
por: Choi, Seoyeon, et al.
Publicado: (2025)
por: Choi, Seoyeon, et al.
Publicado: (2025)
Design of Reward Function on Reinforcement Learning for Automated Driving
por: Goto, Takeru, et al.
Publicado: (2025)
por: Goto, Takeru, et al.
Publicado: (2025)
Low-cost Real-world Implementation of the Swing-up Pendulum for Deep Reinforcement Learning Experiments
por: Böhm, Peter, et al.
Publicado: (2025)
por: Böhm, Peter, et al.
Publicado: (2025)
Aligning Humans and Robots via Reinforcement Learning from Implicit Human Feedback
por: Kim, Suzie, et al.
Publicado: (2025)
por: Kim, Suzie, et al.
Publicado: (2025)
Reward Training Wheels: Adaptive Auxiliary Rewards for Robotics Reinforcement Learning
por: Wang, Linji, et al.
Publicado: (2025)
por: Wang, Linji, et al.
Publicado: (2025)
Evaluating MEDIRL: A Replication and Ablation Study of Maximum Entropy Deep Inverse Reinforcement Learning for Human Social Navigation
por: Gupta, Vinay, et al.
Publicado: (2024)
por: Gupta, Vinay, et al.
Publicado: (2024)
Robust Single Rotation Averaging Revisited
por: Lee, Seong Hun, et al.
Publicado: (2023)
por: Lee, Seong Hun, et al.
Publicado: (2023)
Reward-Augmented Reinforcement Learning for Continuous Control in Precision Autonomous Parking via Policy Optimization Methods
por: Suleman, Ahmad, et al.
Publicado: (2025)
por: Suleman, Ahmad, et al.
Publicado: (2025)
Task-priority Intermediated Hierarchical Distributed Policies: Reinforcement Learning of Adaptive Multi-robot Cooperative Transport
por: Naito, Yusei, et al.
Publicado: (2024)
por: Naito, Yusei, et al.
Publicado: (2024)
Curriculum Reinforcement Learning for Complex Reward Functions
por: Freitag, Kilian, et al.
Publicado: (2024)
por: Freitag, Kilian, et al.
Publicado: (2024)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
por: Biza, Ondrej, et al.
Publicado: (2024)
por: Biza, Ondrej, et al.
Publicado: (2024)
Stage-Wise Reward Shaping for Acrobatic Robots: A Constrained Multi-Objective Reinforcement Learning Approach
por: Kim, Dohyeong, et al.
Publicado: (2024)
por: Kim, Dohyeong, et al.
Publicado: (2024)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
por: Ishihara, Yu, et al.
Publicado: (2025)
por: Ishihara, Yu, et al.
Publicado: (2025)
Wavelet Policy: Lifting Scheme for Policy Learning in Long-Horizon Tasks
por: Huang, Hao, et al.
Publicado: (2025)
por: Huang, Hao, et al.
Publicado: (2025)
Robotic Skill Diversification via Active Mutation of Reward Functions in Reinforcement Learning During a Liquid Pouring Task
por: van Buuren, Jannick, et al.
Publicado: (2025)
por: van Buuren, Jannick, et al.
Publicado: (2025)
Driving Beyond Privilege: Distilling Dense-Reward Knowledge into Sparse-Reward Policies
por: Khanzada, Feeza Khan, et al.
Publicado: (2025)
por: Khanzada, Feeza Khan, et al.
Publicado: (2025)
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks
por: Joshi, Viraj, et al.
Publicado: (2025)
por: Joshi, Viraj, et al.
Publicado: (2025)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
por: Vasan, Gautham, et al.
Publicado: (2024)
por: Vasan, Gautham, et al.
Publicado: (2024)
Learning from Demonstration with Hierarchical Policy Abstractions Toward High-Performance and Courteous Autonomous Racing
por: Chung, Chanyoung, et al.
Publicado: (2024)
por: Chung, Chanyoung, et al.
Publicado: (2024)
Recovering Hidden Reward in Diffusion-Based Policies
por: Ji, Yanbiao, et al.
Publicado: (2026)
por: Ji, Yanbiao, et al.
Publicado: (2026)
Real-Time Reinforcement Learning for Dynamic Tasks with a Parallel Soft Robot
por: Avtges, James, et al.
Publicado: (2025)
por: Avtges, James, et al.
Publicado: (2025)
Ejemplares similares
-
Average-Reward Maximum Entropy Reinforcement Learning for Underactuated Double Pendulum Tasks
por: Choe, Jean Seong Bjorn, et al.
Publicado: (2024) -
Maximum Entropy On-Policy Actor-Critic via Entropy Advantage Estimation
por: Choe, Jean Seong Bjorn, et al.
Publicado: (2024) -
Reinforcement Learning for Robust Athletic Intelligence: Lessons from the 2nd 'AI Olympics with RealAIGym' Competition
por: Wiebe, Felix, et al.
Publicado: (2025) -
Reward-Punishment Reinforcement Learning with Maximum Entropy
por: Wang, Jiexin, et al.
Publicado: (2024) -
Learning Impact-Rich Rotational Maneuvers via Centroidal Velocity Rewards and Sim-to-Real Techniques: A One-Leg Hopper Flip Case Study
por: Kang, Dongyun, et al.
Publicado: (2025)