Saved in:
| Main Authors: | Nguyen, Phu, Polani, Daniel, Tiomkin, Stas |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2411.13613 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Decentralized Traffic Flow Optimization Through Intrinsic Motivation
by: Papala, Himaja, et al.
Published: (2025)
by: Papala, Himaja, et al.
Published: (2025)
Acoustic Wave Manipulation Through Sparse Robotic Actuation
by: Shah, Tristan, et al.
Published: (2025)
by: Shah, Tristan, et al.
Published: (2025)
Multi-Agent Empowerment and Emergence of Complex Behavior in Groups
by: Shah, Tristan, et al.
Published: (2026)
by: Shah, Tristan, et al.
Published: (2026)
Learning telic-controllable state representations
by: Amir, Nadav, et al.
Published: (2024)
by: Amir, Nadav, et al.
Published: (2024)
Emergence of Physical Intelligence via Controllable Information Production
by: Shah, Tristan, et al.
Published: (2026)
by: Shah, Tristan, et al.
Published: (2026)
Bootstrapped Reward Shaping
by: Adamczyk, Jacob, et al.
Published: (2025)
by: Adamczyk, Jacob, et al.
Published: (2025)
Average-Reward Soft Actor-Critic
by: Adamczyk, Jacob, et al.
Published: (2025)
by: Adamczyk, Jacob, et al.
Published: (2025)
Goals and the Structure of Experience
by: Amir, Nadav, et al.
Published: (2025)
by: Amir, Nadav, et al.
Published: (2025)
Reward Training Wheels: Adaptive Auxiliary Rewards for Robotics Reinforcement Learning
by: Wang, Linji, et al.
Published: (2025)
by: Wang, Linji, et al.
Published: (2025)
Boosting Soft Q-Learning by Bounding
by: Adamczyk, Jacob, et al.
Published: (2024)
by: Adamczyk, Jacob, et al.
Published: (2024)
EVAL: EigenVector-based Average-reward Learning
by: Adamczyk, Jacob, et al.
Published: (2025)
by: Adamczyk, Jacob, et al.
Published: (2025)
HALO: Learning Human-Robot Collaboration via Heterogeneous-Agent Lyapunov Policy Optimization
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
CLIP-Motion: Learning Reward Functions for Robotic Actions Using Consecutive Observations
by: Dang, Xuzhe, et al.
Published: (2023)
by: Dang, Xuzhe, et al.
Published: (2023)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
by: Wu, Yanru, et al.
Published: (2026)
by: Wu, Yanru, et al.
Published: (2026)
Robot Policy Learning with Temporal Optimal Transport Reward
by: Fu, Yuwei, et al.
Published: (2024)
by: Fu, Yuwei, et al.
Published: (2024)
Learning Reward for Robot Skills Using Large Language Models via Self-Alignment
by: Zeng, Yuwei, et al.
Published: (2024)
by: Zeng, Yuwei, et al.
Published: (2024)
Automated Hybrid Reward Scheduling via Large Language Models for Robotic Skill Learning
by: Huang, Changxin, et al.
Published: (2025)
by: Huang, Changxin, et al.
Published: (2025)
Multi-Resolution Diffusion for Privacy-Sensitive Recommender Systems
by: Lilienthal, Derek, et al.
Published: (2023)
by: Lilienthal, Derek, et al.
Published: (2023)
Stage-Wise Reward Shaping for Acrobatic Robots: A Constrained Multi-Objective Reinforcement Learning Approach
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery
by: Tran, Chi-Nguyen, et al.
Published: (2026)
by: Tran, Chi-Nguyen, et al.
Published: (2026)
ManeuverNet: A Soft Actor-Critic Framework for Precise Maneuvering of Double-Ackermann-Steering Robots with Optimized Reward Functions
by: Deflesselle, Kohio, et al.
Published: (2026)
by: Deflesselle, Kohio, et al.
Published: (2026)
Learning to Recover: Dynamic Reward Shaping with Wheel-Leg Coordination for Fallen Robots
by: Deng, Boyuan, et al.
Published: (2025)
by: Deng, Boyuan, et al.
Published: (2025)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
by: Patel, Bhrij, et al.
Published: (2023)
by: Patel, Bhrij, et al.
Published: (2023)
Actor-Critic for Continuous Action Chunks: A Reinforcement Learning Framework for Long-Horizon Robotic Manipulation with Sparse Reward
by: Yang, Jiarui, et al.
Published: (2025)
by: Yang, Jiarui, et al.
Published: (2025)
Representation Alignment from Human Feedback for Cross-Embodiment Reward Learning from Mixed-Quality Demonstrations
by: Mattson, Connor, et al.
Published: (2024)
by: Mattson, Connor, et al.
Published: (2024)
EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models
by: Zhao, Zhikai, et al.
Published: (2026)
by: Zhao, Zhikai, et al.
Published: (2026)
Optimizing Crowd-Aware Multi-Agent Path Finding through Local Communication with Graph Neural Networks
by: Pham, Phu, et al.
Published: (2023)
by: Pham, Phu, et al.
Published: (2023)
RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models
by: Wu, Hao, et al.
Published: (2026)
by: Wu, Hao, et al.
Published: (2026)
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
by: Wang, Ruixiang, et al.
Published: (2026)
by: Wang, Ruixiang, et al.
Published: (2026)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
by: Chen, Shirui, et al.
Published: (2026)
by: Chen, Shirui, et al.
Published: (2026)
Not Only Rewards But Also Constraints: Applications on Legged Robot Locomotion
by: Kim, Yunho, et al.
Published: (2023)
by: Kim, Yunho, et al.
Published: (2023)
ARCap: Collecting High-quality Human Demonstrations for Robot Learning with Augmented Reality Feedback
by: Chen, Sirui, et al.
Published: (2024)
by: Chen, Sirui, et al.
Published: (2024)
Reward Learning from Suboptimal Demonstrations with Applications in Surgical Electrocautery
by: Karimi, Zohre, et al.
Published: (2024)
by: Karimi, Zohre, et al.
Published: (2024)
Causally Robust Reward Learning from Reason-Augmented Preference Feedback
by: Hwang, Minjune, et al.
Published: (2026)
by: Hwang, Minjune, et al.
Published: (2026)
CaRL: Learning Scalable Planning Policies with Simple Rewards
by: Jaeger, Bernhard, et al.
Published: (2025)
by: Jaeger, Bernhard, et al.
Published: (2025)
Neural-Network-Driven Reward Prediction as a Heuristic: Advancing Q-Learning for Mobile Robot Path Planning
by: Ji, Yiming, et al.
Published: (2024)
by: Ji, Yiming, et al.
Published: (2024)
OPAL: Encoding Causal Understanding of Physical Systems for Robot Learning
by: Tcheurekdjian, Daniel, et al.
Published: (2025)
by: Tcheurekdjian, Daniel, et al.
Published: (2025)
Reward-Centered ReST-MCTS: A Robust Decision-Making Framework for Robotic Manipulation in High Uncertainty Environments
by: Wang, Xibai
Published: (2025)
by: Wang, Xibai
Published: (2025)
Similar Items
-
Decentralized Traffic Flow Optimization Through Intrinsic Motivation
by: Papala, Himaja, et al.
Published: (2025) -
Acoustic Wave Manipulation Through Sparse Robotic Actuation
by: Shah, Tristan, et al.
Published: (2025) -
Multi-Agent Empowerment and Emergence of Complex Behavior in Groups
by: Shah, Tristan, et al.
Published: (2026) -
Learning telic-controllable state representations
by: Amir, Nadav, et al.
Published: (2024) -
Emergence of Physical Intelligence via Controllable Information Production
by: Shah, Tristan, et al.
Published: (2026)