Automated Reward Design for Gran Turismo
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Michel, Seno, Takuma, Subramanian, Kaushik, Wurman, Peter R., Stone, Peter, Sherstan, Craig |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Super-human Vision-based Reinforcement Learning Agent for Autonomous Racing in Gran Turismo
by: Vasco, Miguel, et al.
Published: (2024)
by: Vasco, Miguel, et al.
Published: (2024)
A Champion-level Vision-based Reinforcement Learning Agent for Competitive Racing in Gran Turismo 7
by: Lee, Hojoon, et al.
Published: (2025)
by: Lee, Hojoon, et al.
Published: (2025)
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024)
by: Lee, Hojoon, et al.
Published: (2024)
Out-of-Distribution Generalization with a SPARC: Racing 100 Unseen Vehicles with a Single Policy
by: Grooten, Bram, et al.
Published: (2025)
by: Grooten, Bram, et al.
Published: (2025)
The Trajectory Alignment Coefficient in Two Acts: From Reward Tuning to Reward Learning
by: Muslimani, Calarina, et al.
Published: (2026)
by: Muslimani, Calarina, et al.
Published: (2026)
Adaptive Splitting of Reusable Temporal Monitors for Rare Traffic Violations
by: Innes, Craig, et al.
Published: (2024)
by: Innes, Craig, et al.
Published: (2024)
MusicSynth: An Automated Pipeline for Generating Violin Fingerboard Animations from Sheet Music Using Optical Music Recognition
by: Kaushik, Abhimanyu
Published: (2026)
by: Kaushik, Abhimanyu
Published: (2026)
Minimum Coverage Sets for Training Robust Ad Hoc Teamwork Agents
by: Rahman, Arrasy, et al.
Published: (2023)
by: Rahman, Arrasy, et al.
Published: (2023)
RF-Agent: Automated Reward Function Design via Language Agent Tree Search
by: Gao, Ning, et al.
Published: (2026)
by: Gao, Ning, et al.
Published: (2026)
An Automated Reinforcement Learning Reward Design Framework with Large Language Model for Cooperative Platoon Coordination
by: Wei, Dixiao, et al.
Published: (2025)
by: Wei, Dixiao, et al.
Published: (2025)
A new proof of non-Cohen-Macaulayness of Bertin's example
by: Seno, Takuma
Published: (2025)
by: Seno, Takuma
Published: (2025)
Auto MC-Reward: Automated Dense Reward Design with Large Language Models for Minecraft
by: Li, Hao, et al.
Published: (2023)
by: Li, Hao, et al.
Published: (2023)
Combining Automated Optimisation of Hyperparameters and Reward Shape
by: Dierkes, Julian, et al.
Published: (2024)
by: Dierkes, Julian, et al.
Published: (2024)
Boosting Universal LLM Reward Design through Heuristic Reward Observation Space Evolution
by: Heng, Zen Kit, et al.
Published: (2025)
by: Heng, Zen Kit, et al.
Published: (2025)
Multilinguality in LLM-Designed Reward Functions for Restless Bandits: Effects on Task Performance and Fairness
by: Parthasarathy, Ambreesh, et al.
Published: (2025)
by: Parthasarathy, Ambreesh, et al.
Published: (2025)
GACL: Grounded Adaptive Curriculum Learning with Active Task and Performance Monitoring
by: Wang, Linji, et al.
Published: (2025)
by: Wang, Linji, et al.
Published: (2025)
Grounded Curriculum Learning
by: Wang, Linji, et al.
Published: (2024)
by: Wang, Linji, et al.
Published: (2024)
Deep Reinforcement Learning in Parameterized Action Space
by: Hausknecht, Matthew, et al.
Published: (2015)
by: Hausknecht, Matthew, et al.
Published: (2015)
Automating Reformulation of Essence Specifications via Graph Rewriting
by: Miguel, Ian, et al.
Published: (2024)
by: Miguel, Ian, et al.
Published: (2024)
Learning from Demonstration with Implicit Nonlinear Dynamics Models
by: Fagan, Peter David, et al.
Published: (2024)
by: Fagan, Peter David, et al.
Published: (2024)
Tiered Reward: Designing Rewards for Specification and Fast Learning of Desired Behavior
by: Zhou, Zhiyuan, et al.
Published: (2022)
by: Zhou, Zhiyuan, et al.
Published: (2022)
SLAC: Simulation-Pretrained Latent Action Space for Whole-Body Real-World RL
by: Hu, Jiaheng, et al.
Published: (2025)
by: Hu, Jiaheng, et al.
Published: (2025)
Proto Successor Measure: Representing the Behavior Space of an RL Agent
by: Agarwal, Siddhant, et al.
Published: (2024)
by: Agarwal, Siddhant, et al.
Published: (2024)
Self-Consistency of the Internal Reward Models Improves Self-Rewarding Language Models
by: Zhou, Xin, et al.
Published: (2025)
by: Zhou, Xin, et al.
Published: (2025)
EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models
by: Zhao, Zhikai, et al.
Published: (2026)
by: Zhao, Zhikai, et al.
Published: (2026)
Automated Design of Agentic Systems
by: Hu, Shengran, et al.
Published: (2024)
by: Hu, Shengran, et al.
Published: (2024)
t-DGR: A Trajectory-Based Deep Generative Replay Method for Continual Learning in Decision Making
by: Yue, William, et al.
Published: (2024)
by: Yue, William, et al.
Published: (2024)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
by: Diaz-Bone, Leander, et al.
Published: (2025)
by: Diaz-Bone, Leander, et al.
Published: (2025)
Robust Reward Design for Markov Decision Processes
by: Wu, Shuo, et al.
Published: (2024)
by: Wu, Shuo, et al.
Published: (2024)
L3M+P: Lifelong Planning with Large Language Models
by: Agarwal, Krish, et al.
Published: (2025)
by: Agarwal, Krish, et al.
Published: (2025)
Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
by: Muslimani, Calarina, et al.
Published: (2025)
by: Muslimani, Calarina, et al.
Published: (2025)
Reward-free Alignment for Conflicting Objectives
by: Chen, Peter, et al.
Published: (2026)
by: Chen, Peter, et al.
Published: (2026)
Multi-Agent Collaborative Reward Design for Enhancing Reasoning in Reinforcement Learning
by: Yang, Pei, et al.
Published: (2025)
by: Yang, Pei, et al.
Published: (2025)
Towards Socially and Morally Aware RL agent: Reward Design With LLM
by: Wang, Zhaoyue
Published: (2024)
by: Wang, Zhaoyue
Published: (2024)
Reward Design for Justifiable Sequential Decision-Making
by: Sukovic, Aleksa, et al.
Published: (2024)
by: Sukovic, Aleksa, et al.
Published: (2024)
SECURE: Semantics-aware Embodied Conversation under Unawareness for Lifelong Robot Learning
by: Rubavicius, Rimvydas, et al.
Published: (2024)
by: Rubavicius, Rimvydas, et al.
Published: (2024)
Representation Stability in a Minimal Continual Learning Agent
by: Subramanian, Vishnu
Published: (2026)
by: Subramanian, Vishnu
Published: (2026)
Automated Hybrid Reward Scheduling via Large Language Models for Robotic Skill Learning
by: Huang, Changxin, et al.
Published: (2025)
by: Huang, Changxin, et al.
Published: (2025)
Human-centric Reward Optimization for Reinforcement Learning-based Automated Driving using Large Language Models
by: Zhou, Ziqi, et al.
Published: (2024)
by: Zhou, Ziqi, et al.
Published: (2024)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
Similar Items
-
A Super-human Vision-based Reinforcement Learning Agent for Autonomous Racing in Gran Turismo
by: Vasco, Miguel, et al.
Published: (2024) -
A Champion-level Vision-based Reinforcement Learning Agent for Competitive Racing in Gran Turismo 7
by: Lee, Hojoon, et al.
Published: (2025) -
SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
by: Lee, Hojoon, et al.
Published: (2024) -
Out-of-Distribution Generalization with a SPARC: Racing 100 Unseen Vehicles with a Single Policy
by: Grooten, Bram, et al.
Published: (2025) -
The Trajectory Alignment Coefficient in Two Acts: From Reward Tuning to Reward Learning
by: Muslimani, Calarina, et al.
Published: (2026)