Eureka: Human-Level Reward Design via Coding Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Yecheng Jason, Liang, William, Wang, Guanzhi, Huang, De-An, Bastani, Osbert, Jayaraman, Dinesh, Zhu, Yuke, Fan, Linxi, Anandkumar, Anima |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DrEureka: Language Model Guided Sim-To-Real Transfer
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)
Eurekaverse: Environment Curriculum Generation via Large Language Models
von: Liang, William, et al.
Veröffentlicht: (2024)
von: Liang, William, et al.
Veröffentlicht: (2024)
Tether: Autonomous Functional Play with Correspondence-Driven Trajectory Warping
von: Liang, William, et al.
Veröffentlicht: (2026)
von: Liang, William, et al.
Veröffentlicht: (2026)
ARDuP: Active Region Video Diffusion for Universal Policies
von: Huang, Shuaiyi, et al.
Veröffentlicht: (2024)
von: Huang, Shuaiyi, et al.
Veröffentlicht: (2024)
Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models
von: Shi, Junyao, et al.
Veröffentlicht: (2024)
von: Shi, Junyao, et al.
Veröffentlicht: (2024)
Vision Language Models are In-Context Value Learners
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
von: Liu, Yuyang, et al.
Veröffentlicht: (2025)
SCIZOR: A Self-Supervised Approach to Data Curation for Large-Scale Imitation Learning
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot Manipulation
von: Fu, Max, et al.
Veröffentlicht: (2026)
von: Fu, Max, et al.
Veröffentlicht: (2026)
Leveling the Playing Field: Carefully Comparing Classical and Learned Controllers for Quadrotor Trajectory Tracking
von: Kunapuli, Pratik, et al.
Veröffentlicht: (2025)
von: Kunapuli, Pratik, et al.
Veröffentlicht: (2025)
Sim-to-Real Reinforcement Learning for Vision-Based Dexterous Manipulation on Humanoids
von: Lin, Toru, et al.
Veröffentlicht: (2025)
von: Lin, Toru, et al.
Veröffentlicht: (2025)
Rethinking Algorithmic Fairness for Human-AI Collaboration
von: Ge, Haosen, et al.
Veröffentlicht: (2023)
von: Ge, Haosen, et al.
Veröffentlicht: (2023)
Asymptotic Normality of Generalized Low-Rank Matrix Sensing via Riemannian Geometry
von: Bastani, Osbert
Veröffentlicht: (2024)
von: Bastani, Osbert
Veröffentlicht: (2024)
Dynamic Grasping with a Learned Meta-Controller
von: Jia, Yinsen, et al.
Veröffentlicht: (2023)
von: Jia, Yinsen, et al.
Veröffentlicht: (2023)
Task-Oriented Hierarchical Object Decomposition for Visuomotor Control
von: Qian, Jianing, et al.
Veröffentlicht: (2024)
von: Qian, Jianing, et al.
Veröffentlicht: (2024)
EgoScale: Scaling Dexterous Manipulation with Diverse Egocentric Human Data
von: Zheng, Ruijie, et al.
Veröffentlicht: (2026)
von: Zheng, Ruijie, et al.
Veröffentlicht: (2026)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
von: Biza, Ondrej, et al.
Veröffentlicht: (2024)
von: Biza, Ondrej, et al.
Veröffentlicht: (2024)
Recasting Generic Pretrained Vision Transformers As Object-Centric Scene Encoders For Manipulation Policies
von: Qian, Jianing, et al.
Veröffentlicht: (2024)
von: Qian, Jianing, et al.
Veröffentlicht: (2024)
Can Transformers Capture Spatial Relations between Objects?
von: Wen, Chuan, et al.
Veröffentlicht: (2024)
von: Wen, Chuan, et al.
Veröffentlicht: (2024)
Improving Human Sequential Decision-Making with Reinforcement Learning
von: Bastani, Hamsa, et al.
Veröffentlicht: (2021)
von: Bastani, Hamsa, et al.
Veröffentlicht: (2021)
RICL: Adding In-Context Adaptability to Pre-Trained Vision-Language-Action Models
von: Sridhar, Kaustubh, et al.
Veröffentlicht: (2025)
von: Sridhar, Kaustubh, et al.
Veröffentlicht: (2025)
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2025)
von: Seneviratne, Gershom, et al.
Veröffentlicht: (2025)
DexUMI: Using Human Hand as the Universal Manipulation Interface for Dexterous Manipulation
von: Xu, Mengda, et al.
Veröffentlicht: (2025)
von: Xu, Mengda, et al.
Veröffentlicht: (2025)
Privileged Sensing Scaffolds Reinforcement Learning
von: Hu, Edward S., et al.
Veröffentlicht: (2024)
von: Hu, Edward S., et al.
Veröffentlicht: (2024)
HOVER: Versatile Neural Whole-Body Controller for Humanoid Robots
von: He, Tairan, et al.
Veröffentlicht: (2024)
von: He, Tairan, et al.
Veröffentlicht: (2024)
HumanoidMimicGen: Data Generation for Loco-Manipulation via Whole-Body Planning
von: Lin, Kevin, et al.
Veröffentlicht: (2026)
von: Lin, Kevin, et al.
Veröffentlicht: (2026)
DexMimicGen: Automated Data Generation for Bimanual Dexterous Manipulation via Imitation Learning
von: Jiang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Jiang, Zhenyu, et al.
Veröffentlicht: (2024)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2023)
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2023)
FLARE: Robot Learning with Implicit World Modeling
von: Zheng, Ruijie, et al.
Veröffentlicht: (2025)
von: Zheng, Ruijie, et al.
Veröffentlicht: (2025)
Stochastic Online Conformal Prediction with Semi-Bandit Feedback
von: Ge, Haosen, et al.
Veröffentlicht: (2024)
von: Ge, Haosen, et al.
Veröffentlicht: (2024)
Are AI Capabilities Increasing Exponentially? A Competing Hypothesis
von: Ge, Haosen, et al.
Veröffentlicht: (2026)
von: Ge, Haosen, et al.
Veröffentlicht: (2026)
One-Step Diffusion Policy: Fast Visuomotor Policies via Diffusion Distillation
von: Wang, Zhendong, et al.
Veröffentlicht: (2024)
von: Wang, Zhendong, et al.
Veröffentlicht: (2024)
ZeroMimic: Distilling Robotic Manipulation Skills from Web Videos
von: Shi, Junyao, et al.
Veröffentlicht: (2025)
von: Shi, Junyao, et al.
Veröffentlicht: (2025)
CHIP: Adaptive Compliance for Humanoid Control through Hindsight Perturbation
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models
von: Zhao, Zhikai, et al.
Veröffentlicht: (2026)
von: Zhao, Zhikai, et al.
Veröffentlicht: (2026)
Memory-Consistent Neural Networks for Imitation Learning
von: Sridhar, Kaustubh, et al.
Veröffentlicht: (2023)
von: Sridhar, Kaustubh, et al.
Veröffentlicht: (2023)
Leveraging Symmetry to Accelerate Learning of Trajectory Tracking Controllers for Free-Flying Robotic Systems
von: Welde, Jake, et al.
Veröffentlicht: (2024)
von: Welde, Jake, et al.
Veröffentlicht: (2024)
VIRAL: Visual Sim-to-Real at Scale for Humanoid Loco-Manipulation
von: He, Tairan, et al.
Veröffentlicht: (2025)
von: He, Tairan, et al.
Veröffentlicht: (2025)
Vision-based Manipulation from Single Human Video with Open-World Object Graphs
von: Zhu, Yifeng, et al.
Veröffentlicht: (2024)
von: Zhu, Yifeng, et al.
Veröffentlicht: (2024)
Understanding Robot Minds: Leveraging Machine Teaching for Transparent Human-Robot Collaboration Across Diverse Groups
von: Jayaraman, Suresh Kumaar, et al.
Veröffentlicht: (2024)
von: Jayaraman, Suresh Kumaar, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DrEureka: Language Model Guided Sim-To-Real Transfer
von: Ma, Yecheng Jason, et al.
Veröffentlicht: (2024) -
Eurekaverse: Environment Curriculum Generation via Large Language Models
von: Liang, William, et al.
Veröffentlicht: (2024) -
Tether: Autonomous Functional Play with Correspondence-Driven Trajectory Warping
von: Liang, William, et al.
Veröffentlicht: (2026) -
ARDuP: Active Region Video Diffusion for Universal Policies
von: Huang, Shuaiyi, et al.
Veröffentlicht: (2024) -
Composing Pre-Trained Object-Centric Representations for Robotics From "What" and "Where" Foundation Models
von: Shi, Junyao, et al.
Veröffentlicht: (2024)