Gespeichert in:
| Hauptverfasser: | Park, Seohong, Kreiman, Tobias, Levine, Sergey |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.15567 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
von: Park, Seohong, et al.
Veröffentlicht: (2023)
von: Park, Seohong, et al.
Veröffentlicht: (2023)
Decoupled Q-Chunking
von: Li, Qiyang, et al.
Veröffentlicht: (2025)
von: Li, Qiyang, et al.
Veröffentlicht: (2025)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023)
von: Park, Seohong, et al.
Veröffentlicht: (2023)
Dual Goal Representations
von: Park, Seohong, et al.
Veröffentlicht: (2025)
von: Park, Seohong, et al.
Veröffentlicht: (2025)
Flow Q-Learning
von: Park, Seohong, et al.
Veröffentlicht: (2025)
von: Park, Seohong, et al.
Veröffentlicht: (2025)
Scalable Offline Model-Based RL with Action Chunks
von: Park, Kwanyoung, et al.
Veröffentlicht: (2025)
von: Park, Kwanyoung, et al.
Veröffentlicht: (2025)
Real-Time Execution of Action Chunking Flow Policies
von: Black, Kevin, et al.
Veröffentlicht: (2025)
von: Black, Kevin, et al.
Veröffentlicht: (2025)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
von: Frans, Kevin, et al.
Veröffentlicht: (2024)
von: Frans, Kevin, et al.
Veröffentlicht: (2024)
Is Value Learning Really the Main Bottleneck in Offline RL?
von: Park, Seohong, et al.
Veröffentlicht: (2024)
von: Park, Seohong, et al.
Veröffentlicht: (2024)
OGBench: Benchmarking Offline Goal-Conditioned RL
von: Park, Seohong, et al.
Veröffentlicht: (2024)
von: Park, Seohong, et al.
Veröffentlicht: (2024)
Transitive RL: Value Learning via Divide and Conquer
von: Park, Seohong, et al.
Veröffentlicht: (2025)
von: Park, Seohong, et al.
Veröffentlicht: (2025)
Intention-Conditioned Flow Occupancy Models
von: Zheng, Chongyi, et al.
Veröffentlicht: (2025)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2025)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
von: Xu, Charles, et al.
Veröffentlicht: (2024)
von: Xu, Charles, et al.
Veröffentlicht: (2024)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
RACER: Epistemic Risk-Sensitive RL Enables Fast Driving with Fewer Crashes
von: Stachowicz, Kyle, et al.
Veröffentlicht: (2024)
von: Stachowicz, Kyle, et al.
Veröffentlicht: (2024)
Q-learning with Adjoint Matching
von: Li, Qiyang, et al.
Veröffentlicht: (2026)
von: Li, Qiyang, et al.
Veröffentlicht: (2026)
Reinforcement Learning with Action Chunking
von: Li, Qiyang, et al.
Veröffentlicht: (2025)
von: Li, Qiyang, et al.
Veröffentlicht: (2025)
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images
von: Hatch, Kyle B., et al.
Veröffentlicht: (2024)
von: Hatch, Kyle B., et al.
Veröffentlicht: (2024)
Steering Your Diffusion Policy with Latent Space Reinforcement Learning
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
von: Wagenmaker, Andrew, et al.
Veröffentlicht: (2025)
Horizon Reduction Makes RL Scalable
von: Park, Seohong, et al.
Veröffentlicht: (2025)
von: Park, Seohong, et al.
Veröffentlicht: (2025)
RAPTOR: A Foundation Policy for Quadrotor Control
von: Eschmann, Jonas, et al.
Veröffentlicht: (2025)
von: Eschmann, Jonas, et al.
Veröffentlicht: (2025)
Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation
von: Feng, Yunhai, et al.
Veröffentlicht: (2025)
von: Feng, Yunhai, et al.
Veröffentlicht: (2025)
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
von: Tang, Grace, et al.
Veröffentlicht: (2024)
von: Tang, Grace, et al.
Veröffentlicht: (2024)
Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow
von: Koo, Juil, et al.
Veröffentlicht: (2026)
von: Koo, Juil, et al.
Veröffentlicht: (2026)
Toward Accurate Long-Horizon Robotic Manipulation: Language-to-Action with Foundation Models via Scene Graphs
von: Dinesh, Sushil Samuel, et al.
Veröffentlicht: (2025)
von: Dinesh, Sushil Samuel, et al.
Veröffentlicht: (2025)
Towards Interpretable Foundation Models of Robot Behavior: A Task Specific Policy Generation Approach
von: Sheidlower, Isaac, et al.
Veröffentlicht: (2024)
von: Sheidlower, Isaac, et al.
Veröffentlicht: (2024)
Yell At Your Robot: Improving On-the-Fly from Language Corrections
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2024)
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2024)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
von: Zhang, Jesse, et al.
Veröffentlicht: (2025)
Language Guided Skill Discovery
von: Rho, Seungeun, et al.
Veröffentlicht: (2024)
von: Rho, Seungeun, et al.
Veröffentlicht: (2024)
Diffusion Guidance Is a Controllable Policy Improvement Operator
von: Frans, Kevin, et al.
Veröffentlicht: (2025)
von: Frans, Kevin, et al.
Veröffentlicht: (2025)
Premier-TACO is a Few-Shot Policy Learner: Pretraining Multitask Representation via Temporal Action-Driven Contrastive Loss
von: Zheng, Ruijie, et al.
Veröffentlicht: (2024)
von: Zheng, Ruijie, et al.
Veröffentlicht: (2024)
Ensembling Prioritized Hybrid Policies for Multi-agent Pathfinding
von: Tang, Huijie, et al.
Veröffentlicht: (2024)
von: Tang, Huijie, et al.
Veröffentlicht: (2024)
An Interactive Agent Foundation Model
von: Durante, Zane, et al.
Veröffentlicht: (2024)
von: Durante, Zane, et al.
Veröffentlicht: (2024)
Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
von: Azeem, Muqsit, et al.
Veröffentlicht: (2024)
von: Azeem, Muqsit, et al.
Veröffentlicht: (2024)
The Ingredients for Robotic Diffusion Transformers
von: Dasari, Sudeep, et al.
Veröffentlicht: (2024)
von: Dasari, Sudeep, et al.
Veröffentlicht: (2024)
Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm
von: Kim, Hyeonjun, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonjun, et al.
Veröffentlicht: (2025)
Online Foundation Model Selection in Robotics
von: Li, Po-han, et al.
Veröffentlicht: (2024)
von: Li, Po-han, et al.
Veröffentlicht: (2024)
Fast Adaptation with Behavioral Foundation Models
von: Sikchi, Harshit, et al.
Veröffentlicht: (2025)
von: Sikchi, Harshit, et al.
Veröffentlicht: (2025)
Policy Decorator: Model-Agnostic Online Refinement for Large Policy Model
von: Yuan, Xiu, et al.
Veröffentlicht: (2024)
von: Yuan, Xiu, et al.
Veröffentlicht: (2024)
Policy-Guided Diffusion
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2024)
von: Jackson, Matthew Thomas, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
von: Park, Seohong, et al.
Veröffentlicht: (2023) -
Decoupled Q-Chunking
von: Li, Qiyang, et al.
Veröffentlicht: (2025) -
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023) -
Dual Goal Representations
von: Park, Seohong, et al.
Veröffentlicht: (2025) -
Flow Q-Learning
von: Park, Seohong, et al.
Veröffentlicht: (2025)