HIQL: Offline Goal-Conditioned RL with Latent States as Actions
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Park, Seohong, Ghosh, Dibya, Eysenbach, Benjamin, Levine, Sergey |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
OGBench: Benchmarking Offline Goal-Conditioned RL
par: Park, Seohong, et autres
Publié: (2024)
par: Park, Seohong, et autres
Publié: (2024)
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
par: Park, Seohong, et autres
Publié: (2023)
par: Park, Seohong, et autres
Publié: (2023)
Scalable Offline Model-Based RL with Action Chunks
par: Park, Kwanyoung, et autres
Publié: (2025)
par: Park, Kwanyoung, et autres
Publié: (2025)
Intention-Conditioned Flow Occupancy Models
par: Zheng, Chongyi, et autres
Publié: (2025)
par: Zheng, Chongyi, et autres
Publié: (2025)
Foundation Policies with Hilbert Representations
par: Park, Seohong, et autres
Publié: (2024)
par: Park, Seohong, et autres
Publié: (2024)
Decoupled Q-Chunking
par: Li, Qiyang, et autres
Publié: (2025)
par: Li, Qiyang, et autres
Publié: (2025)
Horizon Reduction Makes RL Scalable
par: Park, Seohong, et autres
Publié: (2025)
par: Park, Seohong, et autres
Publié: (2025)
Is Value Learning Really the Main Bottleneck in Offline RL?
par: Park, Seohong, et autres
Publié: (2024)
par: Park, Seohong, et autres
Publié: (2024)
Dual Goal Representations
par: Park, Seohong, et autres
Publié: (2025)
par: Park, Seohong, et autres
Publié: (2025)
Stabilizing Contrastive RL: Techniques for Robotic Goal Reaching from Offline Data
par: Zheng, Chongyi, et autres
Publié: (2023)
par: Zheng, Chongyi, et autres
Publié: (2023)
ViVa: Video-Trained Value Functions for Guiding Online RL from Diverse Data
par: Dashora, Nitish, et autres
Publié: (2025)
par: Dashora, Nitish, et autres
Publié: (2025)
Transitive RL: Value Learning via Divide and Conquer
par: Park, Seohong, et autres
Publié: (2025)
par: Park, Seohong, et autres
Publié: (2025)
RACER: Epistemic Risk-Sensitive RL Enables Fast Driving with Fewer Crashes
par: Stachowicz, Kyle, et autres
Publié: (2024)
par: Stachowicz, Kyle, et autres
Publié: (2024)
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images
par: Hatch, Kyle B., et autres
Publié: (2024)
par: Hatch, Kyle B., et autres
Publié: (2024)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
par: Huang, Xingshuai, et autres
Publié: (2024)
par: Huang, Xingshuai, et autres
Publié: (2024)
Language-Conditioned Offline RL for Multi-Robot Navigation
par: Morad, Steven, et autres
Publié: (2024)
par: Morad, Steven, et autres
Publié: (2024)
Exploring the Edges of Latent State Clusters for Goal-Conditioned Reinforcement Learning
par: Duan, Yuanlin, et autres
Publié: (2024)
par: Duan, Yuanlin, et autres
Publié: (2024)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
par: Sikchi, Harshit, et autres
Publié: (2023)
par: Sikchi, Harshit, et autres
Publié: (2023)
Flow Q-Learning
par: Park, Seohong, et autres
Publié: (2025)
par: Park, Seohong, et autres
Publié: (2025)
Multistep Quasimetric Learning for Scalable Goal-conditioned Reinforcement Learning
par: Zheng, Bill Chunyuan, et autres
Publié: (2025)
par: Zheng, Bill Chunyuan, et autres
Publié: (2025)
Reinforcement Learning with Action Chunking
par: Li, Qiyang, et autres
Publié: (2025)
par: Li, Qiyang, et autres
Publié: (2025)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
par: Kim, Changyeon, et autres
Publié: (2025)
par: Kim, Changyeon, et autres
Publié: (2025)
Accelerating Goal-Conditioned RL Algorithms and Research
par: Bortkiewicz, Michał, et autres
Publié: (2024)
par: Bortkiewicz, Michał, et autres
Publié: (2024)
Real-Time Execution of Action Chunking Flow Policies
par: Black, Kevin, et autres
Publié: (2025)
par: Black, Kevin, et autres
Publié: (2025)
Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization
par: Modirshanechi, Alireza, et autres
Publié: (2026)
par: Modirshanechi, Alireza, et autres
Publié: (2026)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
par: Wagenmaker, Andrew, et autres
Publié: (2025)
par: Wagenmaker, Andrew, et autres
Publié: (2025)
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
par: Venugopal, Aravind, et autres
Publié: (2026)
par: Venugopal, Aravind, et autres
Publié: (2026)
Can We Really Learn One Representation to Optimize All Rewards?
par: Zheng, Chongyi, et autres
Publié: (2026)
par: Zheng, Chongyi, et autres
Publié: (2026)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
par: Frans, Kevin, et autres
Publié: (2024)
par: Frans, Kevin, et autres
Publié: (2024)
SLAC: Simulation-Pretrained Latent Action Space for Whole-Body Real-World RL
par: Hu, Jiaheng, et autres
Publié: (2025)
par: Hu, Jiaheng, et autres
Publié: (2025)
Efficient Virtuoso: A Latent Diffusion Transformer Model for Goal-Conditioned Trajectory Planning
par: Guillen-Perez, Antonio
Publié: (2025)
par: Guillen-Perez, Antonio
Publié: (2025)
Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy
par: Cao, Chenyang, et autres
Publié: (2024)
par: Cao, Chenyang, et autres
Publié: (2024)
Q-learning with Adjoint Matching
par: Li, Qiyang, et autres
Publié: (2026)
par: Li, Qiyang, et autres
Publié: (2026)
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
par: Liu, Grace, et autres
Publié: (2024)
par: Liu, Grace, et autres
Publié: (2024)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
par: Xiao, Wei, et autres
Publié: (2025)
par: Xiao, Wei, et autres
Publié: (2025)
CtRL-Sim: Reactive and Controllable Driving Agents with Offline Reinforcement Learning
par: Rowe, Luke, et autres
Publié: (2024)
par: Rowe, Luke, et autres
Publié: (2024)
ReFORM: Reflected Flows for On-support Offline RL via Noise Manipulation
par: Zhang, Songyuan, et autres
Publié: (2026)
par: Zhang, Songyuan, et autres
Publié: (2026)
Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL
par: Choi, Jinwoo, et autres
Publié: (2026)
par: Choi, Jinwoo, et autres
Publié: (2026)
H2O+: An Improved Framework for Hybrid Offline-and-Online RL with Dynamics Gaps
par: Niu, Haoyi, et autres
Publié: (2023)
par: Niu, Haoyi, et autres
Publié: (2023)
Behavior Generation with Latent Actions
par: Lee, Seungjae, et autres
Publié: (2024)
par: Lee, Seungjae, et autres
Publié: (2024)
Documents similaires
-
OGBench: Benchmarking Offline Goal-Conditioned RL
par: Park, Seohong, et autres
Publié: (2024) -
METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
par: Park, Seohong, et autres
Publié: (2023) -
Scalable Offline Model-Based RL with Action Chunks
par: Park, Kwanyoung, et autres
Publié: (2025) -
Intention-Conditioned Flow Occupancy Models
par: Zheng, Chongyi, et autres
Publié: (2025) -
Foundation Policies with Hilbert Representations
par: Park, Seohong, et autres
Publié: (2024)