METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
Fuente:
arXiv
Guardado en:
| Autores principales: | Park, Seohong, Rybkin, Oleh, Levine, Sergey |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023)
por: Park, Seohong, et al.
Publicado: (2023)
Foundation Policies with Hilbert Representations
por: Park, Seohong, et al.
Publicado: (2024)
por: Park, Seohong, et al.
Publicado: (2024)
Decoupled Q-Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Scalable Offline Model-Based RL with Action Chunks
por: Park, Kwanyoung, et al.
Publicado: (2025)
por: Park, Kwanyoung, et al.
Publicado: (2025)
Horizon Reduction Makes RL Scalable
por: Park, Seohong, et al.
Publicado: (2025)
por: Park, Seohong, et al.
Publicado: (2025)
OGBench: Benchmarking Offline Goal-Conditioned RL
por: Park, Seohong, et al.
Publicado: (2024)
por: Park, Seohong, et al.
Publicado: (2024)
Transitive RL: Value Learning via Divide and Conquer
por: Park, Seohong, et al.
Publicado: (2025)
por: Park, Seohong, et al.
Publicado: (2025)
Is Value Learning Really the Main Bottleneck in Offline RL?
por: Park, Seohong, et al.
Publicado: (2024)
por: Park, Seohong, et al.
Publicado: (2024)
Privileged Sensing Scaffolds Reinforcement Learning
por: Hu, Edward S., et al.
Publicado: (2024)
por: Hu, Edward S., et al.
Publicado: (2024)
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings
por: Frans, Kevin, et al.
Publicado: (2024)
por: Frans, Kevin, et al.
Publicado: (2024)
RACER: Epistemic Risk-Sensitive RL Enables Fast Driving with Fewer Crashes
por: Stachowicz, Kyle, et al.
Publicado: (2024)
por: Stachowicz, Kyle, et al.
Publicado: (2024)
Dual Goal Representations
por: Park, Seohong, et al.
Publicado: (2025)
por: Park, Seohong, et al.
Publicado: (2025)
Flow Q-Learning
por: Park, Seohong, et al.
Publicado: (2025)
por: Park, Seohong, et al.
Publicado: (2025)
Video2Policy: Scaling up Manipulation Tasks in Simulation through Internet Videos
por: Ye, Weirui, et al.
Publicado: (2025)
por: Ye, Weirui, et al.
Publicado: (2025)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
por: Wagenmaker, Andrew, et al.
Publicado: (2025)
por: Wagenmaker, Andrew, et al.
Publicado: (2025)
Real-World Reinforcement Learning of Active Perception Behaviors
por: Hu, Edward S., et al.
Publicado: (2025)
por: Hu, Edward S., et al.
Publicado: (2025)
Intention-Conditioned Flow Occupancy Models
por: Zheng, Chongyi, et al.
Publicado: (2025)
por: Zheng, Chongyi, et al.
Publicado: (2025)
Q-learning with Adjoint Matching
por: Li, Qiyang, et al.
Publicado: (2026)
por: Li, Qiyang, et al.
Publicado: (2026)
Scalable Decision-Making in Stochastic Environments through Learned Temporal Abstraction
por: Luo, Baiting, et al.
Publicado: (2025)
por: Luo, Baiting, et al.
Publicado: (2025)
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Sample-efficient and Scalable Exploration in Continuous-Time RL
por: Iten, Klemens, et al.
Publicado: (2025)
por: Iten, Klemens, et al.
Publicado: (2025)
CaRL: Learning Scalable Planning Policies with Simple Rewards
por: Jaeger, Bernhard, et al.
Publicado: (2025)
por: Jaeger, Bernhard, et al.
Publicado: (2025)
Real-Time Execution of Action Chunking Flow Policies
por: Black, Kevin, et al.
Publicado: (2025)
por: Black, Kevin, et al.
Publicado: (2025)
Unsupervised-to-Online Reinforcement Learning
por: Kim, Junsu, et al.
Publicado: (2024)
por: Kim, Junsu, et al.
Publicado: (2024)
DEAS: DEtached value learning with Action Sequence for Scalable Offline RL
por: Kim, Changyeon, et al.
Publicado: (2025)
por: Kim, Changyeon, et al.
Publicado: (2025)
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
por: Xu, Charles, et al.
Publicado: (2024)
por: Xu, Charles, et al.
Publicado: (2024)
A Review of Online Diffusion Policy RL Algorithms for Scalable Robotic Control
por: Choi, Wonhyeok, et al.
Publicado: (2026)
por: Choi, Wonhyeok, et al.
Publicado: (2026)
Automatic Environment Shaping is the Next Frontier in RL
por: Park, Younghyo, et al.
Publicado: (2024)
por: Park, Younghyo, et al.
Publicado: (2024)
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images
por: Hatch, Kyle B., et al.
Publicado: (2024)
por: Hatch, Kyle B., et al.
Publicado: (2024)
Latent Diffusion Planning for Imitation Learning
por: Xie, Amber, et al.
Publicado: (2025)
por: Xie, Amber, et al.
Publicado: (2025)
Preference-Conditioned Language-Guided Abstraction
por: Peng, Andi, et al.
Publicado: (2024)
por: Peng, Andi, et al.
Publicado: (2024)
Learning with Language-Guided State Abstractions
por: Peng, Andi, et al.
Publicado: (2024)
por: Peng, Andi, et al.
Publicado: (2024)
Reconciling Spatial and Temporal Abstractions for Goal Representation
por: Zadem, Mehdi, et al.
Publicado: (2024)
por: Zadem, Mehdi, et al.
Publicado: (2024)
Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation
por: Feng, Yunhai, et al.
Publicado: (2025)
por: Feng, Yunhai, et al.
Publicado: (2025)
Spectral Alignment in Forward-Backward Representations via Temporal Abstraction
por: Azad, Seyed Mahdi B., et al.
Publicado: (2026)
por: Azad, Seyed Mahdi B., et al.
Publicado: (2026)
Failure-Aware RL: Reliable Offline-to-Online Reinforcement Learning with Self-Recovery for Real-World Manipulation
por: Li, Huanyu, et al.
Publicado: (2026)
por: Li, Huanyu, et al.
Publicado: (2026)
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
por: Tang, Grace, et al.
Publicado: (2024)
por: Tang, Grace, et al.
Publicado: (2024)
Neuro-Symbolic Imitation Learning: Discovering Symbolic Abstractions for Skill Learning
por: Keller, Leon, et al.
Publicado: (2025)
por: Keller, Leon, et al.
Publicado: (2025)
Towards Scalable & Efficient Interaction-Aware Planning in Autonomous Vehicles using Knowledge Distillation
por: Gupta, Piyush, et al.
Publicado: (2024)
por: Gupta, Piyush, et al.
Publicado: (2024)
Yell At Your Robot: Improving On-the-Fly from Language Corrections
por: Shi, Lucy Xiaoyang, et al.
Publicado: (2024)
por: Shi, Lucy Xiaoyang, et al.
Publicado: (2024)
Ejemplares similares
-
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
por: Park, Seohong, et al.
Publicado: (2023) -
Foundation Policies with Hilbert Representations
por: Park, Seohong, et al.
Publicado: (2024) -
Decoupled Q-Chunking
por: Li, Qiyang, et al.
Publicado: (2025) -
Scalable Offline Model-Based RL with Action Chunks
por: Park, Kwanyoung, et al.
Publicado: (2025) -
Horizon Reduction Makes RL Scalable
por: Park, Seohong, et al.
Publicado: (2025)