Stabilizing Contrastive RL: Techniques for Robotic Goal Reaching from Offline Data
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Chongyi, Eysenbach, Benjamin, Walke, Homer, Yin, Patrick, Fang, Kuan, Salakhutdinov, Ruslan, Levine, Sergey |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Contrastive Difference Predictive Coding
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
OGBench: Benchmarking Offline Goal-Conditioned RL
von: Park, Seohong, et al.
Veröffentlicht: (2024)
von: Park, Seohong, et al.
Veröffentlicht: (2024)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023)
von: Park, Seohong, et al.
Veröffentlicht: (2023)
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
von: Tang, Grace, et al.
Veröffentlicht: (2024)
von: Tang, Grace, et al.
Veröffentlicht: (2024)
Intention-Conditioned Flow Occupancy Models
von: Zheng, Chongyi, et al.
Veröffentlicht: (2025)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2025)
Inference via Interpolation: Contrastive Representations Provably Enable Planning and Inference
von: Eysenbach, Benjamin, et al.
Veröffentlicht: (2024)
von: Eysenbach, Benjamin, et al.
Veröffentlicht: (2024)
Planning without Search: Refining Frontier LLMs with Offline Goal-Conditioned RL
von: Hong, Joey, et al.
Veröffentlicht: (2025)
von: Hong, Joey, et al.
Veröffentlicht: (2025)
1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities
von: Wang, Kevin, et al.
Veröffentlicht: (2025)
von: Wang, Kevin, et al.
Veröffentlicht: (2025)
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
von: Liu, Grace, et al.
Veröffentlicht: (2024)
von: Liu, Grace, et al.
Veröffentlicht: (2024)
Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration
von: Nimonkar, Chirayu, et al.
Veröffentlicht: (2025)
von: Nimonkar, Chirayu, et al.
Veröffentlicht: (2025)
Horizon Reduction Makes RL Scalable
von: Park, Seohong, et al.
Veröffentlicht: (2025)
von: Park, Seohong, et al.
Veröffentlicht: (2025)
Can We Really Learn One Representation to Optimize All Rewards?
von: Zheng, Chongyi, et al.
Veröffentlicht: (2026)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2026)
Autonomous Improvement of Instruction Following Skills via Foundation Models
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2024)
Consistent Zero-Shot Imitation with Contrastive Goal Inference
von: Wantlin, Kathryn, et al.
Veröffentlicht: (2025)
von: Wantlin, Kathryn, et al.
Veröffentlicht: (2025)
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
von: Zheng, Chongyi, et al.
Veröffentlicht: (2024)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2024)
Accelerating Diffusion Planners in Offline RL via Reward-Aware Consistency Trajectory Distillation
von: Duan, Xintong, et al.
Veröffentlicht: (2025)
von: Duan, Xintong, et al.
Veröffentlicht: (2025)
Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization
von: Modirshanechi, Alireza, et al.
Veröffentlicht: (2026)
von: Modirshanechi, Alireza, et al.
Veröffentlicht: (2026)
Offline Goal-conditioned Reinforcement Learning with Quasimetric Representations
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
Accelerating Goal-Conditioned RL Algorithms and Research
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2024)
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2024)
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
Is Value Learning Really the Main Bottleneck in Offline RL?
von: Park, Seohong, et al.
Veröffentlicht: (2024)
von: Park, Seohong, et al.
Veröffentlicht: (2024)
Scalable Offline Model-Based RL with Action Chunks
von: Park, Kwanyoung, et al.
Veröffentlicht: (2025)
von: Park, Kwanyoung, et al.
Veröffentlicht: (2025)
Value Flows
von: Dong, Perry, et al.
Veröffentlicht: (2025)
von: Dong, Perry, et al.
Veröffentlicht: (2025)
UpSkill: Mutual Information Skill Learning for Structured Response Diversity in LLMs
von: Shah, Devan, et al.
Veröffentlicht: (2026)
von: Shah, Devan, et al.
Veröffentlicht: (2026)
MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting
von: Liu, Fangchen, et al.
Veröffentlicht: (2024)
von: Liu, Fangchen, et al.
Veröffentlicht: (2024)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2025)
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2025)
Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL
von: Choi, Jinwoo, et al.
Veröffentlicht: (2026)
von: Choi, Jinwoo, et al.
Veröffentlicht: (2026)
Training LLM Agents to Empower Humans
von: Ellis, Evan, et al.
Veröffentlicht: (2025)
von: Ellis, Evan, et al.
Veröffentlicht: (2025)
Offline RL for Adaptive Policy Retrieval in Prior Authorization
von: Sharifullin, Ruslan, et al.
Veröffentlicht: (2026)
von: Sharifullin, Ruslan, et al.
Veröffentlicht: (2026)
Dual Goal Representations
von: Park, Seohong, et al.
Veröffentlicht: (2025)
von: Park, Seohong, et al.
Veröffentlicht: (2025)
Learning to Assist Humans without Inferring Rewards
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
Effective Data Augmentation With Diffusion Models
von: Trabucco, Brandon, et al.
Veröffentlicht: (2023)
von: Trabucco, Brandon, et al.
Veröffentlicht: (2023)
Stitching Sub-Trajectories with Conditional Diffusion Model for Goal-Conditioned Offline RL
von: Kim, Sungyoon, et al.
Veröffentlicht: (2024)
von: Kim, Sungyoon, et al.
Veröffentlicht: (2024)
ViVa: Video-Trained Value Functions for Guiding Online RL from Diverse Data
von: Dashora, Nitish, et al.
Veröffentlicht: (2025)
von: Dashora, Nitish, et al.
Veröffentlicht: (2025)
Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2024)
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2024)
Offline Materials Optimization with CliqueFlowmer
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2026)
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2026)
Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2023)
von: Nakamoto, Mitsuhiko, et al.
Veröffentlicht: (2023)
Exploiting Local Dynamics Regularity for Reusable Skills in Offline Hierarchical RL
von: Dayal, Sarthak, et al.
Veröffentlicht: (2026)
von: Dayal, Sarthak, et al.
Veröffentlicht: (2026)
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
von: Venugopal, Aravind, et al.
Veröffentlicht: (2026)
von: Venugopal, Aravind, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Contrastive Difference Predictive Coding
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023) -
OGBench: Benchmarking Offline Goal-Conditioned RL
von: Park, Seohong, et al.
Veröffentlicht: (2024) -
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023) -
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
von: Myers, Vivek, et al.
Veröffentlicht: (2024) -
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
von: Tang, Grace, et al.
Veröffentlicht: (2024)