Consistent Zero-Shot Imitation with Contrastive Goal Inference
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wantlin, Kathryn, Zheng, Chongyi, Eysenbach, Benjamin |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Contrastive Difference Predictive Coding
par: Zheng, Chongyi, et autres
Publié: (2023)
par: Zheng, Chongyi, et autres
Publié: (2023)
Stabilizing Contrastive RL: Techniques for Robotic Goal Reaching from Offline Data
par: Zheng, Chongyi, et autres
Publié: (2023)
par: Zheng, Chongyi, et autres
Publié: (2023)
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
par: Venugopal, Aravind, et autres
Publié: (2026)
par: Venugopal, Aravind, et autres
Publié: (2026)
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
par: Myers, Vivek, et autres
Publié: (2024)
par: Myers, Vivek, et autres
Publié: (2024)
Can We Really Learn One Representation to Optimize All Rewards?
par: Zheng, Chongyi, et autres
Publié: (2026)
par: Zheng, Chongyi, et autres
Publié: (2026)
BuilderBench: The Building Blocks of Intelligent Agents
par: Ghugare, Raj, et autres
Publié: (2025)
par: Ghugare, Raj, et autres
Publié: (2025)
Intention-Conditioned Flow Occupancy Models
par: Zheng, Chongyi, et autres
Publié: (2025)
par: Zheng, Chongyi, et autres
Publié: (2025)
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
par: Zheng, Chongyi, et autres
Publié: (2024)
par: Zheng, Chongyi, et autres
Publié: (2024)
Inference via Interpolation: Contrastive Representations Provably Enable Planning and Inference
par: Eysenbach, Benjamin, et autres
Publié: (2024)
par: Eysenbach, Benjamin, et autres
Publié: (2024)
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
par: Liu, Grace, et autres
Publié: (2024)
par: Liu, Grace, et autres
Publié: (2024)
The "Law" of the Unconscious Contrastive Learner: Probabilistic Alignment of Unpaired Modalities
par: Che, Yongwei, et autres
Publié: (2025)
par: Che, Yongwei, et autres
Publié: (2025)
Value Flows
par: Dong, Perry, et autres
Publié: (2025)
par: Dong, Perry, et autres
Publié: (2025)
UpSkill: Mutual Information Skill Learning for Structured Response Diversity in LLMs
par: Shah, Devan, et autres
Publié: (2026)
par: Shah, Devan, et autres
Publié: (2026)
Offline Goal-conditioned Reinforcement Learning with Quasimetric Representations
par: Myers, Vivek, et autres
Publié: (2025)
par: Myers, Vivek, et autres
Publié: (2025)
Multistep Quasimetric Learning for Scalable Goal-conditioned Reinforcement Learning
par: Zheng, Bill Chunyuan, et autres
Publié: (2025)
par: Zheng, Bill Chunyuan, et autres
Publié: (2025)
OGBench: Benchmarking Offline Goal-Conditioned RL
par: Park, Seohong, et autres
Publié: (2024)
par: Park, Seohong, et autres
Publié: (2024)
Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration
par: Nimonkar, Chirayu, et autres
Publié: (2025)
par: Nimonkar, Chirayu, et autres
Publié: (2025)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
par: Park, Seohong, et autres
Publié: (2023)
par: Park, Seohong, et autres
Publié: (2023)
Demystifying the Mechanisms Behind Emergent Exploration in Goal-conditioned RL
par: Bastankhah, Mahsa, et autres
Publié: (2025)
par: Bastankhah, Mahsa, et autres
Publié: (2025)
Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization
par: Modirshanechi, Alireza, et autres
Publié: (2026)
par: Modirshanechi, Alireza, et autres
Publié: (2026)
Normalizing Flows are Capable Models for RL
par: Ghugare, Raj, et autres
Publié: (2025)
par: Ghugare, Raj, et autres
Publié: (2025)
Contrastive Representations for Temporal Reasoning
par: Ziarko, Alicja, et autres
Publié: (2025)
par: Ziarko, Alicja, et autres
Publié: (2025)
Zero-Shot Offline Imitation Learning via Optimal Transport
par: Rupf, Thomas, et autres
Publié: (2024)
par: Rupf, Thomas, et autres
Publié: (2024)
1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities
par: Wang, Kevin, et autres
Publié: (2025)
par: Wang, Kevin, et autres
Publié: (2025)
Accelerating Goal-Conditioned RL Algorithms and Research
par: Bortkiewicz, Michał, et autres
Publié: (2024)
par: Bortkiewicz, Michał, et autres
Publié: (2024)
Behavior-Consistent Deep Reinforcement Learning
par: Hussing, Marcel, et autres
Publié: (2026)
par: Hussing, Marcel, et autres
Publié: (2026)
Learning to Perceive the World Through Control: Empowerment-Based Representation Learning
par: Bastankhah, Mahsa, et autres
Publié: (2026)
par: Bastankhah, Mahsa, et autres
Publié: (2026)
Generalized Animal Imitator: Agile Locomotion with Versatile Motion Prior
par: Yang, Ruihan, et autres
Publié: (2023)
par: Yang, Ruihan, et autres
Publié: (2023)
Cycle-Consistent Helmholtz Machine: Goal-Seeded Simulation via Inverted Inference
par: Li, Xin
Publié: (2025)
par: Li, Xin
Publié: (2025)
EmbodiSwap for Zero-Shot Robot Imitation Learning
par: Dessalene, Eadom, et autres
Publié: (2025)
par: Dessalene, Eadom, et autres
Publié: (2025)
Horizon Generalization in Reinforcement Learning
par: Myers, Vivek, et autres
Publié: (2025)
par: Myers, Vivek, et autres
Publié: (2025)
On the Role of Iterative Computation in Reinforcement Learning
par: Ghugare, Raj, et autres
Publié: (2026)
par: Ghugare, Raj, et autres
Publié: (2026)
Closing the Gap between TD Learning and Supervised Learning -- A Generalisation Point of View
par: Ghugare, Raj, et autres
Publié: (2024)
par: Ghugare, Raj, et autres
Publié: (2024)
On the Problem of Consistent Anomalies in Zero-Shot Anomaly Detection
par: Le-Gia, Tai
Publié: (2025)
par: Le-Gia, Tai
Publié: (2025)
Zero-Shot Scalable Resilience in UAV Swarms: A Decentralized Imitation Learning Framework with Physics-Informed Graph Interactions
par: Lin, Huan, et autres
Publié: (2026)
par: Lin, Huan, et autres
Publié: (2026)
Deconfounding Imitation Learning with Variational Inference
par: Vuorio, Risto, et autres
Publié: (2022)
par: Vuorio, Risto, et autres
Publié: (2022)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
par: Bortkiewicz, Michał, et autres
Publié: (2025)
par: Bortkiewicz, Michał, et autres
Publié: (2025)
A Rate-Distortion View of Uncertainty Quantification
par: Apostolopoulou, Ifigeneia, et autres
Publié: (2024)
par: Apostolopoulou, Ifigeneia, et autres
Publié: (2024)
Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
par: Mohamed, Faisal, et autres
Publié: (2026)
par: Mohamed, Faisal, et autres
Publié: (2026)
Policy Adaptation via Language Optimization: Decomposing Tasks for Few-Shot Imitation
par: Myers, Vivek, et autres
Publié: (2024)
par: Myers, Vivek, et autres
Publié: (2024)
Documents similaires
-
Contrastive Difference Predictive Coding
par: Zheng, Chongyi, et autres
Publié: (2023) -
Stabilizing Contrastive RL: Techniques for Robotic Goal Reaching from Offline Data
par: Zheng, Chongyi, et autres
Publié: (2023) -
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning
par: Venugopal, Aravind, et autres
Publié: (2026) -
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
par: Myers, Vivek, et autres
Publié: (2024) -
Can We Really Learn One Representation to Optimize All Rewards?
par: Zheng, Chongyi, et autres
Publié: (2026)