Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mohamed, Faisal, Ji, Catherine, Eysenbach, Benjamin, Berseth, Glen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Is Exploration or Optimization the Problem for Deep Reinforcement Learning?
von: Berseth, Glen
Veröffentlicht: (2025)
von: Berseth, Glen
Veröffentlicht: (2025)
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
von: Hugessen, Adriana, et al.
Veröffentlicht: (2024)
von: Hugessen, Adriana, et al.
Veröffentlicht: (2024)
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
von: Liu, Grace, et al.
Veröffentlicht: (2024)
von: Liu, Grace, et al.
Veröffentlicht: (2024)
Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration
von: Nimonkar, Chirayu, et al.
Veröffentlicht: (2025)
von: Nimonkar, Chirayu, et al.
Veröffentlicht: (2025)
Horizon Generalization in Reinforcement Learning
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
von: Myers, Vivek, et al.
Veröffentlicht: (2025)
Can We Really Learn One Representation to Optimize All Rewards?
von: Zheng, Chongyi, et al.
Veröffentlicht: (2026)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2026)
Improving Deep Reinforcement Learning by Reducing the Chain Effect of Value and Policy Churn
von: Tang, Hongyao, et al.
Veröffentlicht: (2024)
von: Tang, Hongyao, et al.
Veröffentlicht: (2024)
Improving Intrinsic Exploration by Creating Stationary Objectives
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2023)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2023)
Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning
von: Lawson, Daniel, et al.
Veröffentlicht: (2025)
von: Lawson, Daniel, et al.
Veröffentlicht: (2025)
Contrastive Representations for Temporal Reasoning
von: Ziarko, Alicja, et al.
Veröffentlicht: (2025)
von: Ziarko, Alicja, et al.
Veröffentlicht: (2025)
Learning to Assist Humans without Inferring Rewards
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
Closing the Gap between TD Learning and Supervised Learning -- A Generalisation Point of View
von: Ghugare, Raj, et al.
Veröffentlicht: (2024)
von: Ghugare, Raj, et al.
Veröffentlicht: (2024)
SegDAC: Visual Generalization in Reinforcement Learning via Dynamic Object Tokens
von: Brown, Alexandre, et al.
Veröffentlicht: (2025)
von: Brown, Alexandre, et al.
Veröffentlicht: (2025)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2025)
von: Bortkiewicz, Michał, et al.
Veröffentlicht: (2025)
Enabling Realtime Reinforcement Learning at Scale with Staggered Asynchronous Inference
von: Riemer, Matthew, et al.
Veröffentlicht: (2024)
von: Riemer, Matthew, et al.
Veröffentlicht: (2024)
Behavior-Consistent Deep Reinforcement Learning
von: Hussing, Marcel, et al.
Veröffentlicht: (2026)
von: Hussing, Marcel, et al.
Veröffentlicht: (2026)
Intelligent Switching for Reset-Free RL
von: Patil, Darshan, et al.
Veröffentlicht: (2024)
von: Patil, Darshan, et al.
Veröffentlicht: (2024)
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
von: Myers, Vivek, et al.
Veröffentlicht: (2024)
Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning
von: Reizinger, Patrik, et al.
Veröffentlicht: (2025)
von: Reizinger, Patrik, et al.
Veröffentlicht: (2025)
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
von: Tang, Hongyao, et al.
Veröffentlicht: (2025)
von: Tang, Hongyao, et al.
Veröffentlicht: (2025)
Diversifying Policy Behaviors with Extrinsic Behavioral Curiosity
von: Wan, Zhenglin, et al.
Veröffentlicht: (2024)
von: Wan, Zhenglin, et al.
Veröffentlicht: (2024)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
von: Jain, Arnav Kumar, et al.
Veröffentlicht: (2024)
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
von: Zheng, Chongyi, et al.
Veröffentlicht: (2024)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2024)
Contrastive Difference Predictive Coding
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
Game-Theoretic Robust Reinforcement Learning Handles Temporally-Coupled Perturbations
von: Liang, Yongyuan, et al.
Veröffentlicht: (2023)
von: Liang, Yongyuan, et al.
Veröffentlicht: (2023)
BuilderBench: The Building Blocks of Intelligent Agents
von: Ghugare, Raj, et al.
Veröffentlicht: (2025)
von: Ghugare, Raj, et al.
Veröffentlicht: (2025)
UpSkill: Mutual Information Skill Learning for Structured Response Diversity in LLMs
von: Shah, Devan, et al.
Veröffentlicht: (2026)
von: Shah, Devan, et al.
Veröffentlicht: (2026)
Bridging State and History Representations: Understanding Self-Predictive RL
von: Ni, Tianwei, et al.
Veröffentlicht: (2024)
von: Ni, Tianwei, et al.
Veröffentlicht: (2024)
A Rate-Distortion View of Uncertainty Quantification
von: Apostolopoulou, Ifigeneia, et al.
Veröffentlicht: (2024)
von: Apostolopoulou, Ifigeneia, et al.
Veröffentlicht: (2024)
OGBench: Benchmarking Offline Goal-Conditioned RL
von: Park, Seohong, et al.
Veröffentlicht: (2024)
von: Park, Seohong, et al.
Veröffentlicht: (2024)
Intention-Conditioned Flow Occupancy Models
von: Zheng, Chongyi, et al.
Veröffentlicht: (2025)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2025)
Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization
von: Modirshanechi, Alireza, et al.
Veröffentlicht: (2026)
von: Modirshanechi, Alireza, et al.
Veröffentlicht: (2026)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
von: Ishihara, Yu, et al.
Veröffentlicht: (2025)
Tiered Reward: Designing Rewards for Specification and Fast Learning of Desired Behavior
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2022)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2022)
Exploration by Random Reward Perturbation
von: Ma, Haozhe, et al.
Veröffentlicht: (2025)
von: Ma, Haozhe, et al.
Veröffentlicht: (2025)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
von: Park, Seohong, et al.
Veröffentlicht: (2023)
von: Park, Seohong, et al.
Veröffentlicht: (2023)
Safe-Support Q-Learning: Learning without Unsafe Exploration
von: Lim, Yeeun, et al.
Veröffentlicht: (2026)
von: Lim, Yeeun, et al.
Veröffentlicht: (2026)
Value Flows
von: Dong, Perry, et al.
Veröffentlicht: (2025)
von: Dong, Perry, et al.
Veröffentlicht: (2025)
1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities
von: Wang, Kevin, et al.
Veröffentlicht: (2025)
von: Wang, Kevin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Is Exploration or Optimization the Problem for Deep Reinforcement Learning?
von: Berseth, Glen
Veröffentlicht: (2025) -
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025) -
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
von: Hugessen, Adriana, et al.
Veröffentlicht: (2024) -
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
von: Liu, Grace, et al.
Veröffentlicht: (2024) -
Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration
von: Nimonkar, Chirayu, et al.
Veröffentlicht: (2025)