Contrastive Difference Predictive Coding
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Chongyi, Salakhutdinov, Ruslan, Eysenbach, Benjamin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Stabilizing Contrastive RL: Techniques for Robotic Goal Reaching from Offline Data
di: Zheng, Chongyi, et al.
Pubblicazione: (2023)
di: Zheng, Chongyi, et al.
Pubblicazione: (2023)
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
di: Myers, Vivek, et al.
Pubblicazione: (2024)
di: Myers, Vivek, et al.
Pubblicazione: (2024)
Can We Really Learn One Representation to Optimize All Rewards?
di: Zheng, Chongyi, et al.
Pubblicazione: (2026)
di: Zheng, Chongyi, et al.
Pubblicazione: (2026)
Intention-Conditioned Flow Occupancy Models
di: Zheng, Chongyi, et al.
Pubblicazione: (2025)
di: Zheng, Chongyi, et al.
Pubblicazione: (2025)
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
di: Zheng, Chongyi, et al.
Pubblicazione: (2024)
di: Zheng, Chongyi, et al.
Pubblicazione: (2024)
Value Flows
di: Dong, Perry, et al.
Pubblicazione: (2025)
di: Dong, Perry, et al.
Pubblicazione: (2025)
UpSkill: Mutual Information Skill Learning for Structured Response Diversity in LLMs
di: Shah, Devan, et al.
Pubblicazione: (2026)
di: Shah, Devan, et al.
Pubblicazione: (2026)
Inference via Interpolation: Contrastive Representations Provably Enable Planning and Inference
di: Eysenbach, Benjamin, et al.
Pubblicazione: (2024)
di: Eysenbach, Benjamin, et al.
Pubblicazione: (2024)
Consistent Zero-Shot Imitation with Contrastive Goal Inference
di: Wantlin, Kathryn, et al.
Pubblicazione: (2025)
di: Wantlin, Kathryn, et al.
Pubblicazione: (2025)
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
di: Liu, Grace, et al.
Pubblicazione: (2024)
di: Liu, Grace, et al.
Pubblicazione: (2024)
Contrastive Representations for Temporal Reasoning
di: Ziarko, Alicja, et al.
Pubblicazione: (2025)
di: Ziarko, Alicja, et al.
Pubblicazione: (2025)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
di: Bortkiewicz, Michał, et al.
Pubblicazione: (2025)
di: Bortkiewicz, Michał, et al.
Pubblicazione: (2025)
InSTA: Towards Internet-Scale Training For Agents
di: Trabucco, Brandon, et al.
Pubblicazione: (2025)
di: Trabucco, Brandon, et al.
Pubblicazione: (2025)
Horizon Generalization in Reinforcement Learning
di: Myers, Vivek, et al.
Pubblicazione: (2025)
di: Myers, Vivek, et al.
Pubblicazione: (2025)
Understanding Visual Concepts Across Models
di: Trabucco, Brandon, et al.
Pubblicazione: (2024)
di: Trabucco, Brandon, et al.
Pubblicazione: (2024)
POPE: Learning to Reason on Hard Problems via Privileged On-Policy Exploration
di: Qu, Yuxiao, et al.
Pubblicazione: (2026)
di: Qu, Yuxiao, et al.
Pubblicazione: (2026)
Tree Search for Language Model Agents
di: Koh, Jing Yu, et al.
Pubblicazione: (2024)
di: Koh, Jing Yu, et al.
Pubblicazione: (2024)
Accelerating Diffusion Planners in Offline RL via Reward-Aware Consistency Trajectory Distillation
di: Duan, Xintong, et al.
Pubblicazione: (2025)
di: Duan, Xintong, et al.
Pubblicazione: (2025)
A Rate-Distortion View of Uncertainty Quantification
di: Apostolopoulou, Ifigeneia, et al.
Pubblicazione: (2024)
di: Apostolopoulou, Ifigeneia, et al.
Pubblicazione: (2024)
Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration
di: Nimonkar, Chirayu, et al.
Pubblicazione: (2025)
di: Nimonkar, Chirayu, et al.
Pubblicazione: (2025)
OGBench: Benchmarking Offline Goal-Conditioned RL
di: Park, Seohong, et al.
Pubblicazione: (2024)
di: Park, Seohong, et al.
Pubblicazione: (2024)
Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
di: Mohamed, Faisal, et al.
Pubblicazione: (2026)
di: Mohamed, Faisal, et al.
Pubblicazione: (2026)
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
di: Dalal, Murtaza, et al.
Pubblicazione: (2024)
di: Dalal, Murtaza, et al.
Pubblicazione: (2024)
Bridging State and History Representations: Understanding Self-Predictive RL
di: Ni, Tianwei, et al.
Pubblicazione: (2024)
di: Ni, Tianwei, et al.
Pubblicazione: (2024)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
di: Park, Seohong, et al.
Pubblicazione: (2023)
di: Park, Seohong, et al.
Pubblicazione: (2023)
1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities
di: Wang, Kevin, et al.
Pubblicazione: (2025)
di: Wang, Kevin, et al.
Pubblicazione: (2025)
Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization
di: Modirshanechi, Alireza, et al.
Pubblicazione: (2026)
di: Modirshanechi, Alireza, et al.
Pubblicazione: (2026)
Rethinking Thinking Tokens: LLMs as Improvement Operators
di: Madaan, Lovish, et al.
Pubblicazione: (2025)
di: Madaan, Lovish, et al.
Pubblicazione: (2025)
Horizon Reduction Makes RL Scalable
di: Park, Seohong, et al.
Pubblicazione: (2025)
di: Park, Seohong, et al.
Pubblicazione: (2025)
Behavior-Consistent Deep Reinforcement Learning
di: Hussing, Marcel, et al.
Pubblicazione: (2026)
di: Hussing, Marcel, et al.
Pubblicazione: (2026)
Training LLM Agents to Empower Humans
di: Ellis, Evan, et al.
Pubblicazione: (2025)
di: Ellis, Evan, et al.
Pubblicazione: (2025)
Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning
di: Reizinger, Patrik, et al.
Pubblicazione: (2025)
di: Reizinger, Patrik, et al.
Pubblicazione: (2025)
RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
di: Qu, Yuxiao, et al.
Pubblicazione: (2025)
di: Qu, Yuxiao, et al.
Pubblicazione: (2025)
Game-Theoretic Robust Reinforcement Learning Handles Temporally-Coupled Perturbations
di: Liang, Yongyuan, et al.
Pubblicazione: (2023)
di: Liang, Yongyuan, et al.
Pubblicazione: (2023)
AgentKit: Structured LLM Reasoning with Dynamic Graphs
di: Wu, Yue, et al.
Pubblicazione: (2024)
di: Wu, Yue, et al.
Pubblicazione: (2024)
UrbanAI 2025 Challenge: Linear vs Transformer Models for Long-Horizon Exogenous Temperature Forecasting
di: Gokhman, Ruslan
Pubblicazione: (2025)
di: Gokhman, Ruslan
Pubblicazione: (2025)
Accelerating Goal-Conditioned RL Algorithms and Research
di: Bortkiewicz, Michał, et al.
Pubblicazione: (2024)
di: Bortkiewicz, Michał, et al.
Pubblicazione: (2024)
Training a Generally Curious Agent
di: Tajwar, Fahim, et al.
Pubblicazione: (2025)
di: Tajwar, Fahim, et al.
Pubblicazione: (2025)
Next Embedding Prediction Makes World Models Stronger
di: Bredis, George, et al.
Pubblicazione: (2026)
di: Bredis, George, et al.
Pubblicazione: (2026)
Blind Inverse Problem Solving Made Easy by Text-to-Image Latent Diffusion
di: Dontas, Michail, et al.
Pubblicazione: (2024)
di: Dontas, Michail, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Stabilizing Contrastive RL: Techniques for Robotic Goal Reaching from Offline Data
di: Zheng, Chongyi, et al.
Pubblicazione: (2023) -
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
di: Myers, Vivek, et al.
Pubblicazione: (2024) -
Can We Really Learn One Representation to Optimize All Rewards?
di: Zheng, Chongyi, et al.
Pubblicazione: (2026) -
Intention-Conditioned Flow Occupancy Models
di: Zheng, Chongyi, et al.
Pubblicazione: (2025) -
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
di: Zheng, Chongyi, et al.
Pubblicazione: (2024)