Normalizing Flows are Capable Models for RL
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ghugare, Raj, Eysenbach, Benjamin |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
On the Role of Iterative Computation in Reinforcement Learning
par: Ghugare, Raj, et autres
Publié: (2026)
par: Ghugare, Raj, et autres
Publié: (2026)
Closing the Gap between TD Learning and Supervised Learning -- A Generalisation Point of View
par: Ghugare, Raj, et autres
Publié: (2024)
par: Ghugare, Raj, et autres
Publié: (2024)
BuilderBench: The Building Blocks of Intelligent Agents
par: Ghugare, Raj, et autres
Publié: (2025)
par: Ghugare, Raj, et autres
Publié: (2025)
1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities
par: Wang, Kevin, et autres
Publié: (2025)
par: Wang, Kevin, et autres
Publié: (2025)
OGBench: Benchmarking Offline Goal-Conditioned RL
par: Park, Seohong, et autres
Publié: (2024)
par: Park, Seohong, et autres
Publié: (2024)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
par: Bortkiewicz, Michał, et autres
Publié: (2025)
par: Bortkiewicz, Michał, et autres
Publié: (2025)
A Single Goal is All You Need: Skills and Exploration Emerge from Contrastive RL without Rewards, Demonstrations, or Subgoals
par: Liu, Grace, et autres
Publié: (2024)
par: Liu, Grace, et autres
Publié: (2024)
Intention-Conditioned Flow Occupancy Models
par: Zheng, Chongyi, et autres
Publié: (2025)
par: Zheng, Chongyi, et autres
Publié: (2025)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
par: Park, Seohong, et autres
Publié: (2023)
par: Park, Seohong, et autres
Publié: (2023)
Demystifying the Mechanisms Behind Emergent Exploration in Goal-conditioned RL
par: Bastankhah, Mahsa, et autres
Publié: (2025)
par: Bastankhah, Mahsa, et autres
Publié: (2025)
The "Law" of the Unconscious Contrastive Learner: Probabilistic Alignment of Unpaired Modalities
par: Che, Yongwei, et autres
Publié: (2025)
par: Che, Yongwei, et autres
Publié: (2025)
Unifying Goal-Conditioned RL and Unsupervised Skill Learning via Control-Maximization
par: Modirshanechi, Alireza, et autres
Publié: (2026)
par: Modirshanechi, Alireza, et autres
Publié: (2026)
Horizon Reduction Makes RL Scalable
par: Park, Seohong, et autres
Publié: (2025)
par: Park, Seohong, et autres
Publié: (2025)
Normalizing Flows are Capable Generative Models
par: Zhai, Shuangfei, et autres
Publié: (2024)
par: Zhai, Shuangfei, et autres
Publié: (2024)
Value Flows
par: Dong, Perry, et autres
Publié: (2025)
par: Dong, Perry, et autres
Publié: (2025)
Accelerating Goal-Conditioned RL Algorithms and Research
par: Bortkiewicz, Michał, et autres
Publié: (2024)
par: Bortkiewicz, Michał, et autres
Publié: (2024)
Consistent Zero-Shot Imitation with Contrastive Goal Inference
par: Wantlin, Kathryn, et autres
Publié: (2025)
par: Wantlin, Kathryn, et autres
Publié: (2025)
Learning to Perceive the World Through Control: Empowerment-Based Representation Learning
par: Bastankhah, Mahsa, et autres
Publié: (2026)
par: Bastankhah, Mahsa, et autres
Publié: (2026)
Bridging State and History Representations: Understanding Self-Predictive RL
par: Ni, Tianwei, et autres
Publié: (2024)
par: Ni, Tianwei, et autres
Publié: (2024)
Stabilizing Contrastive RL: Techniques for Robotic Goal Reaching from Offline Data
par: Zheng, Chongyi, et autres
Publié: (2023)
par: Zheng, Chongyi, et autres
Publié: (2023)
Horizon Generalization in Reinforcement Learning
par: Myers, Vivek, et autres
Publié: (2025)
par: Myers, Vivek, et autres
Publié: (2025)
Contrastive Difference Predictive Coding
par: Zheng, Chongyi, et autres
Publié: (2023)
par: Zheng, Chongyi, et autres
Publié: (2023)
Inference via Interpolation: Contrastive Representations Provably Enable Planning and Inference
par: Eysenbach, Benjamin, et autres
Publié: (2024)
par: Eysenbach, Benjamin, et autres
Publié: (2024)
Offline Goal-conditioned Reinforcement Learning with Quasimetric Representations
par: Myers, Vivek, et autres
Publié: (2025)
par: Myers, Vivek, et autres
Publié: (2025)
Can We Really Learn One Representation to Optimize All Rewards?
par: Zheng, Chongyi, et autres
Publié: (2026)
par: Zheng, Chongyi, et autres
Publié: (2026)
Out-of-Distribution Adaptation in Offline RL: Counterfactual Reasoning via Causal Normalizing Flows
par: Cho, Minjae, et autres
Publié: (2024)
par: Cho, Minjae, et autres
Publié: (2024)
Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration
par: Nimonkar, Chirayu, et autres
Publié: (2025)
par: Nimonkar, Chirayu, et autres
Publié: (2025)
A Rate-Distortion View of Uncertainty Quantification
par: Apostolopoulou, Ifigeneia, et autres
Publié: (2024)
par: Apostolopoulou, Ifigeneia, et autres
Publié: (2024)
Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
par: Mohamed, Faisal, et autres
Publié: (2026)
par: Mohamed, Faisal, et autres
Publié: (2026)
Can a MISL Fly? Analysis and Ingredients for Mutual Information Skill Learning
par: Zheng, Chongyi, et autres
Publié: (2024)
par: Zheng, Chongyi, et autres
Publié: (2024)
Normalizing Flows on Quotient Manifolds via Boundary Quotients
par: Ghanem, William, et autres
Publié: (2025)
par: Ghanem, William, et autres
Publié: (2025)
Multistep Quasimetric Learning for Scalable Goal-conditioned Reinforcement Learning
par: Zheng, Bill Chunyuan, et autres
Publié: (2025)
par: Zheng, Bill Chunyuan, et autres
Publié: (2025)
SLAP: Shortcut Learning for Abstract Planning
par: Liu, Y. Isabel, et autres
Publié: (2025)
par: Liu, Y. Isabel, et autres
Publié: (2025)
$π_\texttt{RL}$: Online RL Fine-tuning for Flow-based Vision-Language-Action Models
par: Chen, Kang, et autres
Publié: (2025)
par: Chen, Kang, et autres
Publié: (2025)
Distilling Normalizing Flows
par: Walton, Steven, et autres
Publié: (2025)
par: Walton, Steven, et autres
Publié: (2025)
Preferential Normalizing Flows
par: Mikkola, Petrus, et autres
Publié: (2024)
par: Mikkola, Petrus, et autres
Publié: (2024)
Piecewise Normalizing Flows
par: Bevins, Harry, et autres
Publié: (2023)
par: Bevins, Harry, et autres
Publié: (2023)
Contrastive Representations for Temporal Reasoning
par: Ziarko, Alicja, et autres
Publié: (2025)
par: Ziarko, Alicja, et autres
Publié: (2025)
Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making
par: Myers, Vivek, et autres
Publié: (2024)
par: Myers, Vivek, et autres
Publié: (2024)
UpSkill: Mutual Information Skill Learning for Structured Response Diversity in LLMs
par: Shah, Devan, et autres
Publié: (2026)
par: Shah, Devan, et autres
Publié: (2026)
Documents similaires
-
On the Role of Iterative Computation in Reinforcement Learning
par: Ghugare, Raj, et autres
Publié: (2026) -
Closing the Gap between TD Learning and Supervised Learning -- A Generalisation Point of View
par: Ghugare, Raj, et autres
Publié: (2024) -
BuilderBench: The Building Blocks of Intelligent Agents
par: Ghugare, Raj, et autres
Publié: (2025) -
1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities
par: Wang, Kevin, et autres
Publié: (2025) -
OGBench: Benchmarking Offline Goal-Conditioned RL
par: Park, Seohong, et autres
Publié: (2024)