Temporal Difference Flows
Fuente:
arXiv
Saved in:
| Main Authors: | Farebrother, Jesse, Pirotta, Matteo, Tirinzoni, Andrea, Munos, Rémi, Lazaric, Alessandro, Touati, Ahmed |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Compositional Planning with Jumpy World Models
by: Farebrother, Jesse, et al.
Published: (2026)
by: Farebrother, Jesse, et al.
Published: (2026)
Simple Ingredients for Offline Reinforcement Learning
by: Cetin, Edoardo, et al.
Published: (2024)
by: Cetin, Edoardo, et al.
Published: (2024)
Zero-Shot Whole-Body Humanoid Control via Behavioral Foundation Models
by: Tirinzoni, Andrea, et al.
Published: (2025)
by: Tirinzoni, Andrea, et al.
Published: (2025)
TD-JEPA: Latent-predictive Representations for Zero-Shot Reinforcement Learning
by: Bagatella, Marco, et al.
Published: (2025)
by: Bagatella, Marco, et al.
Published: (2025)
Fast Adaptation with Behavioral Foundation Models
by: Sikchi, Harshit, et al.
Published: (2025)
by: Sikchi, Harshit, et al.
Published: (2025)
CALE: Continuous Arcade Learning Environment
by: Farebrother, Jesse, et al.
Published: (2024)
by: Farebrother, Jesse, et al.
Published: (2024)
Super-Exponential Regret for UCT, AlphaGo and Variants
by: Orseau, Laurent, et al.
Published: (2024)
by: Orseau, Laurent, et al.
Published: (2024)
Sample Complexity Bounds for Stochastic Shortest Path with a Generative Model
by: Tarbouriech, Jean, et al.
Published: (2026)
by: Tarbouriech, Jean, et al.
Published: (2026)
Non-Adversarial Inverse Reinforcement Learning via Successor Feature Matching
by: Jain, Arnav Kumar, et al.
Published: (2024)
by: Jain, Arnav Kumar, et al.
Published: (2024)
Spectral bandits
by: Kocák, Tomáš, et al.
Published: (2026)
by: Kocák, Tomáš, et al.
Published: (2026)
Temporal-Difference Variational Continual Learning
by: Melo, Luckeciano C., et al.
Published: (2024)
by: Melo, Luckeciano C., et al.
Published: (2024)
A Distributional Analogue to the Successor Representation
by: Wiltzer, Harley, et al.
Published: (2024)
by: Wiltzer, Harley, et al.
Published: (2024)
Soft Policy Optimization: Online Off-Policy RL for Sequence Models
by: Cohen, Taco, et al.
Published: (2025)
by: Cohen, Taco, et al.
Published: (2025)
TKAN: Temporal Kolmogorov-Arnold Networks
by: Genet, Remi, et al.
Published: (2024)
by: Genet, Remi, et al.
Published: (2024)
Mixtures of Experts Unlock Parameter Scaling for Deep RL
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
Discerning Temporal Difference Learning
by: Ma, Jianfei
Published: (2023)
by: Ma, Jianfei
Published: (2023)
Backstepping Temporal Difference Learning
by: Lim, Han-Dong, et al.
Published: (2023)
by: Lim, Han-Dong, et al.
Published: (2023)
Gradient Iterated Temporal-Difference Learning
by: Vincent, Théo, et al.
Published: (2026)
by: Vincent, Théo, et al.
Published: (2026)
Stop Regressing: Training Value Functions via Classification for Scalable Deep RL
by: Farebrother, Jesse, et al.
Published: (2024)
by: Farebrother, Jesse, et al.
Published: (2024)
Demystifying the Recency Heuristic in Temporal-Difference Learning
by: Daley, Brett, et al.
Published: (2024)
by: Daley, Brett, et al.
Published: (2024)
Generalized Preference Optimization: A Unified Approach to Offline Alignment
by: Tang, Yunhao, et al.
Published: (2024)
by: Tang, Yunhao, et al.
Published: (2024)
Understanding the performance gap between online and offline alignment algorithms
by: Tang, Yunhao, et al.
Published: (2024)
by: Tang, Yunhao, et al.
Published: (2024)
Is Temporal Difference Learning the Gold Standard for Stitching in RL?
by: Bortkiewicz, Michał, et al.
Published: (2025)
by: Bortkiewicz, Michał, et al.
Published: (2025)
A Variance Minimization Approach to Temporal-Difference Learning
by: Chen, Xingguo, et al.
Published: (2024)
by: Chen, Xingguo, et al.
Published: (2024)
FlowNet: Modeling Dynamic Spatio-Temporal Systems via Flow Propagation
by: Feng, Yutong, et al.
Published: (2025)
by: Feng, Yutong, et al.
Published: (2025)
Intelligent Video Recording Optimization using Activity Detection for Surveillance Systems
by: Elmir, Youssef, et al.
Published: (2024)
by: Elmir, Youssef, et al.
Published: (2024)
Finite-Time Analysis of Temporal Difference Learning with Experience Replay
by: Lim, Han-Dong, et al.
Published: (2023)
by: Lim, Han-Dong, et al.
Published: (2023)
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
by: Rojas, Juan Sebastian, et al.
Published: (2026)
by: Rojas, Juan Sebastian, et al.
Published: (2026)
Optimizing Return Distributions with Distributional Dynamic Programming
by: Pires, Bernardo Ávila, et al.
Published: (2025)
by: Pires, Bernardo Ávila, et al.
Published: (2025)
An Analysis of Action-Value Temporal-Difference Methods That Learn State Values
by: Daley, Brett, et al.
Published: (2025)
by: Daley, Brett, et al.
Published: (2025)
Finite Sample Analysis of Linear Temporal Difference Learning with Arbitrary Features
by: Xie, Zixuan, et al.
Published: (2025)
by: Xie, Zixuan, et al.
Published: (2025)
Almost Sure Convergence of Linear Temporal Difference Learning with Arbitrary Features
by: Wang, Jiuqi, et al.
Published: (2024)
by: Wang, Jiuqi, et al.
Published: (2024)
An MRP Formulation for Supervised Learning: Generalized Temporal Difference Learning Models
by: Pan, Yangchen, et al.
Published: (2024)
by: Pan, Yangchen, et al.
Published: (2024)
Cochain Perspectives on Temporal-Difference Signals for Learning Beyond Markov Dynamics
by: Zhang, Zuyuan, et al.
Published: (2026)
by: Zhang, Zuyuan, et al.
Published: (2026)
Do It for HER: First-Order Temporal Logic Reward Specification in Reinforcement Learning (Extended Version)
by: Olivieri, Pierriccardo, et al.
Published: (2026)
by: Olivieri, Pierriccardo, et al.
Published: (2026)
Spatio-Temporal Self-Supervised Learning for Traffic Flow Prediction
by: Ji, Jiahao, et al.
Published: (2022)
by: Ji, Jiahao, et al.
Published: (2022)
MicroFlow: An Efficient Rust-Based Inference Engine for TinyML
by: Carnelos, Matteo, et al.
Published: (2024)
by: Carnelos, Matteo, et al.
Published: (2024)
LNUCB-TA: Linear-nonlinear Hybrid Bandit Learning with Temporal Attention
by: Khosravi, Hamed, et al.
Published: (2025)
by: Khosravi, Hamed, et al.
Published: (2025)
Towards Climate Variable Prediction with Conditioned Spatio-Temporal Normalizing Flows
by: Winkler, Christina, et al.
Published: (2023)
by: Winkler, Christina, et al.
Published: (2023)
Similar Items
-
Compositional Planning with Jumpy World Models
by: Farebrother, Jesse, et al.
Published: (2026) -
Simple Ingredients for Offline Reinforcement Learning
by: Cetin, Edoardo, et al.
Published: (2024) -
Zero-Shot Whole-Body Humanoid Control via Behavioral Foundation Models
by: Tirinzoni, Andrea, et al.
Published: (2025) -
TD-JEPA: Latent-predictive Representations for Zero-Shot Reinforcement Learning
by: Bagatella, Marco, et al.
Published: (2025) -
Fast Adaptation with Behavioral Foundation Models
by: Sikchi, Harshit, et al.
Published: (2025)