Multivariate Distributional Reinforcement Learning Using Sliced Divergences
Fuente:
arXiv
Saved in:
| Main Authors: | Debes, Baptiste, Tuytelaars, Tinne |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Distributional value gradients for stochastic environments
by: Debes, Baptiste, et al.
Published: (2026)
by: Debes, Baptiste, et al.
Published: (2026)
Predicting the Susceptibility of Examples to Catastrophic Forgetting
by: Hacohen, Guy, et al.
Published: (2024)
by: Hacohen, Guy, et al.
Published: (2024)
Remembering by Reconstructing: Domain Incremental Learning With Test-Time Training on Video Streams
by: Swinnen, Jonathan, et al.
Published: (2026)
by: Swinnen, Jonathan, et al.
Published: (2026)
Collapse-Proof Non-Contrastive Self-Supervised Learning
by: Sansone, Emanuele, et al.
Published: (2024)
by: Sansone, Emanuele, et al.
Published: (2024)
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
by: Araujo, Vladimir, et al.
Published: (2024)
by: Araujo, Vladimir, et al.
Published: (2024)
Same accuracy, twice as fast: continuous training surpasses retraining from scratch
by: Verwimp, Eli, et al.
Published: (2025)
by: Verwimp, Eli, et al.
Published: (2025)
Eff-GRot: Efficient and Generalizable Rotation Estimation with Transformers
by: Mathioulakis, Fanis, et al.
Published: (2025)
by: Mathioulakis, Fanis, et al.
Published: (2025)
Unsupervised Parameter Efficient Source-free Post-pretraining
by: Jha, Abhishek, et al.
Published: (2025)
by: Jha, Abhishek, et al.
Published: (2025)
Putting a Face to Forgetting: Continual Learning meets Mechanistic Interpretability
by: Masip, Sergi, et al.
Published: (2026)
by: Masip, Sergi, et al.
Published: (2026)
Prediction Error-based Classification for Class-Incremental Learning
by: Zając, Michał, et al.
Published: (2023)
by: Zając, Michał, et al.
Published: (2023)
Two Complementary Perspectives to Continual Learning: Ask Not Only What to Optimize, But Also How
by: Hess, Timm, et al.
Published: (2023)
by: Hess, Timm, et al.
Published: (2023)
The Common Stability Mechanism behind most Self-Supervised Learning Approaches
by: Jha, Abhishek, et al.
Published: (2024)
by: Jha, Abhishek, et al.
Published: (2024)
Self-Learning for Personalized Keyword Spotting on Ultra-Low-Power Audio Sensors
by: Rusci, Manuele, et al.
Published: (2024)
by: Rusci, Manuele, et al.
Published: (2024)
Knowledge Accumulation in Continually Learned Representations and the Issue of Feature Forgetting
by: Hess, Timm, et al.
Published: (2023)
by: Hess, Timm, et al.
Published: (2023)
CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping
by: Lebailly, Tim, et al.
Published: (2023)
by: Lebailly, Tim, et al.
Published: (2023)
Continual Learning of Diffusion Models with Generative Distillation
by: Masip, Sergi, et al.
Published: (2023)
by: Masip, Sergi, et al.
Published: (2023)
DAVE: Diagnostic benchmark for Audio Visual Evaluation
by: Radevski, Gorjan, et al.
Published: (2025)
by: Radevski, Gorjan, et al.
Published: (2025)
Foundations of Multivariate Distributional Reinforcement Learning
by: Wiltzer, Harley, et al.
Published: (2024)
by: Wiltzer, Harley, et al.
Published: (2024)
Co-LoRA: Collaborative Model Personalization on Heterogeneous Multi-Modal Clients
by: Seo, Minhyuk, et al.
Published: (2025)
by: Seo, Minhyuk, et al.
Published: (2025)
Adversarial Dependence Minimization
by: De Plaen, Pierre-François, et al.
Published: (2025)
by: De Plaen, Pierre-François, et al.
Published: (2025)
Gaussian-Smoothed Sliced Probability Divergences
by: Alaya, Mokhtar Z., et al.
Published: (2024)
by: Alaya, Mokhtar Z., et al.
Published: (2024)
A Fast, Robust Elliptical Slice Sampling Implementation for Linearly Truncated Multivariate Normal Distributions
by: Wu, Kaiwen, et al.
Published: (2024)
by: Wu, Kaiwen, et al.
Published: (2024)
Infinite dSprites for Disentangled Continual Learning: Separating Memory Edits from Generalization
by: Dziadzio, Sebastian, et al.
Published: (2023)
by: Dziadzio, Sebastian, et al.
Published: (2023)
Relaxed Triangle Inequality for Kullback-Leibler Divergence Between Multivariate Gaussian Distributions
by: Xiao, Shiji, et al.
Published: (2026)
by: Xiao, Shiji, et al.
Published: (2026)
Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning
by: Li, Zhenghao, et al.
Published: (2025)
by: Li, Zhenghao, et al.
Published: (2025)
The Jacobian and Hessian of the Kullback-Leibler Divergence between Multivariate Gaussian Distributions (Technical Report)
by: Maroñas, Juan
Published: (2025)
by: Maroñas, Juan
Published: (2025)
Model-Free Robust $ϕ$-Divergence Reinforcement Learning Using Both Offline and Online Data
by: Panaganti, Kishan, et al.
Published: (2024)
by: Panaganti, Kishan, et al.
Published: (2024)
Lightweight Multi-System Multivariate Interconnection and Divergence Discovery
by: Asres, Mulugeta Weldezgina, et al.
Published: (2024)
by: Asres, Mulugeta Weldezgina, et al.
Published: (2024)
Diverging Flows: Detecting Extrapolations in Conditional Generation
by: Tsakonas, Constantinos, et al.
Published: (2026)
by: Tsakonas, Constantinos, et al.
Published: (2026)
Convergence Rates for Distribution Matching with Sliced Optimal Transport
by: Thurin, Gauthier, et al.
Published: (2026)
by: Thurin, Gauthier, et al.
Published: (2026)
Geodesic Slice Sampler for Multimodal Distributions with Strong Curvature
by: Williams, Bernardo, et al.
Published: (2025)
by: Williams, Bernardo, et al.
Published: (2025)
Implicit Gaussian Splatting with Efficient Multi-Level Tri-Plane Representation
by: Wu, Minye, et al.
Published: (2024)
by: Wu, Minye, et al.
Published: (2024)
Analysis of Spatial augmentation in Self-supervised models in the purview of training and test distributions
by: Jha, Abhishek, et al.
Published: (2024)
by: Jha, Abhishek, et al.
Published: (2024)
Reinforcement Learning Paycheck Optimization for Multivariate Financial Goals
by: Alaluf, Melda, et al.
Published: (2024)
by: Alaluf, Melda, et al.
Published: (2024)
SafeSlice: Enabling SLA-Compliant O-RAN Slicing via Safe Deep Reinforcement Learning
by: Nagib, Ahmad M., et al.
Published: (2025)
by: Nagib, Ahmad M., et al.
Published: (2025)
Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence
by: Zhu, Lingwei, et al.
Published: (2023)
by: Zhu, Lingwei, et al.
Published: (2023)
Combinatorial Multivariant Multi-Armed Bandits with Applications to Episodic Reinforcement Learning and Beyond
by: Liu, Xutong, et al.
Published: (2024)
by: Liu, Xutong, et al.
Published: (2024)
f-GRPO and Beyond: Divergence-Based Reinforcement Learning Algorithms for General LLM Alignment
by: Haldar, Rajdeep, et al.
Published: (2026)
by: Haldar, Rajdeep, et al.
Published: (2026)
FedKL: Tackling Data Heterogeneity in Federated Reinforcement Learning by Penalizing KL Divergence
by: Xie, Zhijie, et al.
Published: (2022)
by: Xie, Zhijie, et al.
Published: (2022)
Percentile-Based Deep Reinforcement Learning and Reward Based Personalization For Delay Aware RAN Slicing in O-RAN
by: Tehrani, Peyman, et al.
Published: (2025)
by: Tehrani, Peyman, et al.
Published: (2025)
Similar Items
-
Distributional value gradients for stochastic environments
by: Debes, Baptiste, et al.
Published: (2026) -
Predicting the Susceptibility of Examples to Catastrophic Forgetting
by: Hacohen, Guy, et al.
Published: (2024) -
Remembering by Reconstructing: Domain Incremental Learning With Test-Time Training on Video Streams
by: Swinnen, Jonathan, et al.
Published: (2026) -
Collapse-Proof Non-Contrastive Self-Supervised Learning
by: Sansone, Emanuele, et al.
Published: (2024) -
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
by: Araujo, Vladimir, et al.
Published: (2024)