Soft Forward-Backward Representations for Zero-shot Reinforcement Learning with General Utilities
Fuente:
arXiv
Saved in:
| Main Authors: | Bagatella, Marco, Rupf, Thomas, Martius, Georg, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Zero-Shot Offline Imitation Learning via Optimal Transport
by: Rupf, Thomas, et al.
Published: (2024)
by: Rupf, Thomas, et al.
Published: (2024)
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
Optimistic Task Inference for Behavior Foundation Models
by: Rupf, Thomas, et al.
Published: (2025)
by: Rupf, Thomas, et al.
Published: (2025)
Test-time Offline Reinforcement Learning on Goal-related Experience
by: Bagatella, Marco, et al.
Published: (2025)
by: Bagatella, Marco, et al.
Published: (2025)
Active Fine-Tuning of Multi-Task Policies
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
TD-JEPA: Latent-predictive Representations for Zero-Shot Reinforcement Learning
by: Bagatella, Marco, et al.
Published: (2025)
by: Bagatella, Marco, et al.
Published: (2025)
Causal Action Influence Aware Counterfactual Data Augmentation
by: Urpí, Núria Armengol, et al.
Published: (2024)
by: Urpí, Núria Armengol, et al.
Published: (2024)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
by: Diaz-Bone, Leander, et al.
Published: (2025)
by: Diaz-Bone, Leander, et al.
Published: (2025)
Majority Voting for Code Generation
by: Launer, Tim, et al.
Published: (2026)
by: Launer, Tim, et al.
Published: (2026)
Zero-Shot Object-Centric Representation Learning
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
LPGD: A General Framework for Backpropagation through Embedded Optimization Layers
by: Paulus, Anselm, et al.
Published: (2024)
by: Paulus, Anselm, et al.
Published: (2024)
Reinforcement Learning via Self-Distillation
by: Hübotter, Jonas, et al.
Published: (2026)
by: Hübotter, Jonas, et al.
Published: (2026)
SoftJAX & SoftTorch: Empowering Automatic Differentiation Libraries with Informative Gradients
by: Paulus, Anselm, et al.
Published: (2026)
by: Paulus, Anselm, et al.
Published: (2026)
Forecasting in Offline Reinforcement Learning for Non-stationary Environments
by: Ada, Suzan Ece, et al.
Published: (2025)
by: Ada, Suzan Ece, et al.
Published: (2025)
Stochastic Decision Horizons for Constrained Reinforcement Learning
by: Milosevic, Nikola, et al.
Published: (2026)
by: Milosevic, Nikola, et al.
Published: (2026)
Spectral Alignment in Forward-Backward Representations via Temporal Abstraction
by: Azad, Seyed Mahdi B., et al.
Published: (2026)
by: Azad, Seyed Mahdi B., et al.
Published: (2026)
Learning 3D-Gaussian Simulators from RGB Videos
by: Zhobro, Mikel, et al.
Published: (2025)
by: Zhobro, Mikel, et al.
Published: (2025)
Neural Backward Filtering Forward Guiding
by: Yang, Gefan, et al.
Published: (2026)
by: Yang, Gefan, et al.
Published: (2026)
Colored Noise in PPO: Improved Exploration and Performance through Correlated Action Sampling
by: Hollenstein, Jakob, et al.
Published: (2023)
by: Hollenstein, Jakob, et al.
Published: (2023)
Plasma Shape Control via Zero-shot Generative Reinforcement Learning
by: Wu, Niannian, et al.
Published: (2025)
by: Wu, Niannian, et al.
Published: (2025)
Thinking Forward and Backward: Effective Backward Planning with Large Language Models
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
Object-Centric Learning for Real-World Videos by Predicting Temporal Feature Similarities
by: Zadaianchuk, Andrii, et al.
Published: (2023)
by: Zadaianchuk, Andrii, et al.
Published: (2023)
Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning
by: Stojanovic, Stefan, et al.
Published: (2026)
by: Stojanovic, Stefan, et al.
Published: (2026)
Zero-shot Model-based Reinforcement Learning using Large Language Models
by: Benechehab, Abdelhakim, et al.
Published: (2024)
by: Benechehab, Abdelhakim, et al.
Published: (2024)
Multi-Horizon Representations with Hierarchical Forward Models for Reinforcement Learning
by: McInroe, Trevor, et al.
Published: (2022)
by: McInroe, Trevor, et al.
Published: (2022)
Learning First Integrals via Backward-Generated Data and Guided Reinforcement Learning
by: Zhong, Jingfeng, et al.
Published: (2026)
by: Zhong, Jingfeng, et al.
Published: (2026)
Few-shot Multispectral Segmentation with Representations Generated by Reinforcement Learning
by: Jayakody, Dilith, et al.
Published: (2023)
by: Jayakody, Dilith, et al.
Published: (2023)
Uncertainty-Aware Robotic World Model Makes Offline Model-Based Reinforcement Learning Work on Real Robots
by: Li, Chenhao, et al.
Published: (2025)
by: Li, Chenhao, et al.
Published: (2025)
Understanding Transferable Representation Learning and Zero-shot Transfer in CLIP
by: Chen, Zixiang, et al.
Published: (2023)
by: Chen, Zixiang, et al.
Published: (2023)
Understanding Neural Network Binarization with Forward and Backward Proximal Quantizers
by: Lu, Yiwei, et al.
Published: (2024)
by: Lu, Yiwei, et al.
Published: (2024)
Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning
by: Bhardwaj, Arjun, et al.
Published: (2023)
by: Bhardwaj, Arjun, et al.
Published: (2023)
Dual-Force: Enhanced Offline Diversity Maximization under Imitation Constraints
by: Kolev, Pavel, et al.
Published: (2025)
by: Kolev, Pavel, et al.
Published: (2025)
Submodular Reinforcement Learning
by: Prajapat, Manish, et al.
Published: (2023)
by: Prajapat, Manish, et al.
Published: (2023)
CombOptNet: Fit the Right NP-Hard Problem by Learning Integer Programming Constraints
by: Paulus, Anselm, et al.
Published: (2021)
by: Paulus, Anselm, et al.
Published: (2021)
Inductive Biases for Zero-shot Systematic Generalization in Language-informed Reinforcement Learning
by: Dijujin, Negin Hashemi, et al.
Published: (2025)
by: Dijujin, Negin Hashemi, et al.
Published: (2025)
Using Forwards-Backwards Models to Approximate MDP Homomorphisms
by: Mavor-Parker, Augustine N., et al.
Published: (2022)
by: Mavor-Parker, Augustine N., et al.
Published: (2022)
Multi-View Causal Representation Learning with Partial Observability
by: Yao, Dingling, et al.
Published: (2023)
by: Yao, Dingling, et al.
Published: (2023)
Hyperspherical Forward-Forward with Prototypical Representations
by: Sarode, Shalini, et al.
Published: (2026)
by: Sarode, Shalini, et al.
Published: (2026)
Forward-Backward Knowledge Distillation for Continual Clustering
by: Sadeghi, Mohammadreza, et al.
Published: (2024)
by: Sadeghi, Mohammadreza, et al.
Published: (2024)
Conditional Normalizing Flows for Forward and Backward Joint State and Parameter Estimation
by: Lagunowich, Luke S., et al.
Published: (2026)
by: Lagunowich, Luke S., et al.
Published: (2026)
Similar Items
-
Zero-Shot Offline Imitation Learning via Optimal Transport
by: Rupf, Thomas, et al.
Published: (2024) -
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
by: Bagatella, Marco, et al.
Published: (2024) -
Optimistic Task Inference for Behavior Foundation Models
by: Rupf, Thomas, et al.
Published: (2025) -
Test-time Offline Reinforcement Learning on Goal-related Experience
by: Bagatella, Marco, et al.
Published: (2025) -
Active Fine-Tuning of Multi-Task Policies
by: Bagatella, Marco, et al.
Published: (2024)