Saved in:
| Main Authors: | Hollenstein, Jakob, Martius, Georg, Piater, Justus |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2312.11091 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unsupervised Learning of Effective Actions in Robotics
by: Zaric, Marko, et al.
Published: (2024)
by: Zaric, Marko, et al.
Published: (2024)
Dynamic Sparsity: Challenging Common Sparsity Assumptions for Learning World Models in Robotic Reinforcement Learning Benchmarks
by: Pandaram, Muthukumar, et al.
Published: (2025)
by: Pandaram, Muthukumar, et al.
Published: (2025)
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
Effect of Optimizer, Initializer, and Architecture of Hypernetworks on Continual Learning from Demonstration
by: Auddy, Sayantan, et al.
Published: (2023)
by: Auddy, Sayantan, et al.
Published: (2023)
Scalable and Efficient Continual Learning from Demonstration via a Hypernetwork-generated Stable Dynamics Model
by: Auddy, Sayantan, et al.
Published: (2023)
by: Auddy, Sayantan, et al.
Published: (2023)
LPGD: A General Framework for Backpropagation through Embedded Optimization Layers
by: Paulus, Anselm, et al.
Published: (2024)
by: Paulus, Anselm, et al.
Published: (2024)
Causal Action Influence Aware Counterfactual Data Augmentation
by: Urpí, Núria Armengol, et al.
Published: (2024)
by: Urpí, Núria Armengol, et al.
Published: (2024)
Dual-Force: Enhanced Offline Diversity Maximization under Imitation Constraints
by: Kolev, Pavel, et al.
Published: (2025)
by: Kolev, Pavel, et al.
Published: (2025)
Sampling Complexity of TD and PPO in RKHS
by: Zou, Lu, et al.
Published: (2025)
by: Zou, Lu, et al.
Published: (2025)
Object-Centric Learning for Real-World Videos by Predicting Temporal Feature Similarities
by: Zadaianchuk, Andrii, et al.
Published: (2023)
by: Zadaianchuk, Andrii, et al.
Published: (2023)
Soft Forward-Backward Representations for Zero-shot Reinforcement Learning with General Utilities
by: Bagatella, Marco, et al.
Published: (2026)
by: Bagatella, Marco, et al.
Published: (2026)
Backpropagation through Combinatorial Algorithms: Identity with Projection Works
by: Sahoo, Subham Sekhar, et al.
Published: (2022)
by: Sahoo, Subham Sekhar, et al.
Published: (2022)
Active Fine-Tuning of Multi-Task Policies
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
Scaling Laws and Tradeoffs in Recurrent Networks of Expressive Neurons
by: Spieler, Aaron, et al.
Published: (2026)
by: Spieler, Aaron, et al.
Published: (2026)
Turn-PPO: Turn-Level Advantage Estimation with PPO for Improved Multi-Turn RL in Agentic LLMs
by: Li, Junbo, et al.
Published: (2025)
by: Li, Junbo, et al.
Published: (2025)
Offline Diversity Maximization Under Imitation Constraints
by: Vlastelica, Marin, et al.
Published: (2023)
by: Vlastelica, Marin, et al.
Published: (2023)
Test-time Offline Reinforcement Learning on Goal-related Experience
by: Bagatella, Marco, et al.
Published: (2025)
by: Bagatella, Marco, et al.
Published: (2025)
Differentiation of Blackbox Combinatorial Solvers
by: Vlastelica, Marin, et al.
Published: (2019)
by: Vlastelica, Marin, et al.
Published: (2019)
Zero-Shot Offline Imitation Learning via Optimal Transport
by: Rupf, Thomas, et al.
Published: (2024)
by: Rupf, Thomas, et al.
Published: (2024)
CombOptNet: Fit the Right NP-Hard Problem by Learning Integer Programming Constraints
by: Paulus, Anselm, et al.
Published: (2021)
by: Paulus, Anselm, et al.
Published: (2021)
Epistemically-guided forward-backward exploration
by: Urpí, Núria Armengol, et al.
Published: (2025)
by: Urpí, Núria Armengol, et al.
Published: (2025)
TCD-Arena: Assessing Robustness of Time Series Causal Discovery Methods Against Assumption Violations
by: Stein, Gideon, et al.
Published: (2026)
by: Stein, Gideon, et al.
Published: (2026)
Learning 3D-Gaussian Simulators from RGB Videos
by: Zhobro, Mikel, et al.
Published: (2025)
by: Zhobro, Mikel, et al.
Published: (2025)
Forecasting in Offline Reinforcement Learning for Non-stationary Environments
by: Ada, Suzan Ece, et al.
Published: (2025)
by: Ada, Suzan Ece, et al.
Published: (2025)
Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments
by: Beukman, Michael, et al.
Published: (2026)
by: Beukman, Michael, et al.
Published: (2026)
Offline vs. Online Learning in Model-based RL: Lessons for Data Collection Strategies
by: Chen, Jiaqi, et al.
Published: (2025)
by: Chen, Jiaqi, et al.
Published: (2025)
Directional-Clamp PPO
by: Karpel, Gilad, et al.
Published: (2025)
by: Karpel, Gilad, et al.
Published: (2025)
GASP: Guided Asymmetric Self-Play For Coding LLMs
by: Jana, Swadesh, et al.
Published: (2026)
by: Jana, Swadesh, et al.
Published: (2026)
Deep Graph Matching via Blackbox Differentiation of Combinatorial Solvers
by: Rolínek, Michal, et al.
Published: (2020)
by: Rolínek, Michal, et al.
Published: (2020)
Stochastic Decision Horizons for Constrained Reinforcement Learning
by: Milosevic, Nikola, et al.
Published: (2026)
by: Milosevic, Nikola, et al.
Published: (2026)
Optimizing Rank-based Metrics with Blackbox Differentiation
by: Rolínek, Michal, et al.
Published: (2019)
by: Rolínek, Michal, et al.
Published: (2019)
Fault Detection in Solar Thermal Systems using Probabilistic Reconstructions
by: Ebmeier, Florian, et al.
Published: (2025)
by: Ebmeier, Florian, et al.
Published: (2025)
Drifting Fields are not Conservative
by: Franz, Leonard T., et al.
Published: (2026)
by: Franz, Leonard T., et al.
Published: (2026)
Identifying Terrain Physical Parameters from Vision -- Towards Physical-Parameter-Aware Locomotion and Navigation
by: Chen, Jiaqi, et al.
Published: (2024)
by: Chen, Jiaqi, et al.
Published: (2024)
SoftJAX & SoftTorch: Empowering Automatic Differentiation Libraries with Informative Gradients
by: Paulus, Anselm, et al.
Published: (2026)
by: Paulus, Anselm, et al.
Published: (2026)
Imbalanced Classification through the Lens of Spurious Correlations
by: Hackstein, Jakob, et al.
Published: (2025)
by: Hackstein, Jakob, et al.
Published: (2025)
Your Attention Matters: to Improve Model Robustness to Noise and Spurious Correlations
by: Tamayo-Rousseau, Camilo, et al.
Published: (2025)
by: Tamayo-Rousseau, Camilo, et al.
Published: (2025)
Unified Task and Motion Planning using Object-centric Abstractions of Motion Constraints
by: Agostini, Alejandro, et al.
Published: (2023)
by: Agostini, Alejandro, et al.
Published: (2023)
The Expressive Leaky Memory Neuron: an Efficient and Expressive Phenomenological Neuron Model Can Solve Long-Horizon Tasks
by: Spieler, Aaron, et al.
Published: (2023)
by: Spieler, Aaron, et al.
Published: (2023)
Temporally Consistent Object-Centric Learning by Contrasting Slots
by: Manasyan, Anna, et al.
Published: (2024)
by: Manasyan, Anna, et al.
Published: (2024)
Similar Items
-
Unsupervised Learning of Effective Actions in Robotics
by: Zaric, Marko, et al.
Published: (2024) -
Dynamic Sparsity: Challenging Common Sparsity Assumptions for Learning World Models in Robotic Reinforcement Learning Benchmarks
by: Pandaram, Muthukumar, et al.
Published: (2025) -
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
by: Bagatella, Marco, et al.
Published: (2024) -
Effect of Optimizer, Initializer, and Architecture of Hypernetworks on Continual Learning from Demonstration
by: Auddy, Sayantan, et al.
Published: (2023) -
Scalable and Efficient Continual Learning from Demonstration via a Hypernetwork-generated Stable Dynamics Model
by: Auddy, Sayantan, et al.
Published: (2023)