Optimistic Task Inference for Behavior Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Rupf, Thomas, Bagatella, Marco, Vlastelica, Marin, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Soft Forward-Backward Representations for Zero-shot Reinforcement Learning with General Utilities
by: Bagatella, Marco, et al.
Published: (2026)
by: Bagatella, Marco, et al.
Published: (2026)
Zero-Shot Offline Imitation Learning via Optimal Transport
by: Rupf, Thomas, et al.
Published: (2024)
by: Rupf, Thomas, et al.
Published: (2024)
Causal Action Influence Aware Counterfactual Data Augmentation
by: Urpí, Núria Armengol, et al.
Published: (2024)
by: Urpí, Núria Armengol, et al.
Published: (2024)
Active Fine-Tuning of Multi-Task Policies
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
Provable Maximum Entropy Manifold Exploration via Diffusion Models
by: De Santi, Riccardo, et al.
Published: (2025)
by: De Santi, Riccardo, et al.
Published: (2025)
SOMBRL: Scalable and Optimistic Model-Based RL
by: Sukhija, Bhavya, et al.
Published: (2025)
by: Sukhija, Bhavya, et al.
Published: (2025)
Test-time Offline Reinforcement Learning on Goal-related Experience
by: Bagatella, Marco, et al.
Published: (2025)
by: Bagatella, Marco, et al.
Published: (2025)
Majority Voting for Code Generation
by: Launer, Tim, et al.
Published: (2026)
by: Launer, Tim, et al.
Published: (2026)
Flow Density Control: Generative Optimization Beyond Entropy-Regularized Fine-Tuning
by: De Santi, Riccardo, et al.
Published: (2025)
by: De Santi, Riccardo, et al.
Published: (2025)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
by: Diaz-Bone, Leander, et al.
Published: (2025)
by: Diaz-Bone, Leander, et al.
Published: (2025)
Dual-Force: Enhanced Offline Diversity Maximization under Imitation Constraints
by: Kolev, Pavel, et al.
Published: (2025)
by: Kolev, Pavel, et al.
Published: (2025)
Optimistic Online LQR via Intrinsic Rewards
by: Bartos, Marcell, et al.
Published: (2026)
by: Bartos, Marcell, et al.
Published: (2026)
Offline Diversity Maximization Under Imitation Constraints
by: Vlastelica, Marin, et al.
Published: (2023)
by: Vlastelica, Marin, et al.
Published: (2023)
Differentiation of Blackbox Combinatorial Solvers
by: Vlastelica, Marin, et al.
Published: (2019)
by: Vlastelica, Marin, et al.
Published: (2019)
Epistemically-guided forward-backward exploration
by: Urpí, Núria Armengol, et al.
Published: (2025)
by: Urpí, Núria Armengol, et al.
Published: (2025)
Optimistic Games for Combinatorial Bayesian Optimization with Application to Protein Design
by: Bal, Melis Ilayda, et al.
Published: (2024)
by: Bal, Melis Ilayda, et al.
Published: (2024)
Problem Space Transformations for Out-of-Distribution Generalisation in Behavioural Cloning
by: Doshi, Kiran, et al.
Published: (2024)
by: Doshi, Kiran, et al.
Published: (2024)
Optimizing Rank-based Metrics with Blackbox Differentiation
by: Rolínek, Michal, et al.
Published: (2019)
by: Rolínek, Michal, et al.
Published: (2019)
When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited
by: Agrawal, Rishabh, et al.
Published: (2026)
by: Agrawal, Rishabh, et al.
Published: (2026)
Backpropagation through Combinatorial Algorithms: Identity with Projection Works
by: Sahoo, Subham Sekhar, et al.
Published: (2022)
by: Sahoo, Subham Sekhar, et al.
Published: (2022)
Optimistic Dual Averaging Unifies Modern Optimizers
by: Pethick, Thomas, et al.
Published: (2026)
by: Pethick, Thomas, et al.
Published: (2026)
Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning
by: Bhardwaj, Arjun, et al.
Published: (2023)
by: Bhardwaj, Arjun, et al.
Published: (2023)
Resource-efficient Inference with Foundation Model Programs
by: Nie, Lunyiu, et al.
Published: (2025)
by: Nie, Lunyiu, et al.
Published: (2025)
Optimistically Optimistic Exploration for Provably Efficient Infinite-Horizon Reinforcement and Imitation Learning
by: Moulin, Antoine, et al.
Published: (2025)
by: Moulin, Antoine, et al.
Published: (2025)
Task Tokens: A Flexible Approach to Adapting Behavior Foundation Models
by: Vainshtein, Ron, et al.
Published: (2025)
by: Vainshtein, Ron, et al.
Published: (2025)
Uncertainty-Aware Robotic World Model Makes Offline Model-Based Reinforcement Learning Work on Real Robots
by: Li, Chenhao, et al.
Published: (2025)
by: Li, Chenhao, et al.
Published: (2025)
Optimistic Model Rollouts for Pessimistic Offline Policy Optimization
by: Zhai, Yuanzhao, et al.
Published: (2024)
by: Zhai, Yuanzhao, et al.
Published: (2024)
Optimistic Information Directed Sampling
by: Neu, Gergely, et al.
Published: (2024)
by: Neu, Gergely, et al.
Published: (2024)
Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics
by: Li, Chenhao, et al.
Published: (2025)
by: Li, Chenhao, et al.
Published: (2025)
Optimistic Policy Regularization
by: Pham, Mai, et al.
Published: (2026)
by: Pham, Mai, et al.
Published: (2026)
Reinforcement Learning via Self-Distillation
by: Hübotter, Jonas, et al.
Published: (2026)
by: Hübotter, Jonas, et al.
Published: (2026)
TD-JEPA: Latent-predictive Representations for Zero-Shot Reinforcement Learning
by: Bagatella, Marco, et al.
Published: (2025)
by: Bagatella, Marco, et al.
Published: (2025)
Sparse Optimistic Information Directed Sampling
by: Schwartz, Ludovic, et al.
Published: (2025)
by: Schwartz, Ludovic, et al.
Published: (2025)
SLOPE: Optimistic Potential Landscape Shaping for Model-based Reinforcement Learning
by: Li, Yao-Hui, et al.
Published: (2026)
by: Li, Yao-Hui, et al.
Published: (2026)
Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models
by: Hübotter, Jonas, et al.
Published: (2025)
by: Hübotter, Jonas, et al.
Published: (2025)
Omega: Optimistic EMA Gradients
by: Ramirez, Juan, et al.
Published: (2023)
by: Ramirez, Juan, et al.
Published: (2023)
Optimistic critics can empower small actors
by: Mastikhina, Olya, et al.
Published: (2025)
by: Mastikhina, Olya, et al.
Published: (2025)
Bayesian Optimistic Optimisation with Exponentially Decaying Regret
by: Tran-The, Hung, et al.
Published: (2021)
by: Tran-The, Hung, et al.
Published: (2021)
Causal Modeling with Stationary Diffusions
by: Lorch, Lars, et al.
Published: (2023)
by: Lorch, Lars, et al.
Published: (2023)
Similar Items
-
Soft Forward-Backward Representations for Zero-shot Reinforcement Learning with General Utilities
by: Bagatella, Marco, et al.
Published: (2026) -
Zero-Shot Offline Imitation Learning via Optimal Transport
by: Rupf, Thomas, et al.
Published: (2024) -
Causal Action Influence Aware Counterfactual Data Augmentation
by: Urpí, Núria Armengol, et al.
Published: (2024) -
Active Fine-Tuning of Multi-Task Policies
by: Bagatella, Marco, et al.
Published: (2024) -
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
by: Bagatella, Marco, et al.
Published: (2024)