Optimistic Task Inference for Behavior Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rupf, Thomas, Bagatella, Marco, Vlastelica, Marin, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Soft Forward-Backward Representations for Zero-shot Reinforcement Learning with General Utilities
von: Bagatella, Marco, et al.
Veröffentlicht: (2026)
von: Bagatella, Marco, et al.
Veröffentlicht: (2026)
Zero-Shot Offline Imitation Learning via Optimal Transport
von: Rupf, Thomas, et al.
Veröffentlicht: (2024)
von: Rupf, Thomas, et al.
Veröffentlicht: (2024)
Causal Action Influence Aware Counterfactual Data Augmentation
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2024)
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2024)
Active Fine-Tuning of Multi-Task Policies
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)
Provable Maximum Entropy Manifold Exploration via Diffusion Models
von: De Santi, Riccardo, et al.
Veröffentlicht: (2025)
von: De Santi, Riccardo, et al.
Veröffentlicht: (2025)
SOMBRL: Scalable and Optimistic Model-Based RL
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2025)
von: Sukhija, Bhavya, et al.
Veröffentlicht: (2025)
Test-time Offline Reinforcement Learning on Goal-related Experience
von: Bagatella, Marco, et al.
Veröffentlicht: (2025)
von: Bagatella, Marco, et al.
Veröffentlicht: (2025)
Majority Voting for Code Generation
von: Launer, Tim, et al.
Veröffentlicht: (2026)
von: Launer, Tim, et al.
Veröffentlicht: (2026)
Flow Density Control: Generative Optimization Beyond Entropy-Regularized Fine-Tuning
von: De Santi, Riccardo, et al.
Veröffentlicht: (2025)
von: De Santi, Riccardo, et al.
Veröffentlicht: (2025)
DISCOVER: Automated Curricula for Sparse-Reward Reinforcement Learning
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
von: Diaz-Bone, Leander, et al.
Veröffentlicht: (2025)
Dual-Force: Enhanced Offline Diversity Maximization under Imitation Constraints
von: Kolev, Pavel, et al.
Veröffentlicht: (2025)
von: Kolev, Pavel, et al.
Veröffentlicht: (2025)
Optimistic Online LQR via Intrinsic Rewards
von: Bartos, Marcell, et al.
Veröffentlicht: (2026)
von: Bartos, Marcell, et al.
Veröffentlicht: (2026)
Offline Diversity Maximization Under Imitation Constraints
von: Vlastelica, Marin, et al.
Veröffentlicht: (2023)
von: Vlastelica, Marin, et al.
Veröffentlicht: (2023)
Differentiation of Blackbox Combinatorial Solvers
von: Vlastelica, Marin, et al.
Veröffentlicht: (2019)
von: Vlastelica, Marin, et al.
Veröffentlicht: (2019)
Epistemically-guided forward-backward exploration
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2025)
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2025)
Optimistic Games for Combinatorial Bayesian Optimization with Application to Protein Design
von: Bal, Melis Ilayda, et al.
Veröffentlicht: (2024)
von: Bal, Melis Ilayda, et al.
Veröffentlicht: (2024)
Problem Space Transformations for Out-of-Distribution Generalisation in Behavioural Cloning
von: Doshi, Kiran, et al.
Veröffentlicht: (2024)
von: Doshi, Kiran, et al.
Veröffentlicht: (2024)
Optimizing Rank-based Metrics with Blackbox Differentiation
von: Rolínek, Michal, et al.
Veröffentlicht: (2019)
von: Rolínek, Michal, et al.
Veröffentlicht: (2019)
When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited
von: Agrawal, Rishabh, et al.
Veröffentlicht: (2026)
von: Agrawal, Rishabh, et al.
Veröffentlicht: (2026)
Backpropagation through Combinatorial Algorithms: Identity with Projection Works
von: Sahoo, Subham Sekhar, et al.
Veröffentlicht: (2022)
von: Sahoo, Subham Sekhar, et al.
Veröffentlicht: (2022)
Optimistic Dual Averaging Unifies Modern Optimizers
von: Pethick, Thomas, et al.
Veröffentlicht: (2026)
von: Pethick, Thomas, et al.
Veröffentlicht: (2026)
Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning
von: Bhardwaj, Arjun, et al.
Veröffentlicht: (2023)
von: Bhardwaj, Arjun, et al.
Veröffentlicht: (2023)
Resource-efficient Inference with Foundation Model Programs
von: Nie, Lunyiu, et al.
Veröffentlicht: (2025)
von: Nie, Lunyiu, et al.
Veröffentlicht: (2025)
Optimistically Optimistic Exploration for Provably Efficient Infinite-Horizon Reinforcement and Imitation Learning
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
von: Moulin, Antoine, et al.
Veröffentlicht: (2025)
Task Tokens: A Flexible Approach to Adapting Behavior Foundation Models
von: Vainshtein, Ron, et al.
Veröffentlicht: (2025)
von: Vainshtein, Ron, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Robotic World Model Makes Offline Model-Based Reinforcement Learning Work on Real Robots
von: Li, Chenhao, et al.
Veröffentlicht: (2025)
von: Li, Chenhao, et al.
Veröffentlicht: (2025)
Optimistic Model Rollouts for Pessimistic Offline Policy Optimization
von: Zhai, Yuanzhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanzhao, et al.
Veröffentlicht: (2024)
Optimistic Information Directed Sampling
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics
von: Li, Chenhao, et al.
Veröffentlicht: (2025)
von: Li, Chenhao, et al.
Veröffentlicht: (2025)
Optimistic Policy Regularization
von: Pham, Mai, et al.
Veröffentlicht: (2026)
von: Pham, Mai, et al.
Veröffentlicht: (2026)
Reinforcement Learning via Self-Distillation
von: Hübotter, Jonas, et al.
Veröffentlicht: (2026)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2026)
TD-JEPA: Latent-predictive Representations for Zero-Shot Reinforcement Learning
von: Bagatella, Marco, et al.
Veröffentlicht: (2025)
von: Bagatella, Marco, et al.
Veröffentlicht: (2025)
Sparse Optimistic Information Directed Sampling
von: Schwartz, Ludovic, et al.
Veröffentlicht: (2025)
von: Schwartz, Ludovic, et al.
Veröffentlicht: (2025)
SLOPE: Optimistic Potential Landscape Shaping for Model-based Reinforcement Learning
von: Li, Yao-Hui, et al.
Veröffentlicht: (2026)
von: Li, Yao-Hui, et al.
Veröffentlicht: (2026)
Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models
von: Hübotter, Jonas, et al.
Veröffentlicht: (2025)
von: Hübotter, Jonas, et al.
Veröffentlicht: (2025)
Omega: Optimistic EMA Gradients
von: Ramirez, Juan, et al.
Veröffentlicht: (2023)
von: Ramirez, Juan, et al.
Veröffentlicht: (2023)
Optimistic critics can empower small actors
von: Mastikhina, Olya, et al.
Veröffentlicht: (2025)
von: Mastikhina, Olya, et al.
Veröffentlicht: (2025)
Bayesian Optimistic Optimisation with Exponentially Decaying Regret
von: Tran-The, Hung, et al.
Veröffentlicht: (2021)
von: Tran-The, Hung, et al.
Veröffentlicht: (2021)
Causal Modeling with Stationary Diffusions
von: Lorch, Lars, et al.
Veröffentlicht: (2023)
von: Lorch, Lars, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Soft Forward-Backward Representations for Zero-shot Reinforcement Learning with General Utilities
von: Bagatella, Marco, et al.
Veröffentlicht: (2026) -
Zero-Shot Offline Imitation Learning via Optimal Transport
von: Rupf, Thomas, et al.
Veröffentlicht: (2024) -
Causal Action Influence Aware Counterfactual Data Augmentation
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2024) -
Active Fine-Tuning of Multi-Task Policies
von: Bagatella, Marco, et al.
Veröffentlicht: (2024) -
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
von: Bagatella, Marco, et al.
Veröffentlicht: (2024)