In-Context Reinforcement Learning through Bayesian Fusion of Context and Value Prior
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Berkes, Anaïs, Taboga, Vincent, Vakalis, Donna, Rolnick, David, Bengio, Yoshua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HVAC-DPT: A Decision Pretrained Transformer for HVAC Control
von: Berkes, Anaïs
Veröffentlicht: (2024)
von: Berkes, Anaïs
Veröffentlicht: (2024)
In-Context Parametric Inference: Point or Distribution Estimators?
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025)
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025)
Assessing SAM for Tree Crown Instance Segmentation from Drone Imagery
von: Teng, Mélisande, et al.
Veröffentlicht: (2025)
von: Teng, Mélisande, et al.
Veröffentlicht: (2025)
Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement Learning
von: Zhao, Mingde, et al.
Veröffentlicht: (2023)
von: Zhao, Mingde, et al.
Veröffentlicht: (2023)
Shaping Inductive Bias in Diffusion Models through Frequency-Based Noise Control
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2025)
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2025)
GFlowNet Foundations
von: Bengio, Yoshua, et al.
Veröffentlicht: (2021)
von: Bengio, Yoshua, et al.
Veröffentlicht: (2021)
Discrete, compositional, and symbolic representations through attractor dynamics
von: Nam, Andrew, et al.
Veröffentlicht: (2023)
von: Nam, Andrew, et al.
Veröffentlicht: (2023)
Monte Carlo Tree Diffusion for System 2 Planning
von: Yoon, Jaesik, et al.
Veröffentlicht: (2025)
von: Yoon, Jaesik, et al.
Veröffentlicht: (2025)
A Complexity-Based Theory of Compositionality
von: Elmoznino, Eric, et al.
Veröffentlicht: (2024)
von: Elmoznino, Eric, et al.
Veröffentlicht: (2024)
Towards Climate Variable Prediction with Conditioned Spatio-Temporal Normalizing Flows
von: Winkler, Christina, et al.
Veröffentlicht: (2023)
von: Winkler, Christina, et al.
Veröffentlicht: (2023)
Diffusion Generative Flow Samplers: Improving learning signals through partial trajectory optimization
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2023)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2023)
Were RNNs All We Needed?
von: Feng, Leo, et al.
Veröffentlicht: (2024)
von: Feng, Leo, et al.
Veröffentlicht: (2024)
FALCON: Few-step Accurate Likelihoods for Continuous Flows
von: Rehman, Danyal, et al.
Veröffentlicht: (2025)
von: Rehman, Danyal, et al.
Veröffentlicht: (2025)
Mixture-of-Experts Meets In-Context Reinforcement Learning
von: Wu, Wenhao, et al.
Veröffentlicht: (2025)
von: Wu, Wenhao, et al.
Veröffentlicht: (2025)
Towards Monotonic Improvement in In-Context Reinforcement Learning
von: Zhang, Wenhao, et al.
Veröffentlicht: (2025)
von: Zhang, Wenhao, et al.
Veröffentlicht: (2025)
In-Context Reinforcement Learning for Variable Action Spaces
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2023)
von: Sinii, Viacheslav, et al.
Veröffentlicht: (2023)
In-Context Compositional Q-Learning for Offline Reinforcement Learning
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
von: Xu, Qiushui, et al.
Veröffentlicht: (2025)
Efficient Causal Graph Discovery Using Large Language Models
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2024)
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2024)
Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity
von: Muni, Aneri, et al.
Veröffentlicht: (2026)
von: Muni, Aneri, et al.
Veröffentlicht: (2026)
Expert-Guided LLM Reasoning for Battery Discovery: From AI-Driven Hypothesis to Synthesis and Characterization
von: Liu, Shengchao, et al.
Veröffentlicht: (2025)
von: Liu, Shengchao, et al.
Veröffentlicht: (2025)
Active Attacks: Red-teaming LLMs via Adaptive Environments
von: Yun, Taeyoung, et al.
Veröffentlicht: (2025)
von: Yun, Taeyoung, et al.
Veröffentlicht: (2025)
Context-Sensitive Abstractions for Reinforcement Learning with Parameterized Actions
von: Nayyar, Rashmeet Kaur, et al.
Veröffentlicht: (2025)
von: Nayyar, Rashmeet Kaur, et al.
Veröffentlicht: (2025)
In-Context Reinforcement Learning via Communicative World Models
von: Martinez-Lopez, Fernando, et al.
Veröffentlicht: (2025)
von: Martinez-Lopez, Fernando, et al.
Veröffentlicht: (2025)
Statistical Context Detection for Deep Lifelong Reinforcement Learning
von: Dick, Jeffery, et al.
Veröffentlicht: (2024)
von: Dick, Jeffery, et al.
Veröffentlicht: (2024)
Heuristic Transformer: Belief Augmented In-Context Reinforcement Learning
von: Dippel, Oliver, et al.
Veröffentlicht: (2025)
von: Dippel, Oliver, et al.
Veröffentlicht: (2025)
Amortized In-Context Bayesian Posterior Estimation
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025)
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025)
Distributional GFlowNets with Quantile Flows
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2023)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2023)
FSP-Laplace: Function-Space Priors for the Laplace Approximation in Bayesian Deep Learning
von: Cinquin, Tristan, et al.
Veröffentlicht: (2024)
von: Cinquin, Tristan, et al.
Veröffentlicht: (2024)
Learning What Matters: Steering Diffusion via Spectrally Anisotropic Forward Noise
von: Scimeca, Luca, et al.
Veröffentlicht: (2025)
von: Scimeca, Luca, et al.
Veröffentlicht: (2025)
Relational In-Context Learning via Synthetic Pre-training with Structural Prior
von: Wang, Yanbo, et al.
Veröffentlicht: (2026)
von: Wang, Yanbo, et al.
Veröffentlicht: (2026)
Structure-Aligned Protein Language Model
von: Chen, Can, et al.
Veröffentlicht: (2025)
von: Chen, Can, et al.
Veröffentlicht: (2025)
Transformers Provably Implement In-Context Reinforcement Learning with Policy Improvement
von: Liang, Haodong, et al.
Veröffentlicht: (2026)
von: Liang, Haodong, et al.
Veröffentlicht: (2026)
Distilling Reinforcement Learning Algorithms for In-Context Model-Based Planning
von: Son, Jaehyeon, et al.
Veröffentlicht: (2025)
von: Son, Jaehyeon, et al.
Veröffentlicht: (2025)
Online Reinforcement Learning in Non-Stationary Context-Driven Environments
von: Hamadanian, Pouya, et al.
Veröffentlicht: (2023)
von: Hamadanian, Pouya, et al.
Veröffentlicht: (2023)
Generative Diffusion Prior Distillation for Long-Context Knowledge Transfer
von: Udayangani, Nilushika, et al.
Veröffentlicht: (2026)
von: Udayangani, Nilushika, et al.
Veröffentlicht: (2026)
Action abstractions for amortized sampling
von: Boussif, Oussama, et al.
Veröffentlicht: (2024)
von: Boussif, Oussama, et al.
Veröffentlicht: (2024)
Can Safety Fine-Tuning Be More Principled? Lessons Learned from Cybersecurity
von: Williams-King, David, et al.
Veröffentlicht: (2025)
von: Williams-King, David, et al.
Veröffentlicht: (2025)
Solving Bayesian inverse problems with diffusion priors and off-policy RL
von: Scimeca, Luca, et al.
Veröffentlicht: (2025)
von: Scimeca, Luca, et al.
Veröffentlicht: (2025)
Mitigating Shortcut Learning with Diffusion Counterfactuals and Diverse Ensembles
von: Scimeca, Luca, et al.
Veröffentlicht: (2023)
von: Scimeca, Luca, et al.
Veröffentlicht: (2023)
Vintix: Action Model via In-Context Reinforcement Learning
von: Polubarov, Andrey, et al.
Veröffentlicht: (2025)
von: Polubarov, Andrey, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
HVAC-DPT: A Decision Pretrained Transformer for HVAC Control
von: Berkes, Anaïs
Veröffentlicht: (2024) -
In-Context Parametric Inference: Point or Distribution Estimators?
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025) -
Assessing SAM for Tree Crown Instance Segmentation from Drone Imagery
von: Teng, Mélisande, et al.
Veröffentlicht: (2025) -
Consciousness-Inspired Spatio-Temporal Abstractions for Better Generalization in Reinforcement Learning
von: Zhao, Mingde, et al.
Veröffentlicht: (2023) -
Shaping Inductive Bias in Diffusion Models through Frequency-Based Noise Control
von: Jiralerspong, Thomas, et al.
Veröffentlicht: (2025)