Bayesian Exploration Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Fellows, Mattie, Kaplowitz, Brandon, de Witt, Christian Schroeder, Whiteson, Shimon |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Bayesian Solution To The Imitation Gap
by: Vuorio, Risto, et al.
Published: (2024)
by: Vuorio, Risto, et al.
Published: (2024)
Equivariant Networks for Zero-Shot Coordination
by: Muglich, Darius, et al.
Published: (2022)
by: Muglich, Darius, et al.
Published: (2022)
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
by: Ellis, Benjamin, et al.
Published: (2024)
by: Ellis, Benjamin, et al.
Published: (2024)
GoalLadder: Incremental Goal Discovery with Vision-Language Models
by: Zakharov, Alexey, et al.
Published: (2025)
by: Zakharov, Alexey, et al.
Published: (2025)
Reinforcement Learning and Consumption-Savings Behavior
by: Kaplowitz, Brandon
Published: (2025)
by: Kaplowitz, Brandon
Published: (2025)
Detecting Multi-Agent Collusion Through Multi-Agent Interpretability
by: Rose, Aaron, et al.
Published: (2026)
by: Rose, Aaron, et al.
Published: (2026)
Rate-Informed Discovery via Bayesian Adaptive Multifidelity Sampling
by: Sinha, Aman, et al.
Published: (2024)
by: Sinha, Aman, et al.
Published: (2024)
IGDrivSim: A Benchmark for the Imitation Gap in Autonomous Driving
by: Grislain, Clémence, et al.
Published: (2024)
by: Grislain, Clémence, et al.
Published: (2024)
Learning to Reason at the Frontier of Learnability
by: Foster, Thomas, et al.
Published: (2025)
by: Foster, Thomas, et al.
Published: (2025)
SplAgger: Split Aggregation for Meta-Reinforcement Learning
by: Beck, Jacob, et al.
Published: (2024)
by: Beck, Jacob, et al.
Published: (2024)
Simplifying Deep Temporal Difference Learning
by: Gallici, Matteo, et al.
Published: (2024)
by: Gallici, Matteo, et al.
Published: (2024)
Refining Minimax Regret for Unsupervised Environment Design
by: Beukman, Michael, et al.
Published: (2024)
by: Beukman, Michael, et al.
Published: (2024)
SOReL and TOReL: Two Methods for Fully Offline Reinforcement Learning
by: Fellows, Mattie, et al.
Published: (2025)
by: Fellows, Mattie, et al.
Published: (2025)
Distilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology Control
by: Xiong, Zheng, et al.
Published: (2024)
by: Xiong, Zheng, et al.
Published: (2024)
SAGE: Scalable Ground Truth Evaluations for Large Sparse Autoencoders
by: Venhoff, Constantin, et al.
Published: (2024)
by: Venhoff, Constantin, et al.
Published: (2024)
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025)
by: Goldie, Alexander David, et al.
Published: (2025)
Mirror Learning: A Unifying Framework of Policy Optimisation
by: Kuba, Jakub Grudzien, et al.
Published: (2022)
by: Kuba, Jakub Grudzien, et al.
Published: (2022)
A Survey of In-Context Reinforcement Learning
by: Moeini, Amir, et al.
Published: (2025)
by: Moeini, Amir, et al.
Published: (2025)
HyperVLA: Efficient Inference in Vision-Language-Action Models via Hypernetworks
by: Xiong, Zheng, et al.
Published: (2025)
by: Xiong, Zheng, et al.
Published: (2025)
A Clean Slate for Offline Reinforcement Learning
by: Jackson, Matthew Thomas, et al.
Published: (2025)
by: Jackson, Matthew Thomas, et al.
Published: (2025)
Can Learned Optimization Make Reinforcement Learning Less Difficult?
by: Goldie, Alexander David, et al.
Published: (2024)
by: Goldie, Alexander David, et al.
Published: (2024)
UniGen: Unified Modeling of Initial Agent States and Trajectories for Generating Autonomous Driving Scenarios
by: Mahjourian, Reza, et al.
Published: (2024)
by: Mahjourian, Reza, et al.
Published: (2024)
A Tutorial on Meta-Reinforcement Learning
by: Beck, Jacob, et al.
Published: (2023)
by: Beck, Jacob, et al.
Published: (2023)
DEEDEE: Fast and Scalable Out-of-Distribution Dynamics Detection
by: Aljaafari, Tala, et al.
Published: (2025)
by: Aljaafari, Tala, et al.
Published: (2025)
Efficient Dictionary Learning with Switch Sparse Autoencoders
by: Mudide, Anish, et al.
Published: (2024)
by: Mudide, Anish, et al.
Published: (2024)
Policy-Guided Diffusion
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Evolution Strategies at the Hyperscale
by: Sarkar, Bidipta, et al.
Published: (2025)
by: Sarkar, Bidipta, et al.
Published: (2025)
Discovering Temporally-Aware Reinforcement Learning Algorithms
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
Rethinking Out-of-Distribution Detection for Reinforcement Learning: Advancing Methods for Evaluation and Detection
by: Nasvytis, Linas, et al.
Published: (2024)
by: Nasvytis, Linas, et al.
Published: (2024)
Architecture Matters for Multi-Agent Security
by: Hagag, Ben, et al.
Published: (2026)
by: Hagag, Ben, et al.
Published: (2026)
OpenSanctions Pairs: Large-Scale Entity Matching with LLMs
by: Smith, Chandler, et al.
Published: (2026)
by: Smith, Chandler, et al.
Published: (2026)
Learning to Drive in New Cities Without Human Demonstrations
by: Wang, Zilin, et al.
Published: (2026)
by: Wang, Zilin, et al.
Published: (2026)
Toward Robust Real-World Audio Deepfake Detection: Closing the Explainability Gap
by: Channing, Georgia, et al.
Published: (2024)
by: Channing, Georgia, et al.
Published: (2024)
Exploring Exploration in Bayesian Optimization
by: Papenmeier, Leonard, et al.
Published: (2025)
by: Papenmeier, Leonard, et al.
Published: (2025)
Contraction and Hourglass Persistence for Learning on Graphs, Simplices, and Cells
by: Ji, Mattie, et al.
Published: (2026)
by: Ji, Mattie, et al.
Published: (2026)
Mitigating Goal Misgeneralization via Minimax Regret
by: Sadek, Karim Abdel, et al.
Published: (2025)
by: Sadek, Karim Abdel, et al.
Published: (2025)
Unelicitable Backdoors in Language Models via Cryptographic Transformer Circuits
by: Draguns, Andis, et al.
Published: (2024)
by: Draguns, Andis, et al.
Published: (2024)
PSyDUCK: Training-Free Steganography for Latent Diffusion
by: Mahfuz, Aqib, et al.
Published: (2025)
by: Mahfuz, Aqib, et al.
Published: (2025)
Fundamental Limitations in Pointwise Defences of LLM Finetuning APIs
by: Davies, Xander, et al.
Published: (2025)
by: Davies, Xander, et al.
Published: (2025)
On topological descriptors for graph products
by: Ji, Mattie, et al.
Published: (2025)
by: Ji, Mattie, et al.
Published: (2025)
Similar Items
-
A Bayesian Solution To The Imitation Gap
by: Vuorio, Risto, et al.
Published: (2024) -
Equivariant Networks for Zero-Shot Coordination
by: Muglich, Darius, et al.
Published: (2022) -
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
by: Ellis, Benjamin, et al.
Published: (2024) -
GoalLadder: Incremental Goal Discovery with Vision-Language Models
by: Zakharov, Alexey, et al.
Published: (2025) -
Reinforcement Learning and Consumption-Savings Behavior
by: Kaplowitz, Brandon
Published: (2025)