On the Equivalence of Random Network Distillation, Deep Ensembles, and Bayesian Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Zanger, Moritz A., Wu, Yijun, Van der Vaart, Pascal R., Böhmer, Wendelin, Spaan, Matthijs T. J. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Contextual Similarity Distillation: Ensemble Uncertainties with a Single Model
by: Zanger, Moritz A., et al.
Published: (2025)
by: Zanger, Moritz A., et al.
Published: (2025)
Diverse Projection Ensembles for Distributional Reinforcement Learning
by: Zanger, Moritz A., et al.
Published: (2023)
by: Zanger, Moritz A., et al.
Published: (2023)
How Ensembles of Distilled Policies Improve Generalisation in Reinforcement Learning
by: Weltevrede, Max, et al.
Published: (2025)
by: Weltevrede, Max, et al.
Published: (2025)
Universal Value-Function Uncertainties
by: Zanger, Moritz A., et al.
Published: (2025)
by: Zanger, Moritz A., et al.
Published: (2025)
Value Improved Actor Critic Algorithms
by: Oren, Yaniv, et al.
Published: (2024)
by: Oren, Yaniv, et al.
Published: (2024)
Priors Matter: Addressing Misspecification in Bayesian Deep Q-Learning
by: van der Vaart, Pascal R., et al.
Published: (2025)
by: van der Vaart, Pascal R., et al.
Published: (2025)
Twice Sequential Monte Carlo for Tree Search
by: Oren, Yaniv, et al.
Published: (2025)
by: Oren, Yaniv, et al.
Published: (2025)
Explore-Go: Leveraging Exploration for Generalisation in Deep Reinforcement Learning
by: Weltevrede, Max, et al.
Published: (2024)
by: Weltevrede, Max, et al.
Published: (2024)
Epistemic Monte Carlo Tree Search
by: Oren, Yaniv, et al.
Published: (2022)
by: Oren, Yaniv, et al.
Published: (2022)
Exploration Implies Data Augmentation: Reachability and Generalisation in Contextual MDPs
by: Weltevrede, Max, et al.
Published: (2024)
by: Weltevrede, Max, et al.
Published: (2024)
Bayesian Inference with Deep Weakly Nonlinear Networks
by: Hanin, Boris, et al.
Published: (2024)
by: Hanin, Boris, et al.
Published: (2024)
TransZero: Parallel Tree Expansion in MuZero using Transformer Networks
by: Malmsten, Emil, et al.
Published: (2025)
by: Malmsten, Emil, et al.
Published: (2025)
Improving Robustness of AlphaZero Algorithms to Test-Time Environment Changes
by: Tamassia, Isidoro, et al.
Published: (2025)
by: Tamassia, Isidoro, et al.
Published: (2025)
Quantitative CLTs in Deep Neural Networks
by: Favaro, Stefano, et al.
Published: (2023)
by: Favaro, Stefano, et al.
Published: (2023)
Sparse Masked Attention Policies for Reliable Generalization
by: Horsch, Caroline, et al.
Published: (2026)
by: Horsch, Caroline, et al.
Published: (2026)
Generalisation to unseen topologies: Towards control of biological neural network activity
by: Engwegen, Laurens, et al.
Published: (2024)
by: Engwegen, Laurens, et al.
Published: (2024)
Feature Learning Dynamics in Infinite-Depth Neural Networks
by: Yao, Zihan, et al.
Published: (2025)
by: Yao, Zihan, et al.
Published: (2025)
RecBayes: Recurrent Bayesian Ad Hoc Teamwork in Large Partially Observable Domains
by: Ribeiro, João G., et al.
Published: (2025)
by: Ribeiro, João G., et al.
Published: (2025)
Deep Conditional Measure Quantization
by: Turinici, Gabriel
Published: (2023)
by: Turinici, Gabriel
Published: (2023)
Positive Experience Reflection for Agents in Interactive Text Environments
by: Lippmann, Philip, et al.
Published: (2024)
by: Lippmann, Philip, et al.
Published: (2024)
PMCTS: Particle Monte Carlo Tree Search for Principled Parallelized Inference Time Scaling
by: Oren, Yaniv, et al.
Published: (2026)
by: Oren, Yaniv, et al.
Published: (2026)
Advancing Deep Learning through Probability Engineering: A Pragmatic Paradigm for Modern AI
by: Zhang, Jianyi
Published: (2025)
by: Zhang, Jianyi
Published: (2025)
EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control
by: Evers, Thomas, et al.
Published: (2026)
by: Evers, Thomas, et al.
Published: (2026)
Collapsed Inference for Bayesian Deep Learning
by: Zeng, Zhe, et al.
Published: (2023)
by: Zeng, Zhe, et al.
Published: (2023)
Neural Network Parameter-optimization of Gaussian pmDAGs
by: Saremi, Mehrzad
Published: (2023)
by: Saremi, Mehrzad
Published: (2023)
A Neural Scaling Law from Lottery Ticket Ensembling
by: Liu, Ziming, et al.
Published: (2023)
by: Liu, Ziming, et al.
Published: (2023)
Bayesian Networks for Causal Analysis in Socioecological Systems
by: Cabañas, Rafael, et al.
Published: (2024)
by: Cabañas, Rafael, et al.
Published: (2024)
Modular Recurrence in Contextual MDPs for Universal Morphology Control
by: Engwegen, Laurens, et al.
Published: (2025)
by: Engwegen, Laurens, et al.
Published: (2025)
Probabilistic Geometric Alignment via Bayesian Latent Transport for Domain-Adaptive Foundation Models
by: Aueawatthanaphisut, Aueaphum, et al.
Published: (2026)
by: Aueawatthanaphisut, Aueaphum, et al.
Published: (2026)
Virtual Parameter Sharpening: Dynamic Low-Rank Perturbations for Inference-Time Reasoning Enhancement
by: Kublashvili, Saba
Published: (2025)
by: Kublashvili, Saba
Published: (2025)
Physics-Informed Inference Time Scaling for Solving High-Dimensional PDE via Defect Correction
by: Fan, Zexi, et al.
Published: (2025)
by: Fan, Zexi, et al.
Published: (2025)
Uncertainty Quantification with Bayesian Higher Order ReLU KANs
by: Giroux, James, et al.
Published: (2024)
by: Giroux, James, et al.
Published: (2024)
When Attention Collapses: Residual Evidence Modeling for Compositional Inference
by: Houba, Niklas
Published: (2026)
by: Houba, Niklas
Published: (2026)
Multi-hop Upstream Anticipatory Traffic Signal Control with Deep Reinforcement Learning
by: Li, Xiaocan, et al.
Published: (2024)
by: Li, Xiaocan, et al.
Published: (2024)
On the consistent reasoning paradox of intelligence and optimal trust in AI: The power of 'I don't know'
by: Bastounis, Alexander, et al.
Published: (2024)
by: Bastounis, Alexander, et al.
Published: (2024)
DistillKac: Few-Step Image Generation via Damped Wave Equations
by: Han, Weiqiao, et al.
Published: (2025)
by: Han, Weiqiao, et al.
Published: (2025)
Reinforcement Learning by Guided Safe Exploration
by: Yang, Qisong, et al.
Published: (2023)
by: Yang, Qisong, et al.
Published: (2023)
Smart Ensemble Learning Framework for Predicting Groundwater Heavy Metal Pollution
by: Ansah-Narh, T., et al.
Published: (2026)
by: Ansah-Narh, T., et al.
Published: (2026)
$α$-TCAV: A Unified Framework for Testing with Concept Activation Vectors
by: Schnoor, Ekkehard, et al.
Published: (2026)
by: Schnoor, Ekkehard, et al.
Published: (2026)
Causal Effect Identification in Heterogeneous Environments from Higher-Order Moments
by: Kivva, Yaroslav, et al.
Published: (2025)
by: Kivva, Yaroslav, et al.
Published: (2025)
Similar Items
-
Contextual Similarity Distillation: Ensemble Uncertainties with a Single Model
by: Zanger, Moritz A., et al.
Published: (2025) -
Diverse Projection Ensembles for Distributional Reinforcement Learning
by: Zanger, Moritz A., et al.
Published: (2023) -
How Ensembles of Distilled Policies Improve Generalisation in Reinforcement Learning
by: Weltevrede, Max, et al.
Published: (2025) -
Universal Value-Function Uncertainties
by: Zanger, Moritz A., et al.
Published: (2025) -
Value Improved Actor Critic Algorithms
by: Oren, Yaniv, et al.
Published: (2024)