Deep Exploration with PAC-Bayes
Fuente:
arXiv
Saved in:
| Main Authors: | Tasdighi, Bahareh, Haussmann, Manuel, Werge, Nicklas, Wu, Yi-Shan, Kandemir, Melih |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Ensemble Aggregation for Actor-Critics
by: Werge, Nicklas, et al.
Published: (2025)
by: Werge, Nicklas, et al.
Published: (2025)
Improving Actor-Critic Training with Steerable Action-Value Approximation Errors
by: Tasdighi, Bahareh, et al.
Published: (2024)
by: Tasdighi, Bahareh, et al.
Published: (2024)
ObjectRL: An Object-Oriented Reinforcement Learning Codebase
by: Baykal, Gulcin, et al.
Published: (2025)
by: Baykal, Gulcin, et al.
Published: (2025)
PAC-Bayesian Soft Actor-Critic Learning
by: Tasdighi, Bahareh, et al.
Published: (2023)
by: Tasdighi, Bahareh, et al.
Published: (2023)
Deep Actor-Critics with Tight Risk Certificates
by: Tasdighi, Bahareh, et al.
Published: (2025)
by: Tasdighi, Bahareh, et al.
Published: (2025)
Weighted Sequential Bayesian Inference for Non-Stationary Linear Contextual Bandits
by: Werge, Nicklas, et al.
Published: (2023)
by: Werge, Nicklas, et al.
Published: (2023)
Deterministic Uncertainty Propagation for Improved Model-Based Offline Reinforcement Learning
by: Akgül, Abdullah, et al.
Published: (2024)
by: Akgül, Abdullah, et al.
Published: (2024)
A Measure-Theoretic Finite-Sample Theory for Adaptive-Data Fitted Q-Iteration
by: Haussmann, Manuel, et al.
Published: (2026)
by: Haussmann, Manuel, et al.
Published: (2026)
Overcoming Non-stationary Dynamics with Evidential Proximal Policy Optimization
by: Akgül, Abdullah, et al.
Published: (2025)
by: Akgül, Abdullah, et al.
Published: (2025)
Distributional Active Inference
by: Akgül, Abdullah, et al.
Published: (2026)
by: Akgül, Abdullah, et al.
Published: (2026)
Continual Learning of Multi-modal Dynamics with External Memory
by: Akgül, Abdullah, et al.
Published: (2022)
by: Akgül, Abdullah, et al.
Published: (2022)
Recursive PAC-Bayes: A Frequentist Approach to Sequential Prior Updates with No Information Loss
by: Wu, Yi-Shan, et al.
Published: (2024)
by: Wu, Yi-Shan, et al.
Published: (2024)
Disentanglement with Factor Quantized Variational Autoencoders
by: Baykal, Gulcin, et al.
Published: (2024)
by: Baykal, Gulcin, et al.
Published: (2024)
EdVAE: Mitigating Codebook Collapse with Evidential Discrete Variational Autoencoders
by: Baykal, Gulcin, et al.
Published: (2023)
by: Baykal, Gulcin, et al.
Published: (2023)
Improved Algorithms for Stochastic Linear Bandits Using Tail Bounds for Martingale Mixtures
by: Flynn, Hamish, et al.
Published: (2023)
by: Flynn, Hamish, et al.
Published: (2023)
Better-than-KL PAC-Bayes Bounds
by: Kuzborskij, Ilja, et al.
Published: (2024)
by: Kuzborskij, Ilja, et al.
Published: (2024)
Hitting Time Isomorphism for Multi-Stage Planning with Foundation Policies
by: Boock, Magnus Victor, et al.
Published: (2026)
by: Boock, Magnus Victor, et al.
Published: (2026)
PAC-Bayes Analysis for Recalibration in Classification
by: Fujisawa, Masahiro, et al.
Published: (2024)
by: Fujisawa, Masahiro, et al.
Published: (2024)
Learning via Surrogate PAC-Bayes
by: Picard-Weibel, Antoine, et al.
Published: (2024)
by: Picard-Weibel, Antoine, et al.
Published: (2024)
Policy-based Tuning of Autoregressive Image Models with Instance- and Distribution-Level Rewards
by: Baran, Orhun Buğra, et al.
Published: (2026)
by: Baran, Orhun Buğra, et al.
Published: (2026)
PAC-Bayes-Chernoff bounds for unbounded losses
by: Casado, Ioar, et al.
Published: (2024)
by: Casado, Ioar, et al.
Published: (2024)
Empirical PAC-Bayes Bounds for Markov Chains
by: Karagulyan, Vahe, et al.
Published: (2025)
by: Karagulyan, Vahe, et al.
Published: (2025)
How good is PAC-Bayes at explaining generalisation?
by: Picard-Weibel, Antoine, et al.
Published: (2025)
by: Picard-Weibel, Antoine, et al.
Published: (2025)
Refined PAC-Bayes Bounds for Offline Bandits
by: Gouverneur, Amaury, et al.
Published: (2025)
by: Gouverneur, Amaury, et al.
Published: (2025)
Latent variable model for high-dimensional point process with structured missingness
by: Sinelnikov, Maksim, et al.
Published: (2024)
by: Sinelnikov, Maksim, et al.
Published: (2024)
Calibrating Bayesian UNet++ for Sub-Seasonal Forecasting
by: Asan, Busra, et al.
Published: (2024)
by: Asan, Busra, et al.
Published: (2024)
User-friendly introduction to PAC-Bayes bounds
by: Alquier, Pierre
Published: (2021)
by: Alquier, Pierre
Published: (2021)
PAC-Bayes Meets Online Contextual Optimization
by: Xie, Zhuojun, et al.
Published: (2025)
by: Xie, Zhuojun, et al.
Published: (2025)
PAC-Bayes Bounds for Multivariate Linear Regression and Linear Autoencoders
by: Guo, Ruixin, et al.
Published: (2025)
by: Guo, Ruixin, et al.
Published: (2025)
Generalisation under gradient descent via deterministic PAC-Bayes
by: Clerico, Eugenio, et al.
Published: (2022)
by: Clerico, Eugenio, et al.
Published: (2022)
The Challenge of Achieving Attributability in Multilingual Table-to-Text Generation with Question-Answer Blueprints
by: Haussmann, Aden
Published: (2025)
by: Haussmann, Aden
Published: (2025)
Bayes meets Bernstein at the Meta Level: an Analysis of Fast Rates in Meta-Learning with PAC-Bayes
by: Riou, Charles, et al.
Published: (2023)
by: Riou, Charles, et al.
Published: (2023)
On Uniform, Bayesian, and PAC-Bayesian Deep Ensembles
by: Hauptvogel, Nick, et al.
Published: (2024)
by: Hauptvogel, Nick, et al.
Published: (2024)
PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory
by: Wang, Chenyang, et al.
Published: (2026)
by: Wang, Chenyang, et al.
Published: (2026)
Randomized Pairwise Learning with Adaptive Sampling: A PAC-Bayes Analysis
by: Zhou, Sijia, et al.
Published: (2025)
by: Zhou, Sijia, et al.
Published: (2025)
Sheaf Graph Neural Networks via PAC-Bayes Spectral Optimization
by: Choi, Yoonhyuk, et al.
Published: (2025)
by: Choi, Yoonhyuk, et al.
Published: (2025)
PAC-Bayes Generalisation Bounds for Dynamical Systems Including Stable RNNs
by: Eringis, Deividas, et al.
Published: (2023)
by: Eringis, Deividas, et al.
Published: (2023)
Emotion Detection in Twitter Messages Using Combination of Long Short-Term Memory and Convolutional Deep Neural Networks
by: Golchin, Bahareh, et al.
Published: (2025)
by: Golchin, Bahareh, et al.
Published: (2025)
Controlling Multiple Errors Simultaneously with a PAC-Bayes Bound
by: Adams, Reuben, et al.
Published: (2022)
by: Adams, Reuben, et al.
Published: (2022)
Leveraging PAC-Bayes Theory and Gibbs Distributions for Generalization Bounds with Complexity Measures
by: Viallard, Paul, et al.
Published: (2024)
by: Viallard, Paul, et al.
Published: (2024)
Similar Items
-
Adaptive Ensemble Aggregation for Actor-Critics
by: Werge, Nicklas, et al.
Published: (2025) -
Improving Actor-Critic Training with Steerable Action-Value Approximation Errors
by: Tasdighi, Bahareh, et al.
Published: (2024) -
ObjectRL: An Object-Oriented Reinforcement Learning Codebase
by: Baykal, Gulcin, et al.
Published: (2025) -
PAC-Bayesian Soft Actor-Critic Learning
by: Tasdighi, Bahareh, et al.
Published: (2023) -
Deep Actor-Critics with Tight Risk Certificates
by: Tasdighi, Bahareh, et al.
Published: (2025)