Better-than-KL PAC-Bayes Bounds
Fuente:
arXiv
Saved in:
| Main Authors: | Kuzborskij, Ilja, Jun, Kwang-Sung, Wu, Yulian, Jang, Kyoungseok, Orabona, Francesco |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Offline and Online KL-Regularized RLHF under Differential Privacy
by: Wu, Yulian, et al.
Published: (2025)
by: Wu, Yulian, et al.
Published: (2025)
Efficient Low-Rank Matrix Estimation, Experimental Design, and Arm-Set-Dependent Low-Rank Bandits
by: Jang, Kyoungseok, et al.
Published: (2024)
by: Jang, Kyoungseok, et al.
Published: (2024)
A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO
by: Zhou, Xingyu, et al.
Published: (2025)
by: Zhou, Xingyu, et al.
Published: (2025)
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
by: Kuzborskij, Ilja, et al.
Published: (2025)
by: Kuzborskij, Ilja, et al.
Published: (2025)
Low-rank bias, weight decay, and model merging in neural networks
by: Kuzborskij, Ilja, et al.
Published: (2025)
by: Kuzborskij, Ilja, et al.
Published: (2025)
Square$χ$PO: Differentially Private and Robust $χ^2$-Preference Optimization in Offline Direct Alignment
by: Zhou, Xingyu, et al.
Published: (2025)
by: Zhou, Xingyu, et al.
Published: (2025)
A Note on How to Remove the $\ln\ln T$ Term from the Squint Bound
by: Orabona, Francesco
Published: (2026)
by: Orabona, Francesco
Published: (2026)
GL-LowPopArt: A Nearly Instance-Wise Minimax-Optimal Estimator for Generalized Low-Rank Trace Regression
by: Lee, Junghyun, et al.
Published: (2025)
by: Lee, Junghyun, et al.
Published: (2025)
From Betting to Empirical Bernstein LIL
by: Orabona, Francesco
Published: (2026)
by: Orabona, Francesco
Published: (2026)
A Modern Introduction to Online Learning
by: Orabona, Francesco
Published: (2019)
by: Orabona, Francesco
Published: (2019)
Stability and Generalization for Bellman Residuals
by: Kang, Enoch H., et al.
Published: (2025)
by: Kang, Enoch H., et al.
Published: (2025)
Empirical PAC-Bayes Bounds for Markov Chains
by: Karagulyan, Vahe, et al.
Published: (2025)
by: Karagulyan, Vahe, et al.
Published: (2025)
Refined PAC-Bayes Bounds for Offline Bandits
by: Gouverneur, Amaury, et al.
Published: (2025)
by: Gouverneur, Amaury, et al.
Published: (2025)
Second-Order Bounds for [0,1]-Valued Regression via Betting Loss
by: Li, Yinan, et al.
Published: (2025)
by: Li, Yinan, et al.
Published: (2025)
To Believe or Not to Believe Your LLM
by: Yadkori, Yasin Abbasi, et al.
Published: (2024)
by: Yadkori, Yasin Abbasi, et al.
Published: (2024)
Fixed Confidence Best Arm Identification in the Bayesian Setting
by: Jang, Kyoungseok, et al.
Published: (2024)
by: Jang, Kyoungseok, et al.
Published: (2024)
Rate-optimal Design for Anytime Best Arm Identification
by: Komiyama, Junpei, et al.
Published: (2025)
by: Komiyama, Junpei, et al.
Published: (2025)
STaR-Bets: Sequential Target-Recalculating Bets for Tighter Confidence Intervals
by: Voráček, Václav, et al.
Published: (2025)
by: Voráček, Václav, et al.
Published: (2025)
PAC-Bayes Bounds for Multivariate Linear Regression and Linear Autoencoders
by: Guo, Ruixin, et al.
Published: (2025)
by: Guo, Ruixin, et al.
Published: (2025)
Optimal Regret of Bernoulli Bandits under Global Differential Privacy
by: Azize, Achraf, et al.
Published: (2025)
by: Azize, Achraf, et al.
Published: (2025)
DAG-Math: Graph-of-Thought Guided Mathematical Reasoning in LLMs
by: Zhang, Yuanhe, et al.
Published: (2025)
by: Zhang, Yuanhe, et al.
Published: (2025)
Kullback-Leibler Maillard Sampling for Multi-armed Bandits with Bounded Rewards
by: Qin, Hao, et al.
Published: (2023)
by: Qin, Hao, et al.
Published: (2023)
Sparsity-Agnostic Linear Bandits with Adaptive Adversaries
by: Jin, Tianyuan, et al.
Published: (2024)
by: Jin, Tianyuan, et al.
Published: (2024)
Online Linear Regression with Paid Stochastic Features
by: Merlis, Nadav, et al.
Published: (2025)
by: Merlis, Nadav, et al.
Published: (2025)
Deep Exploration with PAC-Bayes
by: Tasdighi, Bahareh, et al.
Published: (2024)
by: Tasdighi, Bahareh, et al.
Published: (2024)
New Lower Bounds for Stochastic Non-Convex Optimization through Divergence Decomposition
by: Saad, El Mehdi, et al.
Published: (2025)
by: Saad, El Mehdi, et al.
Published: (2025)
An Equivalence Between Static and Dynamic Regret Minimization
by: Jacobsen, Andrew, et al.
Published: (2024)
by: Jacobsen, Andrew, et al.
Published: (2024)
PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory
by: Wang, Chenyang, et al.
Published: (2026)
by: Wang, Chenyang, et al.
Published: (2026)
PAC-Bayes Generalisation Bounds for Dynamical Systems Including Stable RNNs
by: Eringis, Deividas, et al.
Published: (2023)
by: Eringis, Deividas, et al.
Published: (2023)
Controlling Multiple Errors Simultaneously with a PAC-Bayes Bound
by: Adams, Reuben, et al.
Published: (2022)
by: Adams, Reuben, et al.
Published: (2022)
Leveraging PAC-Bayes Theory and Gibbs Distributions for Generalization Bounds with Complexity Measures
by: Viallard, Paul, et al.
Published: (2024)
by: Viallard, Paul, et al.
Published: (2024)
A Framework for Bounding Deterministic Risk with PAC-Bayes: Applications to Majority Votes
by: Leblanc, Benjamin, et al.
Published: (2025)
by: Leblanc, Benjamin, et al.
Published: (2025)
Improved Regret Bounds of (Multinomial) Logistic Bandits via Regret-to-Confidence-Set Conversion
by: Lee, Junghyun, et al.
Published: (2023)
by: Lee, Junghyun, et al.
Published: (2023)
Sufficient Conditions for Stability of Minimum-Norm Interpolating Deep ReLU Networks
by: Harzli, Ouns El, et al.
Published: (2026)
by: Harzli, Ouns El, et al.
Published: (2026)
New Perspectives on the Polyak Stepsize: Surrogate Functions and Negative Results
by: Orabona, Francesco, et al.
Published: (2025)
by: Orabona, Francesco, et al.
Published: (2025)
A Best-of-Both-Worlds Proof for Tsallis-INF without Fenchel Conjugates
by: Lee, Wei-Cheng, et al.
Published: (2025)
by: Lee, Wei-Cheng, et al.
Published: (2025)
A Finite-Time Analysis of TD Learning with Linear Function Approximation without Projections or Strong Convexity
by: Lee, Wei-Cheng, et al.
Published: (2025)
by: Lee, Wei-Cheng, et al.
Published: (2025)
Online Conformal Prediction via Universal Portfolio Algorithms
by: Liu, Tuo, et al.
Published: (2026)
by: Liu, Tuo, et al.
Published: (2026)
HAVER: Instance-Dependent Error Bounds for Maximum Mean Estimation and Applications to Q-Learning and Monte Carlo Tree Search
by: Nguyen, Tuan Ngo, et al.
Published: (2024)
by: Nguyen, Tuan Ngo, et al.
Published: (2024)
On-Average Stability of Multipass Preconditioned SGD and Effective Dimension
by: Vary, Simon, et al.
Published: (2026)
by: Vary, Simon, et al.
Published: (2026)
Similar Items
-
Offline and Online KL-Regularized RLHF under Differential Privacy
by: Wu, Yulian, et al.
Published: (2025) -
Efficient Low-Rank Matrix Estimation, Experimental Design, and Arm-Set-Dependent Low-Rank Bandits
by: Jang, Kyoungseok, et al.
Published: (2024) -
A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO
by: Zhou, Xingyu, et al.
Published: (2025) -
Pointwise confidence estimation in the non-linear $\ell^2$-regularized least squares
by: Kuzborskij, Ilja, et al.
Published: (2025) -
Low-rank bias, weight decay, and model merging in neural networks
by: Kuzborskij, Ilja, et al.
Published: (2025)