Bayesian Computation in Deep Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Wenlong, Li, Bolian, Zhang, Ruqi, Li, Yingzhen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Making Reliable and Flexible Decisions in Long-tailed Classification
di: Li, Bolian, et al.
Pubblicazione: (2025)
di: Li, Bolian, et al.
Pubblicazione: (2025)
Entropy-MCMC: Sampling from Flat Basins with Ease
di: Li, Bolian, et al.
Pubblicazione: (2023)
di: Li, Bolian, et al.
Pubblicazione: (2023)
Uniform-Correct Policy Optimization: Breaking RLVR's Indifference to Diversity
di: Lochab, Anamika, et al.
Pubblicazione: (2026)
di: Lochab, Anamika, et al.
Pubblicazione: (2026)
ETA: Evaluating Then Aligning Safety of Vision Language Models at Inference Time
di: Ding, Yi, et al.
Pubblicazione: (2024)
di: Ding, Yi, et al.
Pubblicazione: (2024)
HiBBO: HiPPO-based Space Consistency for High-dimensional Bayesian Optimisation
di: Xuan, Junyu, et al.
Pubblicazione: (2025)
di: Xuan, Junyu, et al.
Pubblicazione: (2025)
Calibrating Transformers via Sparse Gaussian Processes
di: Chen, Wenlong, et al.
Pubblicazione: (2023)
di: Chen, Wenlong, et al.
Pubblicazione: (2023)
Learning Self-Correction in Vision-Language Models via Rollout Augmentation
di: Ding, Yi, et al.
Pubblicazione: (2026)
di: Ding, Yi, et al.
Pubblicazione: (2026)
Cascade Reward Sampling for Efficient Decoding-Time Alignment
di: Li, Bolian, et al.
Pubblicazione: (2024)
di: Li, Bolian, et al.
Pubblicazione: (2024)
Stacey: Promoting Stochastic Steepest Descent via Accelerated $\ell_p$-Smooth Nonconvex Optimization
di: Luo, Xinyu, et al.
Pubblicazione: (2025)
di: Luo, Xinyu, et al.
Pubblicazione: (2025)
Addressing Performance Saturation for LLM RL via Precise Entropy Curve Control
di: Li, Bolian, et al.
Pubblicazione: (2026)
di: Li, Bolian, et al.
Pubblicazione: (2026)
Your Image is Secretly the Last Frame of a Pseudo Video
di: Chen, Wenlong, et al.
Pubblicazione: (2024)
di: Chen, Wenlong, et al.
Pubblicazione: (2024)
Training Bayesian Neural Networks with Sparse Subspace Variational Inference
di: Li, Junbo, et al.
Pubblicazione: (2024)
di: Li, Junbo, et al.
Pubblicazione: (2024)
Variational Uncertainty Decomposition for In-Context Learning
di: Jayasekera, I. Shavindra, et al.
Pubblicazione: (2025)
di: Jayasekera, I. Shavindra, et al.
Pubblicazione: (2025)
Probabilistic Learning and Generation in Deep Sequence Models
di: Chen, Wenlong
Pubblicazione: (2026)
di: Chen, Wenlong
Pubblicazione: (2026)
Position: Bayesian Deep Learning is Needed in the Age of Large-Scale AI
di: Papamarkou, Theodore, et al.
Pubblicazione: (2024)
di: Papamarkou, Theodore, et al.
Pubblicazione: (2024)
Reheated Gradient-based Discrete Sampling for Combinatorial Optimization
di: Li, Muheng, et al.
Pubblicazione: (2025)
di: Li, Muheng, et al.
Pubblicazione: (2025)
Gradient Descent with Projection Finds Over-Parameterized Neural Networks for Learning Low-Degree Polynomials with Nearly Minimax Optimal Rate
di: Yang, Yingzhen, et al.
Pubblicazione: (2026)
di: Yang, Yingzhen, et al.
Pubblicazione: (2026)
Shallow Neural Networks Learn Low-Degree Spherical Polynomials with Feature Learning by Learnable Channel Attention
di: Yang, Yingzhen
Pubblicazione: (2025)
di: Yang, Yingzhen
Pubblicazione: (2025)
Neural Flow Samplers with Shortcut Models
di: Chen, Wuhao, et al.
Pubblicazione: (2025)
di: Chen, Wuhao, et al.
Pubblicazione: (2025)
Discrete Neural Flow Samplers with Locally Equivariant Transformer
di: Ou, Zijing, et al.
Pubblicazione: (2025)
di: Ou, Zijing, et al.
Pubblicazione: (2025)
Improved Generalization Bounds for Transductive Learning by Transductive Local Complexity and Its Applications
di: Yang, Yingzhen
Pubblicazione: (2023)
di: Yang, Yingzhen
Pubblicazione: (2023)
Spectral Transformer Neural Processes
di: Chen, Xianhe, et al.
Pubblicazione: (2026)
di: Chen, Xianhe, et al.
Pubblicazione: (2026)
Compact Memory for Continual Logistic Regression
di: Jung, Yohan, et al.
Pubblicazione: (2025)
di: Jung, Yohan, et al.
Pubblicazione: (2025)
Recurrent Memory for Online Interdomain Gaussian Processes
di: Chen, Wenlong, et al.
Pubblicazione: (2025)
di: Chen, Wenlong, et al.
Pubblicazione: (2025)
Sharp Generalization for Nonparametric Regression in Interpolation Space by Over-Parameterized Neural Networks Trained with Preconditioned Gradient Descent and Early Stopping
di: Yang, Yingzhen, et al.
Pubblicazione: (2024)
di: Yang, Yingzhen, et al.
Pubblicazione: (2024)
Gradient Descent Finds Over-Parameterized Neural Networks with Sharp Generalization for Nonparametric Regression
di: Yang, Yingzhen, et al.
Pubblicazione: (2024)
di: Yang, Yingzhen, et al.
Pubblicazione: (2024)
Projective Proximal Gradient Descent for A Class of Nonconvex Nonsmooth Optimization Problems: Fast Convergence Without Kurdyka-Lojasiewicz (KL) Property
di: Yang, Yingzhen, et al.
Pubblicazione: (2023)
di: Yang, Yingzhen, et al.
Pubblicazione: (2023)
Neural Stochastic Flows: Solver-Free Modelling and Inference for SDE Solutions
di: Kiyohara, Naoki, et al.
Pubblicazione: (2025)
di: Kiyohara, Naoki, et al.
Pubblicazione: (2025)
Inference-Time Scaling of Discrete Diffusion Models via Importance Weighting and Optimal Proposal Design
di: Ou, Zijing, et al.
Pubblicazione: (2025)
di: Ou, Zijing, et al.
Pubblicazione: (2025)
Slithering Through Gaps: Capturing Discrete Isolated Modes via Logistic Bridging
di: Mohanty, Pinaki, et al.
Pubblicazione: (2026)
di: Mohanty, Pinaki, et al.
Pubblicazione: (2026)
CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought
di: Zhang, Boxuan, et al.
Pubblicazione: (2025)
di: Zhang, Boxuan, et al.
Pubblicazione: (2025)
Streamlining Prediction in Bayesian Deep Learning
di: Li, Rui, et al.
Pubblicazione: (2024)
di: Li, Rui, et al.
Pubblicazione: (2024)
Interdomain Attention: Beyond Token-Level Key-Value Memory
di: Kiyohara, Naoki, et al.
Pubblicazione: (2026)
di: Kiyohara, Naoki, et al.
Pubblicazione: (2026)
Controlled LLM Decoding via Discrete Auto-regressive Biasing
di: Pynadath, Patrick, et al.
Pubblicazione: (2025)
di: Pynadath, Patrick, et al.
Pubblicazione: (2025)
Energy-Based Reward Models for Robust Language Model Alignment
di: Lochab, Anamika, et al.
Pubblicazione: (2025)
di: Lochab, Anamika, et al.
Pubblicazione: (2025)
Why Any-Order Autoregressive Models Need Two-Stream Attention: A Structural-Semantic Tradeoff
di: Pynadath, Patrick, et al.
Pubblicazione: (2026)
di: Pynadath, Patrick, et al.
Pubblicazione: (2026)
Structured Coupling for Flow Matching
di: Sumba, Xavier, et al.
Pubblicazione: (2026)
di: Sumba, Xavier, et al.
Pubblicazione: (2026)
On the Identifiability of Switching Dynamical Systems
di: Balsells-Rodas, Carles, et al.
Pubblicazione: (2023)
di: Balsells-Rodas, Carles, et al.
Pubblicazione: (2023)
Addressing the Inconsistency in Bayesian Deep Learning via Generalized Laplace Approximation
di: Chen, Yinsong, et al.
Pubblicazione: (2024)
di: Chen, Yinsong, et al.
Pubblicazione: (2024)
Graph Contrastive Learning with Low-Rank Regularization and Low-Rank Attention for Noisy Node Classification
di: Wang, Yancheng, et al.
Pubblicazione: (2024)
di: Wang, Yancheng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Making Reliable and Flexible Decisions in Long-tailed Classification
di: Li, Bolian, et al.
Pubblicazione: (2025) -
Entropy-MCMC: Sampling from Flat Basins with Ease
di: Li, Bolian, et al.
Pubblicazione: (2023) -
Uniform-Correct Policy Optimization: Breaking RLVR's Indifference to Diversity
di: Lochab, Anamika, et al.
Pubblicazione: (2026) -
ETA: Evaluating Then Aligning Safety of Vision Language Models at Inference Time
di: Ding, Yi, et al.
Pubblicazione: (2024) -
HiBBO: HiPPO-based Space Consistency for High-dimensional Bayesian Optimisation
di: Xuan, Junyu, et al.
Pubblicazione: (2025)