Constrained Exploration via Reflected Replica Exchange Stochastic Gradient Langevin Dynamics
Fuente:
arXiv
Saved in:
| Main Authors: | Zheng, Haoyang, Du, Hengrong, Feng, Qi, Deng, Wei, Lin, Guang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring Non-Convex Discrete Energy Landscapes: An Efficient Langevin-Like Sampler with Replica Exchange
by: Zheng, Haoyang, et al.
Published: (2025)
by: Zheng, Haoyang, et al.
Published: (2025)
Rethinking Langevin Thompson Sampling from A Stochastic Approximation Perspective
by: Wang, Weixin, et al.
Published: (2025)
by: Wang, Weixin, et al.
Published: (2025)
Accelerating Approximate Thompson Sampling with Underdamped Langevin Monte Carlo
by: Zheng, Haoyang, et al.
Published: (2024)
by: Zheng, Haoyang, et al.
Published: (2024)
BLADE: Bayesian Langevin Active Discovery with Replica Exchange for Identification of Complex Systems
by: Kong, Cindy Xiangrui, et al.
Published: (2025)
by: Kong, Cindy Xiangrui, et al.
Published: (2025)
Non-Reversible Langevin Algorithms for Constrained Sampling
by: Du, Hengrong, et al.
Published: (2025)
by: Du, Hengrong, et al.
Published: (2025)
Reflected Schrödinger Bridge for Constrained Generative Modeling
by: Deng, Wei, et al.
Published: (2024)
by: Deng, Wei, et al.
Published: (2024)
Direct Soft-Policy Sampling via Langevin Dynamics
by: Ki, Donghyeon, et al.
Published: (2026)
by: Ki, Donghyeon, et al.
Published: (2026)
Ultra-Fast Language Generation via Discrete Diffusion Divergence Instruct
by: Zheng, Haoyang, et al.
Published: (2025)
by: Zheng, Haoyang, et al.
Published: (2025)
Cascaded two-stage feature clustering and selection via separability and consistency in fuzzy decision systems
by: Chen, Yuepeng, et al.
Published: (2024)
by: Chen, Yuepeng, et al.
Published: (2024)
Stochastic Collapse: How Gradient Noise Attracts SGD Dynamics Towards Simpler Subnetworks
by: Chen, Feng, et al.
Published: (2023)
by: Chen, Feng, et al.
Published: (2023)
Margin-aware Fuzzy Rough Feature Selection: Bridging Uncertainty Characterization and Pattern Classification
by: Xu, Suping, et al.
Published: (2025)
by: Xu, Suping, et al.
Published: (2025)
Optimal Stochastic Trace Estimation in Generative Modeling
by: Liu, Xinyang, et al.
Published: (2025)
by: Liu, Xinyang, et al.
Published: (2025)
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying
by: Nishimori, Soichiro, et al.
Published: (2026)
by: Nishimori, Soichiro, et al.
Published: (2026)
HomPINNs: homotopy physics-informed neural networks for solving the inverse problems of nonlinear differential equations with multiple solutions
by: Zheng, Haoyang, et al.
Published: (2023)
by: Zheng, Haoyang, et al.
Published: (2023)
Ensuring Safety in an Uncertain Environment: Constrained MDPs via Stochastic Thresholds
by: Zuo, Qian, et al.
Published: (2025)
by: Zuo, Qian, et al.
Published: (2025)
Stochastic Re-weighted Gradient Descent via Distributionally Robust Optimization
by: Kumar, Ramnath, et al.
Published: (2023)
by: Kumar, Ramnath, et al.
Published: (2023)
Incentivized Exploration of Non-Stationary Stochastic Bandits
by: Chakraborty, Sourav, et al.
Published: (2024)
by: Chakraborty, Sourav, et al.
Published: (2024)
Almost Bayesian: The Fractal Dynamics of Stochastic Gradient Descent
by: Hennick, Max, et al.
Published: (2025)
by: Hennick, Max, et al.
Published: (2025)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
by: Yue, Bo, et al.
Published: (2024)
by: Yue, Bo, et al.
Published: (2024)
Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning
by: McCarthy, James, et al.
Published: (2025)
by: McCarthy, James, et al.
Published: (2025)
Fast Explanations via Policy Gradient-Optimized Explainer
by: Pan, Deng, et al.
Published: (2024)
by: Pan, Deng, et al.
Published: (2024)
Scaling up Dynamic Edge Partition Models via Stochastic Gradient MCMC
by: Yang, Sikun, et al.
Published: (2024)
by: Yang, Sikun, et al.
Published: (2024)
Unlocking Reasoning Capabilities in LLMs via Reinforcement Learning Exploration
by: Deng, Wenhao, et al.
Published: (2025)
by: Deng, Wenhao, et al.
Published: (2025)
Variational Learning of Gaussian Process Latent Variable Models through Stochastic Gradient Annealed Importance Sampling
by: Xu, Jian, et al.
Published: (2024)
by: Xu, Jian, et al.
Published: (2024)
Adaptive Heavy-Tailed Stochastic Gradient Descent
by: Gong, Bodu, et al.
Published: (2025)
by: Gong, Bodu, et al.
Published: (2025)
Stochastic Gradient Descent with Momentum is Algorithmically Stable
by: Lei, Yunwen, et al.
Published: (2026)
by: Lei, Yunwen, et al.
Published: (2026)
Stochastic Penalty-Barrier Methods for Constrained Machine Learning
by: Bosák, Adam, et al.
Published: (2026)
by: Bosák, Adam, et al.
Published: (2026)
Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL
by: Wu, Junyi, et al.
Published: (2026)
by: Wu, Junyi, et al.
Published: (2026)
ASTRO: Adaptive Stitching via Dynamics-Guided Trajectory Rollouts
by: Yu, Hang, et al.
Published: (2025)
by: Yu, Hang, et al.
Published: (2025)
QF: Quick Feedforward AI Model Training without Gradient Back Propagation
by: Qi, Feng
Published: (2025)
by: Qi, Feng
Published: (2025)
An Effective Dynamic Gradient Calibration Method for Continual Learning
by: Lin, Weichen, et al.
Published: (2024)
by: Lin, Weichen, et al.
Published: (2024)
Sequential Controlled Langevin Diffusions
by: Chen, Junhua, et al.
Published: (2024)
by: Chen, Junhua, et al.
Published: (2024)
Tilted Quantile Gradient Updates for Quantile-Constrained Reinforcement Learning
by: Li, Chenglin, et al.
Published: (2024)
by: Li, Chenglin, et al.
Published: (2024)
Can LLMs predict the convergence of Stochastic Gradient Descent?
by: Zekri, Oussama, et al.
Published: (2024)
by: Zekri, Oussama, et al.
Published: (2024)
Noise Balance and Stationary Distribution of Stochastic Gradient Descent
by: Ziyin, Liu, et al.
Published: (2023)
by: Ziyin, Liu, et al.
Published: (2023)
Beyond Markovian: Reflective Exploration via Bayes-Adaptive RL for LLM Reasoning
by: Zhang, Shenao, et al.
Published: (2025)
by: Zhang, Shenao, et al.
Published: (2025)
Enhancing Stochastic Gradient Descent: A Unified Framework and Novel Acceleration Methods for Faster Convergence
by: Deng, Yichuan, et al.
Published: (2024)
by: Deng, Yichuan, et al.
Published: (2024)
Gradient Flow Drifting: Generative Modeling via Wasserstein Gradient Flows of KDE-Approximated Divergences
by: Cao, Jiarui, et al.
Published: (2026)
by: Cao, Jiarui, et al.
Published: (2026)
Geometric Neural Operators via Lie Group-Constrained Latent Dynamics
by: Zhang, Jiaquan, et al.
Published: (2026)
by: Zhang, Jiaquan, et al.
Published: (2026)
Generalized Discrete Diffusion with Self-Correction
by: Wang, Linxuan, et al.
Published: (2026)
by: Wang, Linxuan, et al.
Published: (2026)
Similar Items
-
Exploring Non-Convex Discrete Energy Landscapes: An Efficient Langevin-Like Sampler with Replica Exchange
by: Zheng, Haoyang, et al.
Published: (2025) -
Rethinking Langevin Thompson Sampling from A Stochastic Approximation Perspective
by: Wang, Weixin, et al.
Published: (2025) -
Accelerating Approximate Thompson Sampling with Underdamped Langevin Monte Carlo
by: Zheng, Haoyang, et al.
Published: (2024) -
BLADE: Bayesian Langevin Active Discovery with Replica Exchange for Identification of Complex Systems
by: Kong, Cindy Xiangrui, et al.
Published: (2025) -
Non-Reversible Langevin Algorithms for Constrained Sampling
by: Du, Hengrong, et al.
Published: (2025)