Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiao, Yuchen, Chen, Yuxin, Li, Gen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards a unified framework for guided diffusion models
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025)
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025)
Provable Efficiency of Guidance in Diffusion Models for General Data Distribution
von: Li, Gen, et al.
Veröffentlicht: (2025)
von: Li, Gen, et al.
Veröffentlicht: (2025)
Transformers Meet In-Context Learning: A Universal Approximation Theory
von: Li, Gen, et al.
Veröffentlicht: (2025)
von: Li, Gen, et al.
Veröffentlicht: (2025)
Optimal Convergence Analysis of DDPM for General Distributions
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025)
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025)
Are First-Order Diffusion Samplers Really Slower? A Fast Forward-Value Approach
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025)
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025)
Faster Diffusion Models via Higher-Order Approximation
von: Li, Gen, et al.
Veröffentlicht: (2025)
von: Li, Gen, et al.
Veröffentlicht: (2025)
Towards Faster Non-Asymptotic Convergence for Diffusion-Based Generative Models
von: Li, Gen, et al.
Veröffentlicht: (2023)
von: Li, Gen, et al.
Veröffentlicht: (2023)
Towards a mathematical theory for consistency training in diffusion models
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Deflated HeteroPCA: Overcoming the curse of ill-conditioning in heteroskedastic PCA
von: Zhou, Yuchen, et al.
Veröffentlicht: (2023)
von: Zhou, Yuchen, et al.
Veröffentlicht: (2023)
Minimax-Optimal Reward-Agnostic Exploration in Reinforcement Learning
von: Li, Gen, et al.
Veröffentlicht: (2023)
von: Li, Gen, et al.
Veröffentlicht: (2023)
Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
von: Li, Gen, et al.
Veröffentlicht: (2020)
von: Li, Gen, et al.
Veröffentlicht: (2020)
The Curious Price of Distributional Robustness in Reinforcement Learning with a Generative Model
von: Shi, Laixi, et al.
Veröffentlicht: (2023)
von: Shi, Laixi, et al.
Veröffentlicht: (2023)
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
von: Yan, Yuling, et al.
Veröffentlicht: (2022)
von: Yan, Yuling, et al.
Veröffentlicht: (2022)
Is Q-Learning Minimax Optimal? A Tight Sample Complexity Analysis
von: Li, Gen, et al.
Veröffentlicht: (2021)
von: Li, Gen, et al.
Veröffentlicht: (2021)
Settling the Sample Complexity of Model-Based Offline Reinforcement Learning
von: Li, Gen, et al.
Veröffentlicht: (2022)
von: Li, Gen, et al.
Veröffentlicht: (2022)
Tensor Cookbook: Mastering Tensors through Diagrams
von: Rakhshan, Beheshteh T., et al.
Veröffentlicht: (2026)
von: Rakhshan, Beheshteh T., et al.
Veröffentlicht: (2026)
Towards a better list of citation superstars: compiling a multidisciplinary list of highly cited researchers
von: Podlubny, Igor, et al.
Veröffentlicht: (2006)
von: Podlubny, Igor, et al.
Veröffentlicht: (2006)
A Sharp Convergence Theory for The Probability Flow ODEs of Diffusion Models
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Breaking AR's Sampling Bottleneck: Provable Acceleration via Diffusion Language Models
von: Li, Gen, et al.
Veröffentlicht: (2025)
von: Li, Gen, et al.
Veröffentlicht: (2025)
Minimax Optimality of the Probability Flow ODE for Diffusion Models
von: Cai, Changxiao, et al.
Veröffentlicht: (2025)
von: Cai, Changxiao, et al.
Veröffentlicht: (2025)
Optimal structure learning and conditional independence testing
von: Gao, Ming, et al.
Veröffentlicht: (2025)
von: Gao, Ming, et al.
Veröffentlicht: (2025)
Continuous-time reinforcement learning: ellipticity enables model-free value function approximation
von: Mou, Wenlong
Veröffentlicht: (2026)
von: Mou, Wenlong
Veröffentlicht: (2026)
Prior-dependent analysis of posterior sampling reinforcement learning with function approximation
von: Li, Yingru, et al.
Veröffentlicht: (2024)
von: Li, Yingru, et al.
Veröffentlicht: (2024)
Learning single index model with gradient descent: spectral initialization and precise asymptotics
von: Chen, Yuchen, et al.
Veröffentlicht: (2025)
von: Chen, Yuchen, et al.
Veröffentlicht: (2025)
Breaking Training Bottlenecks: Effective and Stable Reinforcement Learning for Coding Models
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
Scaling Data Difficulty: Improving Coding Models via Reinforcement Learning on Fresh and Challenging Problems
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
Denoising diffusion probabilistic models are optimally adaptive to unknown low dimensionality
von: Huang, Zhihan, et al.
Veröffentlicht: (2024)
von: Huang, Zhihan, et al.
Veröffentlicht: (2024)
O(d/T) Convergence Theory for Diffusion Probabilistic Models under Minimal Assumptions
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Adapting to Unknown Low-Dimensional Structures in Score-Based Diffusion Models
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
High-accuracy and dimension-free sampling with diffusions
von: Gatmiry, Khashayar, et al.
Veröffentlicht: (2026)
von: Gatmiry, Khashayar, et al.
Veröffentlicht: (2026)
A non-asymptotic distributional theory of approximate message passing for sparse and robust regression
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Maximum diffusion reinforcement learning
von: Berrueta, Thomas A., et al.
Veröffentlicht: (2023)
von: Berrueta, Thomas A., et al.
Veröffentlicht: (2023)
Adaptive deep learning for nonlinear time series models
von: Kurisu, Daisuke, et al.
Veröffentlicht: (2022)
von: Kurisu, Daisuke, et al.
Veröffentlicht: (2022)
Kaggle Chronicles: 15 Years of Competitions, Community and Data Science Innovation
von: Bönisch, Kevin, et al.
Veröffentlicht: (2025)
von: Bönisch, Kevin, et al.
Veröffentlicht: (2025)
Mesterséges Intelligencia Kutatások Magyarországon
von: Benczúr, András A., et al.
Veröffentlicht: (2025)
von: Benczúr, András A., et al.
Veröffentlicht: (2025)
Spike-timing-dependent Hebbian learning as noisy gradient descent
von: Dexheimer, Niklas, et al.
Veröffentlicht: (2025)
von: Dexheimer, Niklas, et al.
Veröffentlicht: (2025)
Optimal training-conditional regret for online conformal prediction
von: Liang, Jiadong, et al.
Veröffentlicht: (2026)
von: Liang, Jiadong, et al.
Veröffentlicht: (2026)
High-accuracy sampling for diffusion models and log-concave distributions
von: Chen, Fan, et al.
Veröffentlicht: (2026)
von: Chen, Fan, et al.
Veröffentlicht: (2026)
A Score-Based Density Formula, with Applications in Diffusion Generative Models
von: Li, Gen, et al.
Veröffentlicht: (2024)
von: Li, Gen, et al.
Veröffentlicht: (2024)
Dimension-Free Convergence of Diffusion Models for Approximate Gaussian Mixtures
von: Li, Gen, et al.
Veröffentlicht: (2025)
von: Li, Gen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards a unified framework for guided diffusion models
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025) -
Provable Efficiency of Guidance in Diffusion Models for General Data Distribution
von: Li, Gen, et al.
Veröffentlicht: (2025) -
Transformers Meet In-Context Learning: A Universal Approximation Theory
von: Li, Gen, et al.
Veröffentlicht: (2025) -
Optimal Convergence Analysis of DDPM for General Distributions
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025) -
Are First-Order Diffusion Samplers Really Slower? A Fast Forward-Value Approach
von: Jiao, Yuchen, et al.
Veröffentlicht: (2025)