Intrinsic Wasserstein Rates for Score-Based Generative Models on Smooth Manifolds
Fuente:
arXiv
Guardado en:
| Autores principales: | Fu, Guoji, Suzuki, Taiji, Lee, Wee Sun, Nitanda, Atsushi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Approximation and Generalization Abilities of Score-based Neural Network Generative Models for Sub-Gaussian Distributions
por: Fu, Guoji, et al.
Publicado: (2025)
por: Fu, Guoji, et al.
Publicado: (2025)
Towards a Unified Analysis of Neural Networks in Nonparametric Instrumental Variable Regression: Optimization and Generalization
por: Chen, Zonghao, et al.
Publicado: (2025)
por: Chen, Zonghao, et al.
Publicado: (2025)
Direct Distributional Optimization for Provable Alignment of Diffusion Models
por: Kawata, Ryotaro, et al.
Publicado: (2025)
por: Kawata, Ryotaro, et al.
Publicado: (2025)
Implicit Graph Neural Diffusion Networks: Convergence, Generalization, and Over-Smoothing
por: Fu, Guoji, et al.
Publicado: (2023)
por: Fu, Guoji, et al.
Publicado: (2023)
Propagation of Chaos for Mean-Field Langevin Dynamics and its Application to Model Ensemble
por: Nitanda, Atsushi, et al.
Publicado: (2025)
por: Nitanda, Atsushi, et al.
Publicado: (2025)
Koopman-based generalization bound: New aspect for full-rank weights
por: Hashimoto, Yuka, et al.
Publicado: (2023)
por: Hashimoto, Yuka, et al.
Publicado: (2023)
Improved Particle Approximation Error for Mean Field Neural Networks
por: Nitanda, Atsushi
Publicado: (2024)
por: Nitanda, Atsushi
Publicado: (2024)
Continual Reinforcement Learning by Planning with Online World Models
por: Liu, Zichen, et al.
Publicado: (2025)
por: Liu, Zichen, et al.
Publicado: (2025)
DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Models
por: Bu, Dake, et al.
Publicado: (2026)
por: Bu, Dake, et al.
Publicado: (2026)
Provably Transformers Harness Multi-Concept Word Semantics for Efficient In-Context Learning
por: Bu, Dake, et al.
Publicado: (2024)
por: Bu, Dake, et al.
Publicado: (2024)
Provable Benefit of Curriculum in Transformer Tree-Reasoning Post-Training
por: Bu, Dake, et al.
Publicado: (2025)
por: Bu, Dake, et al.
Publicado: (2025)
Alternating Diffusion for Proximal Sampling with Zeroth Order Queries
por: Takagi, Hirohane, et al.
Publicado: (2026)
por: Takagi, Hirohane, et al.
Publicado: (2026)
Uniform convergence of the smooth calibration error and its relationship with functional gradient
por: Futami, Futoshi, et al.
Publicado: (2025)
por: Futami, Futoshi, et al.
Publicado: (2025)
How Does Preconditioning Guide Feature Learning in Deep Neural Networks?
por: Yoshida, Kotaro, et al.
Publicado: (2025)
por: Yoshida, Kotaro, et al.
Publicado: (2025)
Provable In-Context Vector Arithmetic via Retrieving Task Concepts
por: Bu, Dake, et al.
Publicado: (2025)
por: Bu, Dake, et al.
Publicado: (2025)
From Saddle Points Toward Global Minima: A Newton-Type Method on Wasserstein Space
por: Lascu, Razvan-Andrei, et al.
Publicado: (2026)
por: Lascu, Razvan-Andrei, et al.
Publicado: (2026)
Post-Training as Reweighting: A Stochastic View of Reasoning Trajectories in Language Models
por: Bu, Dake, et al.
Publicado: (2025)
por: Bu, Dake, et al.
Publicado: (2025)
Hessian-guided Perturbed Wasserstein Gradient Flows for Escaping Saddle Points
por: Yamamoto, Naoya, et al.
Publicado: (2025)
por: Yamamoto, Naoya, et al.
Publicado: (2025)
The Mechanism of Weak-to-Strong Generalization: Feature Elicitation from Latent Knowledge
por: Awano, Ryoya, et al.
Publicado: (2026)
por: Awano, Ryoya, et al.
Publicado: (2026)
Statistical Analysis of the Sinkhorn Iterations for Two-Sample Schrödinger Bridge Estimation
por: Maeda, Ibuki, et al.
Publicado: (2025)
por: Maeda, Ibuki, et al.
Publicado: (2025)
In-Context Learning Is Provably Bayesian Inference: A Generalization Theory for Meta-Learning
por: Wakayama, Tomoya, et al.
Publicado: (2025)
por: Wakayama, Tomoya, et al.
Publicado: (2025)
Slowly Annealed Langevin Dynamics: Theory and Applications to Training-Free Guided Generation
por: Nitanda, Atsushi, et al.
Publicado: (2026)
por: Nitanda, Atsushi, et al.
Publicado: (2026)
Mirror Descent Policy Optimisation for Robust Constrained Markov Decision Processes
por: Bossens, David M., et al.
Publicado: (2025)
por: Bossens, David M., et al.
Publicado: (2025)
State Space Models are Provably Comparable to Transformers in Dynamic Token Selection
por: Nishikawa, Naoki, et al.
Publicado: (2024)
por: Nishikawa, Naoki, et al.
Publicado: (2024)
Direct Density Ratio Optimization: A Statistically Consistent Approach to Aligning Large Language Models
por: Higuchi, Rei, et al.
Publicado: (2025)
por: Higuchi, Rei, et al.
Publicado: (2025)
Transformers as Measure-Theoretic Associative Memory: A Statistical Perspective and Minimax Optimality
por: Kawata, Ryotaro, et al.
Publicado: (2026)
por: Kawata, Ryotaro, et al.
Publicado: (2026)
Transformers Learn Nonlinear Features In Context: Nonconvex Mean-field Dynamics on the Attention Landscape
por: Kim, Juno, et al.
Publicado: (2024)
por: Kim, Juno, et al.
Publicado: (2024)
Deep Two-Way Matrix Reordering for Relational Data Analysis
por: Watanabe, Chihiro, et al.
Publicado: (2021)
por: Watanabe, Chihiro, et al.
Publicado: (2021)
Transformers Provably Solve Parity Efficiently with Chain of Thought
por: Kim, Juno, et al.
Publicado: (2024)
por: Kim, Juno, et al.
Publicado: (2024)
Mean-field Analysis on Two-layer Neural Networks from a Kernel Perspective
por: Takakura, Shokichi, et al.
Publicado: (2024)
por: Takakura, Shokichi, et al.
Publicado: (2024)
AutoLL: Automatic Linear Layout of Graphs based on Deep Neural Network
por: Watanabe, Chihiro, et al.
Publicado: (2021)
por: Watanabe, Chihiro, et al.
Publicado: (2021)
Test time training enhances in-context learning of nonlinear functions
por: Kuwataka, Kento, et al.
Publicado: (2025)
por: Kuwataka, Kento, et al.
Publicado: (2025)
Approximation and Estimation Ability of Transformers for Sequence-to-Sequence Functions with Infinite Dimensional Input
por: Takakura, Shokichi, et al.
Publicado: (2023)
por: Takakura, Shokichi, et al.
Publicado: (2023)
Why is parameter averaging beneficial in SGD? An objective smoothing perspective
por: Nitanda, Atsushi, et al.
Publicado: (2023)
por: Nitanda, Atsushi, et al.
Publicado: (2023)
Constrained Layout Generation with Factor Graphs
por: Dupty, Mohammed Haroon, et al.
Publicado: (2024)
por: Dupty, Mohammed Haroon, et al.
Publicado: (2024)
Wasserstein Convergence Guarantees for a General Class of Score-Based Generative Models
por: Gao, Xuefeng, et al.
Publicado: (2023)
por: Gao, Xuefeng, et al.
Publicado: (2023)
Provably Learning Diffusion Models under the Manifold Hypothesis: Collapse and Refine
por: Huang, Wei, et al.
Publicado: (2026)
por: Huang, Wei, et al.
Publicado: (2026)
EVaDE : Event-Based Variational Thompson Sampling for Model-Based Reinforcement Learning
por: Aravindan, Siddharth, et al.
Publicado: (2025)
por: Aravindan, Siddharth, et al.
Publicado: (2025)
When Scores Learn Geometry: Rate Separations under the Manifold Hypothesis
por: Li, Xiang, et al.
Publicado: (2025)
por: Li, Xiang, et al.
Publicado: (2025)
Degrees of Freedom for Linear Attention: Distilling Softmax Attention with Optimal Feature Efficiency
por: Nishikawa, Naoki, et al.
Publicado: (2025)
por: Nishikawa, Naoki, et al.
Publicado: (2025)
Ejemplares similares
-
Approximation and Generalization Abilities of Score-based Neural Network Generative Models for Sub-Gaussian Distributions
por: Fu, Guoji, et al.
Publicado: (2025) -
Towards a Unified Analysis of Neural Networks in Nonparametric Instrumental Variable Regression: Optimization and Generalization
por: Chen, Zonghao, et al.
Publicado: (2025) -
Direct Distributional Optimization for Provable Alignment of Diffusion Models
por: Kawata, Ryotaro, et al.
Publicado: (2025) -
Implicit Graph Neural Diffusion Networks: Convergence, Generalization, and Over-Smoothing
por: Fu, Guoji, et al.
Publicado: (2023) -
Propagation of Chaos for Mean-Field Langevin Dynamics and its Application to Model Ensemble
por: Nitanda, Atsushi, et al.
Publicado: (2025)