Transformers Meet In-Context Learning: A Universal Approximation Theory
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Gen, Jiao, Yuchen, Huang, Yu, Wei, Yuting, Chen, Yuxin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Faster Diffusion Models via Higher-Order Approximation
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Towards a unified framework for guided diffusion models
by: Jiao, Yuchen, et al.
Published: (2025)
by: Jiao, Yuchen, et al.
Published: (2025)
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology
by: Jiao, Yuchen, et al.
Published: (2025)
by: Jiao, Yuchen, et al.
Published: (2025)
Provable Efficiency of Guidance in Diffusion Models for General Data Distribution
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
A Sharp Convergence Theory for The Probability Flow ODEs of Diffusion Models
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Towards Faster Non-Asymptotic Convergence for Diffusion-Based Generative Models
by: Li, Gen, et al.
Published: (2023)
by: Li, Gen, et al.
Published: (2023)
Optimal Convergence Analysis of DDPM for General Distributions
by: Jiao, Yuchen, et al.
Published: (2025)
by: Jiao, Yuchen, et al.
Published: (2025)
Breaking the Sample Size Barrier in Model-Based Reinforcement Learning with a Generative Model
by: Li, Gen, et al.
Published: (2020)
by: Li, Gen, et al.
Published: (2020)
Dimension-Free Convergence of Diffusion Models for Approximate Gaussian Mixtures
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Is Q-Learning Minimax Optimal? A Tight Sample Complexity Analysis
by: Li, Gen, et al.
Published: (2021)
by: Li, Gen, et al.
Published: (2021)
The Curious Price of Distributional Robustness in Reinforcement Learning with a Generative Model
by: Shi, Laixi, et al.
Published: (2023)
by: Shi, Laixi, et al.
Published: (2023)
Settling the Sample Complexity of Model-Based Offline Reinforcement Learning
by: Li, Gen, et al.
Published: (2022)
by: Li, Gen, et al.
Published: (2022)
Towards a mathematical theory for consistency training in diffusion models
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Are First-Order Diffusion Samplers Really Slower? A Fast Forward-Value Approach
by: Jiao, Yuchen, et al.
Published: (2025)
by: Jiao, Yuchen, et al.
Published: (2025)
A non-asymptotic distributional theory of approximate message passing for sparse and robust regression
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
Minimax-Optimal Reward-Agnostic Exploration in Reinforcement Learning
by: Li, Gen, et al.
Published: (2023)
by: Li, Gen, et al.
Published: (2023)
Statistical and Algorithmic Foundations of Reinforcement Learning
by: Chi, Yuejie, et al.
Published: (2025)
by: Chi, Yuejie, et al.
Published: (2025)
Denoising diffusion probabilistic models are optimally adaptive to unknown low dimensionality
by: Huang, Zhihan, et al.
Published: (2024)
by: Huang, Zhihan, et al.
Published: (2024)
Deflated HeteroPCA: Overcoming the curse of ill-conditioning in heteroskedastic PCA
by: Zhou, Yuchen, et al.
Published: (2023)
by: Zhou, Yuchen, et al.
Published: (2023)
Model-Based Reinforcement Learning for Offline Zero-Sum Markov Games
by: Yan, Yuling, et al.
Published: (2022)
by: Yan, Yuling, et al.
Published: (2022)
A Theory of Universal Agnostic Learning
by: Hanneke, Steve, et al.
Published: (2026)
by: Hanneke, Steve, et al.
Published: (2026)
A Statistical Theory of Contrastive Learning via Approximate Sufficient Statistics
by: Lin, Licong, et al.
Published: (2025)
by: Lin, Licong, et al.
Published: (2025)
O(d/T) Convergence Theory for Diffusion Probabilistic Models under Minimal Assumptions
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
A Survey on Statistical Theory of Deep Learning: Approximation, Training Dynamics, and Generative Models
by: Suh, Namjoon, et al.
Published: (2024)
by: Suh, Namjoon, et al.
Published: (2024)
Efficient Sampling with Discrete Diffusion Models: Sharp and Adaptive Guarantees
by: Dmitriev, Daniil, et al.
Published: (2026)
by: Dmitriev, Daniil, et al.
Published: (2026)
Learning single index model with gradient descent: spectral initialization and precise asymptotics
by: Chen, Yuchen, et al.
Published: (2025)
by: Chen, Yuchen, et al.
Published: (2025)
High-probability sample complexities for policy evaluation with linear function approximation
by: Li, Gen, et al.
Published: (2023)
by: Li, Gen, et al.
Published: (2023)
Breaking AR's Sampling Bottleneck: Provable Acceleration via Diffusion Language Models
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Minimax Optimality of the Probability Flow ODE for Diffusion Models
by: Cai, Changxiao, et al.
Published: (2025)
by: Cai, Changxiao, et al.
Published: (2025)
Learning Spectral Methods by Transformers
by: He, Yihan, et al.
Published: (2025)
by: He, Yihan, et al.
Published: (2025)
Diffusion Transformer Captures Spatial-Temporal Dependencies: A Theory for Gaussian Process Data
by: Fu, Hengyu, et al.
Published: (2024)
by: Fu, Hengyu, et al.
Published: (2024)
On Transferring Transferability: Towards a Theory for Size Generalization
by: Levin, Eitan, et al.
Published: (2025)
by: Levin, Eitan, et al.
Published: (2025)
Universal Inference Meets Random Projections: A Scalable Test for Log-concavity
by: Dunn, Robin, et al.
Published: (2021)
by: Dunn, Robin, et al.
Published: (2021)
Statistical Inference under Adaptive Sampling with LinUCB
by: Fan, Wei, et al.
Published: (2025)
by: Fan, Wei, et al.
Published: (2025)
Transformers for Learning on Noisy and Task-Level Manifolds: Approximation and Generalization Insights
by: Shen, Zhaiming, et al.
Published: (2025)
by: Shen, Zhaiming, et al.
Published: (2025)
Ordinary Least Squares is a Special Case of Transformer
by: Tan, Xiaojun, et al.
Published: (2026)
by: Tan, Xiaojun, et al.
Published: (2026)
When Does Model Collapse Occur in Structured Interactive Learning?
by: Wu, Yuchen, et al.
Published: (2026)
by: Wu, Yuchen, et al.
Published: (2026)
Adapting to Unknown Low-Dimensional Structures in Score-Based Diffusion Models
by: Li, Gen, et al.
Published: (2024)
by: Li, Gen, et al.
Published: (2024)
On Universality of Non-Separable Approximate Message Passing Algorithms
by: Lovig, Max, et al.
Published: (2025)
by: Lovig, Max, et al.
Published: (2025)
A Theory of Feature Learning in Kernel Models
by: Chen, Yunlu, et al.
Published: (2023)
by: Chen, Yunlu, et al.
Published: (2023)
Similar Items
-
Faster Diffusion Models via Higher-Order Approximation
by: Li, Gen, et al.
Published: (2025) -
Towards a unified framework for guided diffusion models
by: Jiao, Yuchen, et al.
Published: (2025) -
Connections between reinforcement learning with feedback,test-time scaling, and diffusion guidance: An anthology
by: Jiao, Yuchen, et al.
Published: (2025) -
Provable Efficiency of Guidance in Diffusion Models for General Data Distribution
by: Li, Gen, et al.
Published: (2025) -
A Sharp Convergence Theory for The Probability Flow ODEs of Diffusion Models
by: Li, Gen, et al.
Published: (2024)