Transformers for Learning on Noisy and Task-Level Manifolds: Approximation and Generalization Insights
Fuente:
arXiv
Guardado en:
| Autores principales: | Shen, Zhaiming, Havrilla, Alex, Lai, Rongjie, Cloninger, Alexander, Liao, Wenjing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Understanding In-Context Learning for Nonlinear Regression with Transformers: Attention as Featurizer
por: Hsu, Alexander, et al.
Publicado: (2026)
por: Hsu, Alexander, et al.
Publicado: (2026)
Understanding In-Context Learning on Structured Manifolds: Bridging Attention to Kernel Methods
por: Shen, Zhaiming, et al.
Publicado: (2025)
por: Shen, Zhaiming, et al.
Publicado: (2025)
Semi-Supervised Manifold Learning with Complexity Decoupled Chart Autoencoders
por: Schonsheck, Stefan C., et al.
Publicado: (2022)
por: Schonsheck, Stefan C., et al.
Publicado: (2022)
Maximal Volume Matrix Cross Approximation for Image Compression and Least Squares Solution
por: Allen, Kenneth, et al.
Publicado: (2023)
por: Allen, Kenneth, et al.
Publicado: (2023)
A Kernel-based Stochastic Approximation Framework for Nonlinear Operator Learning
por: Yang, Jia-Qi, et al.
Publicado: (2025)
por: Yang, Jia-Qi, et al.
Publicado: (2025)
Faster Diffusion Models via Higher-Order Approximation
por: Li, Gen, et al.
Publicado: (2025)
por: Li, Gen, et al.
Publicado: (2025)
Dimension-Free Convergence of Diffusion Models for Approximate Gaussian Mixtures
por: Li, Gen, et al.
Publicado: (2025)
por: Li, Gen, et al.
Publicado: (2025)
Simultaneous Approximation of the Score Function and Its Derivatives by Deep Neural Networks
por: Yakovlev, Konstantin, et al.
Publicado: (2025)
por: Yakovlev, Konstantin, et al.
Publicado: (2025)
Data-Driven Model Reduction using WeldNet: Windowed Encoders for Learning Dynamics
por: Dahal, Biraj, et al.
Publicado: (2025)
por: Dahal, Biraj, et al.
Publicado: (2025)
Graph-based Semi-supervised Local Clustering with Few Labeled Nodes
por: Shen, Zhaiming, et al.
Publicado: (2022)
por: Shen, Zhaiming, et al.
Publicado: (2022)
Convergence Rates for Learning Pseudo-Differential Operators
por: Chen, Jiaheng, et al.
Publicado: (2026)
por: Chen, Jiaheng, et al.
Publicado: (2026)
Towards Sharp Minimax Risk Bounds for Operator Learning
por: Adcock, Ben, et al.
Publicado: (2025)
por: Adcock, Ben, et al.
Publicado: (2025)
On Spectral Learning for Odeco Tensors: Perturbation, Initialization, and Algorithms
por: Auddy, Arnab, et al.
Publicado: (2025)
por: Auddy, Arnab, et al.
Publicado: (2025)
Data-driven Learning of Interaction Laws in Multispecies Particle Systems with Gaussian Processes: Convergence Theory and Applications
por: Feng, Jinchao, et al.
Publicado: (2025)
por: Feng, Jinchao, et al.
Publicado: (2025)
Characteristic Learning for Provable One Step Generation
por: Ding, Zhao, et al.
Publicado: (2024)
por: Ding, Zhao, et al.
Publicado: (2024)
Enabling stratified sampling in high dimensions via nonlinear dimensionality reduction
por: Geraci, Gianluca, et al.
Publicado: (2025)
por: Geraci, Gianluca, et al.
Publicado: (2025)
Posterior Covariance Structures in Gaussian Processes
por: Cai, Difeng, et al.
Publicado: (2024)
por: Cai, Difeng, et al.
Publicado: (2024)
Linear cost and exponentially convergent approximation of Gaussian Matérn processes on intervals
por: Bolin, David, et al.
Publicado: (2024)
por: Bolin, David, et al.
Publicado: (2024)
Gaussian Process Regression under Computational and Epistemic Misspecification
por: Sanz-Alonso, Daniel, et al.
Publicado: (2023)
por: Sanz-Alonso, Daniel, et al.
Publicado: (2023)
Optimal Recovery Meets Minimax Estimation
por: DeVore, Ronald, et al.
Publicado: (2025)
por: DeVore, Ronald, et al.
Publicado: (2025)
Solving PDEs on Spheres with Physics-Informed Convolutional Neural Networks
por: Lei, Guanhang, et al.
Publicado: (2023)
por: Lei, Guanhang, et al.
Publicado: (2023)
Can Linear Probes Measure LLM Uncertainty?
por: Dakhmouche, Ramzi, et al.
Publicado: (2025)
por: Dakhmouche, Ramzi, et al.
Publicado: (2025)
Score Operator Newton transport
por: Chandramoorthy, Nisha, et al.
Publicado: (2023)
por: Chandramoorthy, Nisha, et al.
Publicado: (2023)
Weighted least-squares approximation with determinantal point processes and generalized volume sampling
por: Nouy, Anthony, et al.
Publicado: (2023)
por: Nouy, Anthony, et al.
Publicado: (2023)
Posterior Concentration of Bayesian Physics-Informed Neural Networks for Elliptic PDEs
por: Zhao, Yuxuan, et al.
Publicado: (2026)
por: Zhao, Yuxuan, et al.
Publicado: (2026)
Benign overfitting in Fixed Dimension via Physics-Informed Learning with Smooth Inductive Bias
por: Wong, Honam, et al.
Publicado: (2024)
por: Wong, Honam, et al.
Publicado: (2024)
The Optimal Linear B-splines Approximation via Kolmogorov Superposition Theorem and its Application
por: Lai, Ming-Jun, et al.
Publicado: (2024)
por: Lai, Ming-Jun, et al.
Publicado: (2024)
The Kolmogorov Superposition Theorem can Break the Curse of Dimensionality When Approximating High Dimensional Functions
por: Lai, Ming-Jun, et al.
Publicado: (2021)
por: Lai, Ming-Jun, et al.
Publicado: (2021)
Which Spaces can be Embedded in $L_p$-type Reproducing Kernel Banach Space? A Characterization via Metric Entropy
por: Lu, Yiping, et al.
Publicado: (2024)
por: Lu, Yiping, et al.
Publicado: (2024)
Stochastic Regret Guarantees for Online Zeroth- and First-Order Bilevel Optimization
por: Nazari, Parvin, et al.
Publicado: (2025)
por: Nazari, Parvin, et al.
Publicado: (2025)
Sampling and estimation on manifolds using the Langevin diffusion
por: Bharath, Karthik, et al.
Publicado: (2023)
por: Bharath, Karthik, et al.
Publicado: (2023)
Perturbation Bounds for Low-Rank Inverse Approximations under Noise
por: Tran, Phuc, et al.
Publicado: (2025)
por: Tran, Phuc, et al.
Publicado: (2025)
Stein transport for Bayesian inference
por: Nüsken, Nikolas
Publicado: (2024)
por: Nüsken, Nikolas
Publicado: (2024)
Gaussian Processes and Reproducing Kernels: Connections and Equivalences
por: Kanagawa, Motonobu, et al.
Publicado: (2025)
por: Kanagawa, Motonobu, et al.
Publicado: (2025)
Samplet limits and multiwavelets
por: Giacchi, Gianluca, et al.
Publicado: (2026)
por: Giacchi, Gianluca, et al.
Publicado: (2026)
Analysis of singular subspaces under random perturbations
por: Wang, Ke
Publicado: (2024)
por: Wang, Ke
Publicado: (2024)
Tensor Methods in High Dimensional Data Analysis: Opportunities and Challenges
por: Auddy, Arnab, et al.
Publicado: (2024)
por: Auddy, Arnab, et al.
Publicado: (2024)
Revisit CP Tensor Decomposition: Statistical Optimality and Fast Convergence
por: Tang, Runshi, et al.
Publicado: (2025)
por: Tang, Runshi, et al.
Publicado: (2025)
Provable Diffusion Posterior Sampling for Bayesian Inversion
por: Chang, Jinyuan, et al.
Publicado: (2025)
por: Chang, Jinyuan, et al.
Publicado: (2025)
Wedge Sampling: Efficient Tensor Completion with Nearly-Linear Sample Complexity
por: Luo, Hengrui, et al.
Publicado: (2026)
por: Luo, Hengrui, et al.
Publicado: (2026)
Ejemplares similares
-
Understanding In-Context Learning for Nonlinear Regression with Transformers: Attention as Featurizer
por: Hsu, Alexander, et al.
Publicado: (2026) -
Understanding In-Context Learning on Structured Manifolds: Bridging Attention to Kernel Methods
por: Shen, Zhaiming, et al.
Publicado: (2025) -
Semi-Supervised Manifold Learning with Complexity Decoupled Chart Autoencoders
por: Schonsheck, Stefan C., et al.
Publicado: (2022) -
Maximal Volume Matrix Cross Approximation for Image Compression and Least Squares Solution
por: Allen, Kenneth, et al.
Publicado: (2023) -
A Kernel-based Stochastic Approximation Framework for Nonlinear Operator Learning
por: Yang, Jia-Qi, et al.
Publicado: (2025)