Measure-to-measure interpolation using Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Geshkovski, Borjan, Rigollet, Philippe, Ruiz-Balet, Domènec |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Constructive conditional normalizing flows
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2026)
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2026)
Constructive approximate transport maps with normalizing flows
von: Álvarez-López, Antonio, et al.
Veröffentlicht: (2024)
von: Álvarez-López, Antonio, et al.
Veröffentlicht: (2024)
Perceptrons and localization of attention's mean-field landscape
von: Álvarez-López, Antonio, et al.
Veröffentlicht: (2026)
von: Álvarez-López, Antonio, et al.
Veröffentlicht: (2026)
Attention's forward pass and Frank-Wolfe
von: Alcalde, Albert, et al.
Veröffentlicht: (2025)
von: Alcalde, Albert, et al.
Veröffentlicht: (2025)
Homogenized Transformers
von: Koubbi, Hugo, et al.
Veröffentlicht: (2026)
von: Koubbi, Hugo, et al.
Veröffentlicht: (2026)
Pattern control via Diffussion interaction
von: Ruiz-Balet, Domènec, et al.
Veröffentlicht: (2024)
von: Ruiz-Balet, Domènec, et al.
Veröffentlicht: (2024)
On the number of modes of Gaussian kernel density estimators
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2024)
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2024)
Splat Regression Models
von: Daniels, Mara, et al.
Veröffentlicht: (2025)
von: Daniels, Mara, et al.
Veröffentlicht: (2025)
YuriiFormer: A Suite of Nesterov-Accelerated Transformers
von: Zimin, Aleksandr, et al.
Veröffentlicht: (2026)
von: Zimin, Aleksandr, et al.
Veröffentlicht: (2026)
A mathematical perspective on Transformers
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2023)
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2023)
A Total Variation Flow Scheme for Ergodic Mean Field Games
von: Kalise, Dante, et al.
Veröffentlicht: (2024)
von: Kalise, Dante, et al.
Veröffentlicht: (2024)
The emergence of clusters in self-attention dynamics
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2023)
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2023)
Mean-field games for harvesting problems: Uniqueness, long-time behaviour and weak KAM theory
von: Kobeissi, Ziad, et al.
Veröffentlicht: (2024)
von: Kobeissi, Ziad, et al.
Veröffentlicht: (2024)
Dynamic metastability in the self-attention model
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2024)
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2024)
Synchronization of mean-field models on the circle
von: Polyanskiy, Yury, et al.
Veröffentlicht: (2025)
von: Polyanskiy, Yury, et al.
Veröffentlicht: (2025)
Propagation of Chaos in Contextual Flow Maps
von: Chen, Shi, et al.
Veröffentlicht: (2026)
von: Chen, Shi, et al.
Veröffentlicht: (2026)
Size and depth of monotone neural networks: interpolation and approximation
von: Mikulincer, Dan, et al.
Veröffentlicht: (2022)
von: Mikulincer, Dan, et al.
Veröffentlicht: (2022)
Approximation and interpolation of deep neural networks
von: Constantinescu, Vlad-Raul, et al.
Veröffentlicht: (2023)
von: Constantinescu, Vlad-Raul, et al.
Veröffentlicht: (2023)
Kinetic theory for Transformers and the lost-in-the-middle phenomenon
von: Duerinckx, Mitia, et al.
Veröffentlicht: (2026)
von: Duerinckx, Mitia, et al.
Veröffentlicht: (2026)
Optimal Transport with Tempered Exponential Measures
von: Amid, Ehsan, et al.
Veröffentlicht: (2023)
von: Amid, Ehsan, et al.
Veröffentlicht: (2023)
Unraveling the Gradient Descent Dynamics of Transformers
von: Song, Bingqing, et al.
Veröffentlicht: (2024)
von: Song, Bingqing, et al.
Veröffentlicht: (2024)
An Optimal Control Approach To Transformer Training
von: Akman, Kağan, et al.
Veröffentlicht: (2026)
von: Akman, Kağan, et al.
Veröffentlicht: (2026)
Stochastic Approximation Methods for Distortion Risk Measure Optimization
von: Jiang, Jinyang, et al.
Veröffentlicht: (2025)
von: Jiang, Jinyang, et al.
Veröffentlicht: (2025)
Random Coordinate Descent on the Wasserstein Space of Probability Measures
von: Xu, Yewei, et al.
Veröffentlicht: (2026)
von: Xu, Yewei, et al.
Veröffentlicht: (2026)
Sliced Wasserstein Steering between Gaussian Measures
von: Ito, Kaito, et al.
Veröffentlicht: (2026)
von: Ito, Kaito, et al.
Veröffentlicht: (2026)
Incremental Learning of Sparse Attention Patterns in Transformers
von: Yüksel, Oğuz Kaan, et al.
Veröffentlicht: (2026)
von: Yüksel, Oğuz Kaan, et al.
Veröffentlicht: (2026)
On the Effectiveness of the z-Transform Method in Quadratic Optimization
von: Bach, Francis
Veröffentlicht: (2025)
von: Bach, Francis
Veröffentlicht: (2025)
Understanding Lookahead Dynamics Through Laplace Transform
von: Sanyal, Aniket, et al.
Veröffentlicht: (2025)
von: Sanyal, Aniket, et al.
Veröffentlicht: (2025)
DLMMPR:Deep Learning-based Measurement Matrix for Phase Retrieval
von: Liu, Jing, et al.
Veröffentlicht: (2025)
von: Liu, Jing, et al.
Veröffentlicht: (2025)
How Transformers Get Rich: Approximation and Dynamics Analysis
von: Wang, Mingze, et al.
Veröffentlicht: (2024)
von: Wang, Mingze, et al.
Veröffentlicht: (2024)
Near-Optimal Real-Time Personalization with Simple Transformers
von: An, Lin, et al.
Veröffentlicht: (2025)
von: An, Lin, et al.
Veröffentlicht: (2025)
Probabilistic Smoothing with Ratio-Monotone Transforms for Global Optimization
von: Jang, Kukyoung, et al.
Veröffentlicht: (2026)
von: Jang, Kukyoung, et al.
Veröffentlicht: (2026)
On the Convergence of Gradient Descent on Learning Transformers with Residual Connections
von: Qin, Zhen, et al.
Veröffentlicht: (2025)
von: Qin, Zhen, et al.
Veröffentlicht: (2025)
Q-Measure-Learning for Continuous State RL: Efficient Implementation and Convergence
von: Wang, Shengbo
Veröffentlicht: (2026)
von: Wang, Shengbo
Veröffentlicht: (2026)
Mixtures Closest to a Given Measure: A Semidefinite Programming Approach
von: Đurašinović, Srećko, et al.
Veröffentlicht: (2025)
von: Đurašinović, Srećko, et al.
Veröffentlicht: (2025)
Bayesian Optimization with Structured Measurements: A Vector-Valued RKHS Framework
von: Wang, Wenbin, et al.
Veröffentlicht: (2026)
von: Wang, Wenbin, et al.
Veröffentlicht: (2026)
A Similarity Measure Between Functions with Applications to Statistical Learning and Optimization
von: Huang, Chengpiao, et al.
Veröffentlicht: (2025)
von: Huang, Chengpiao, et al.
Veröffentlicht: (2025)
A Theoretical Analysis of Self-Supervised Learning for Vision Transformers
von: Huang, Yu, et al.
Veröffentlicht: (2024)
von: Huang, Yu, et al.
Veröffentlicht: (2024)
TwIST: Rigging the Lottery in Transformers with Independent Subnetwork Training
von: Menezes, Michael, et al.
Veröffentlicht: (2025)
von: Menezes, Michael, et al.
Veröffentlicht: (2025)
Mean-Field Langevin Dynamics for Signed Measures via a Bilevel Approach
von: Wang, Guillaume, et al.
Veröffentlicht: (2024)
von: Wang, Guillaume, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Constructive conditional normalizing flows
von: Geshkovski, Borjan, et al.
Veröffentlicht: (2026) -
Constructive approximate transport maps with normalizing flows
von: Álvarez-López, Antonio, et al.
Veröffentlicht: (2024) -
Perceptrons and localization of attention's mean-field landscape
von: Álvarez-López, Antonio, et al.
Veröffentlicht: (2026) -
Attention's forward pass and Frank-Wolfe
von: Alcalde, Albert, et al.
Veröffentlicht: (2025) -
Homogenized Transformers
von: Koubbi, Hugo, et al.
Veröffentlicht: (2026)