Perceptrons and localization of attention's mean-field landscape
Fuente:
arXiv
Saved in:
| Main Authors: | Álvarez-López, Antonio, Geshkovski, Borjan, Ruiz-Balet, Domènec |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Constructive approximate transport maps with normalizing flows
by: Álvarez-López, Antonio, et al.
Published: (2024)
by: Álvarez-López, Antonio, et al.
Published: (2024)
Constructive conditional normalizing flows
by: Geshkovski, Borjan, et al.
Published: (2026)
by: Geshkovski, Borjan, et al.
Published: (2026)
Measure-to-measure interpolation using Transformers
by: Geshkovski, Borjan, et al.
Published: (2024)
by: Geshkovski, Borjan, et al.
Published: (2024)
Attention's forward pass and Frank-Wolfe
by: Alcalde, Albert, et al.
Published: (2025)
by: Alcalde, Albert, et al.
Published: (2025)
Pattern control via Diffussion interaction
by: Ruiz-Balet, Domènec, et al.
Published: (2024)
by: Ruiz-Balet, Domènec, et al.
Published: (2024)
A Total Variation Flow Scheme for Ergodic Mean Field Games
by: Kalise, Dante, et al.
Published: (2024)
by: Kalise, Dante, et al.
Published: (2024)
Mean-field games for harvesting problems: Uniqueness, long-time behaviour and weak KAM theory
by: Kobeissi, Ziad, et al.
Published: (2024)
by: Kobeissi, Ziad, et al.
Published: (2024)
Morphological Perceptron with Competitive Layer: Training Using Convex-Concave Procedure
by: Cunha, Iara, et al.
Published: (2025)
by: Cunha, Iara, et al.
Published: (2025)
Meta-reinforcement learning with minimum attention
by: Gupta, Shashank, et al.
Published: (2025)
by: Gupta, Shashank, et al.
Published: (2025)
Gradient Descent as a Perceptron Algorithm: Understanding Dynamics and Implicit Acceleration
by: Tyurin, Alexander
Published: (2025)
by: Tyurin, Alexander
Published: (2025)
The emergence of clusters in self-attention dynamics
by: Geshkovski, Borjan, et al.
Published: (2023)
by: Geshkovski, Borjan, et al.
Published: (2023)
Implicit Bias and Fast Convergence Rates for Self-attention
by: Vasudeva, Bhavya, et al.
Published: (2024)
by: Vasudeva, Bhavya, et al.
Published: (2024)
Unified continuous-time q-learning for mean-field game and mean-field control problems
by: Wei, Xiaoli, et al.
Published: (2024)
by: Wei, Xiaoli, et al.
Published: (2024)
Nesterov acceleration in benignly non-convex landscapes
by: Gupta, Kanan, et al.
Published: (2024)
by: Gupta, Kanan, et al.
Published: (2024)
Mirror Descent-Ascent for mean-field min-max problems
by: Lascu, Razvan-Andrei, et al.
Published: (2024)
by: Lascu, Razvan-Andrei, et al.
Published: (2024)
Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems
by: Yang, Xianjin, et al.
Published: (2025)
by: Yang, Xianjin, et al.
Published: (2025)
Dynamic metastability in the self-attention model
by: Geshkovski, Borjan, et al.
Published: (2024)
by: Geshkovski, Borjan, et al.
Published: (2024)
From NeurODEs to AutoencODEs: a mean-field control framework for width-varying Neural Networks
by: Cipriani, Cristina, et al.
Published: (2023)
by: Cipriani, Cristina, et al.
Published: (2023)
On propagation of chaos for the Fisher-Rao gradient flow in entropic mean-field optimization
by: Lazić, Petra, et al.
Published: (2026)
by: Lazić, Petra, et al.
Published: (2026)
Non-convex entropic mean-field optimization via Best Response flow
by: Lascu, Razvan-Andrei, et al.
Published: (2025)
by: Lascu, Razvan-Andrei, et al.
Published: (2025)
A Fisher-Rao gradient flow for entropic mean-field min-max games
by: Lascu, Razvan-Andrei, et al.
Published: (2024)
by: Lascu, Razvan-Andrei, et al.
Published: (2024)
Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise
by: Frikha, Noufel, et al.
Published: (2024)
by: Frikha, Noufel, et al.
Published: (2024)
Continuous-time q-learning for mean-field control problems
by: Wei, Xiaoli, et al.
Published: (2023)
by: Wei, Xiaoli, et al.
Published: (2023)
Function approximation by neural nets in the mean-field regime: Entropic regularization and controlled McKean-Vlasov dynamics
by: Tzen, Belinda, et al.
Published: (2020)
by: Tzen, Belinda, et al.
Published: (2020)
Modified K-means Algorithm with Local Optimality Guarantees
by: Li, Mingyi, et al.
Published: (2025)
by: Li, Mingyi, et al.
Published: (2025)
Bayesian inference of mean velocity fields and turbulence models from flow MRI
by: Kontogiannis, A., et al.
Published: (2024)
by: Kontogiannis, A., et al.
Published: (2024)
Algorithms for mean-field variational inference via polyhedral optimization in the Wasserstein space
by: Jiang, Yiheng, et al.
Published: (2023)
by: Jiang, Yiheng, et al.
Published: (2023)
Scalable Second-order Riemannian Optimization for $K$-means Clustering
by: Xu, Peng, et al.
Published: (2025)
by: Xu, Peng, et al.
Published: (2025)
Sinkhorn doubly stochastic attention rank decay analysis
by: Lapenna, Michela, et al.
Published: (2026)
by: Lapenna, Michela, et al.
Published: (2026)
Kernel-based potential mean-field games with unbiased random Fourier $U$-statistics
by: Nakano, Yumiharu
Published: (2026)
by: Nakano, Yumiharu
Published: (2026)
Convergence of two-timescale gradient descent ascent dynamics: finite-dimensional and mean-field perspectives
by: An, Jing, et al.
Published: (2025)
by: An, Jing, et al.
Published: (2025)
Synchronization of mean-field models on the circle
by: Polyanskiy, Yury, et al.
Published: (2025)
by: Polyanskiy, Yury, et al.
Published: (2025)
Statistically Optimal K-means Clustering via Nonnegative Low-rank Semidefinite Programming
by: Zhuang, Yubo, et al.
Published: (2023)
by: Zhuang, Yubo, et al.
Published: (2023)
Linear attention is (maybe) all you need (to understand transformer optimization)
by: Ahn, Kwangjun, et al.
Published: (2023)
by: Ahn, Kwangjun, et al.
Published: (2023)
Continuous-time q-learning for mean-field control with common noise, part-I: Theoretical foundations
by: Ren, Zhenjie, et al.
Published: (2026)
by: Ren, Zhenjie, et al.
Published: (2026)
Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms
by: Ren, Zhenjie, et al.
Published: (2026)
by: Ren, Zhenjie, et al.
Published: (2026)
Novel clustered federated learning based on local loss
by: Gu, Endong, et al.
Published: (2024)
by: Gu, Endong, et al.
Published: (2024)
Symmetric Mean-field Langevin Dynamics for Distributional Minimax Problems
by: Kim, Juno, et al.
Published: (2023)
by: Kim, Juno, et al.
Published: (2023)
Rolling Ball Optimizer: Learning by ironing out loss landscape wrinkles
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2025)
by: Belgoumri, Mohammed Djameleddine, et al.
Published: (2025)
Population-Aware Imitation Learning in Mean-field Games with Common Noise
by: Lambrecht, Grégoire, et al.
Published: (2026)
by: Lambrecht, Grégoire, et al.
Published: (2026)
Similar Items
-
Constructive approximate transport maps with normalizing flows
by: Álvarez-López, Antonio, et al.
Published: (2024) -
Constructive conditional normalizing flows
by: Geshkovski, Borjan, et al.
Published: (2026) -
Measure-to-measure interpolation using Transformers
by: Geshkovski, Borjan, et al.
Published: (2024) -
Attention's forward pass and Frank-Wolfe
by: Alcalde, Albert, et al.
Published: (2025) -
Pattern control via Diffussion interaction
by: Ruiz-Balet, Domènec, et al.
Published: (2024)