The Asymptotic Behavior of Attention in Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Abella, Álvaro Rodríguez, Silvestre, João Pedro, Tabuada, Paulo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gain Scheduling with a Neural Operator for a Transport PDE with Nonlinear Recirculation
by: Lamarque, Maxence, et al.
Published: (2024)
by: Lamarque, Maxence, et al.
Published: (2024)
Distributed Optimization via Energy Conservation Laws in Dilated Coordinates
by: Baranwal, Mayank, et al.
Published: (2024)
by: Baranwal, Mayank, et al.
Published: (2024)
Adaptive Neural-Operator Backstepping Control of a Benchmark Hyperbolic PDE
by: Lamarque, Maxence, et al.
Published: (2024)
by: Lamarque, Maxence, et al.
Published: (2024)
Analysis of Thompson Sampling for Controlling Unknown Linear Diffusion Processes
by: Faradonbeh, Mohamad Kazem Shirani, et al.
Published: (2022)
by: Faradonbeh, Mohamad Kazem Shirani, et al.
Published: (2022)
Universal Approximation Power of Deep Residual Neural Networks via Nonlinear Control Theory
by: Tabuada, Paulo, et al.
Published: (2020)
by: Tabuada, Paulo, et al.
Published: (2020)
A Framework for Time-Varying Optimization via Derivative Estimation
by: Marchi, Matteo, et al.
Published: (2024)
by: Marchi, Matteo, et al.
Published: (2024)
Heat Death of Generative Models in Closed-Loop Learning
by: Marchi, Matteo, et al.
Published: (2024)
by: Marchi, Matteo, et al.
Published: (2024)
Koopman-Assisted Reinforcement Learning
by: Rozwood, Preston, et al.
Published: (2024)
by: Rozwood, Preston, et al.
Published: (2024)
Neural Hamiltonian Operator
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Stable Walking for Bipedal Locomotion under Foot-Slip via Virtual Nonholonomic Constraints
by: Colombo, Leonardo, et al.
Published: (2026)
by: Colombo, Leonardo, et al.
Published: (2026)
Control Barrier Function based Quadratic Programs Introduce Undesirable Asymptotically Stable Equilibria
by: Reis, Matheus F., et al.
Published: (2020)
by: Reis, Matheus F., et al.
Published: (2020)
SINDy-RL: Interpretable and Efficient Model-Based Reinforcement Learning
by: Zolman, Nicholas, et al.
Published: (2024)
by: Zolman, Nicholas, et al.
Published: (2024)
Data-driven Nonlinear Model Reduction using Koopman Theory: Integrated Control Form and NMPC Case Study
by: Schulze, Jan C., et al.
Published: (2024)
by: Schulze, Jan C., et al.
Published: (2024)
Stability properties of gradient flow dynamics for the symmetric low-rank matrix factorization problem
by: Mohammadi, Hesameddin, et al.
Published: (2024)
by: Mohammadi, Hesameddin, et al.
Published: (2024)
From exponential to finite/fixed-time stability: Applications to optimization
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
Neural Operators for Predictor Feedback Control of Nonlinear Delay Systems
by: Bhan, Luke, et al.
Published: (2024)
by: Bhan, Luke, et al.
Published: (2024)
Identifiability of Differential-Algebraic Systems
by: Montanari, Arthur N., et al.
Published: (2024)
by: Montanari, Arthur N., et al.
Published: (2024)
Safe and Robust Domains of Attraction for Discrete-Time Systems: A Set-Based Characterization and Certifiable Neural Network Estimation
by: Serry, Mohamed, et al.
Published: (2026)
by: Serry, Mohamed, et al.
Published: (2026)
Near-Optimal Distributed Linear-Quadratic Regulator for Networked Systems
by: Shin, Sungho, et al.
Published: (2022)
by: Shin, Sungho, et al.
Published: (2022)
Learning Dissipative Neural Dynamical Systems
by: Xu, Yuezhu, et al.
Published: (2023)
by: Xu, Yuezhu, et al.
Published: (2023)
Tradeoffs between convergence rate and noise amplification for momentum-based accelerated optimization algorithms
by: Mohammadi, Hesameddin, et al.
Published: (2022)
by: Mohammadi, Hesameddin, et al.
Published: (2022)
Observability conditions for neural state-space models with eigenvalues and their roots of unity
by: Gracyk, Andrew
Published: (2025)
by: Gracyk, Andrew
Published: (2025)
Safely Learning Dynamical Systems
by: Ahmadi, Amir Ali, et al.
Published: (2023)
by: Ahmadi, Amir Ali, et al.
Published: (2023)
On the Hopf-Cole Transform for Control-affine Schrödinger Bridge
by: Teter, Alexis, et al.
Published: (2025)
by: Teter, Alexis, et al.
Published: (2025)
Model Predictive Control of Hybrid Dynamical Systems
by: Sanfelice, Ricardo G., et al.
Published: (2026)
by: Sanfelice, Ricardo G., et al.
Published: (2026)
Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization
by: Murillo-Gonzalez, Alejandro, et al.
Published: (2026)
by: Murillo-Gonzalez, Alejandro, et al.
Published: (2026)
Online Optimization and Ambiguity-based Learning of Distributionally Uncertain Dynamic Systems
by: Li, Dan, et al.
Published: (2021)
by: Li, Dan, et al.
Published: (2021)
Safe and Near-Optimal Control with Online Dynamics Learning
by: Prajapat, Manish, et al.
Published: (2025)
by: Prajapat, Manish, et al.
Published: (2025)
Asymptotic stability equals exponential stability -- while you twist your eyes
by: Jongeneel, Wouter
Published: (2024)
by: Jongeneel, Wouter
Published: (2024)
Moving-Horizon Estimators for Hyperbolic and Parabolic PDEs in 1-D
by: Bhan, Luke, et al.
Published: (2024)
by: Bhan, Luke, et al.
Published: (2024)
Nonadaptive Output Regulation of Second-Order Nonlinear Uncertain Systems
by: Lu, Maobin, et al.
Published: (2025)
by: Lu, Maobin, et al.
Published: (2025)
Stochastic Approximation with Delayed Updates: Finite-Time Rates under Markovian Sampling
by: Adibi, Arman, et al.
Published: (2024)
by: Adibi, Arman, et al.
Published: (2024)
DeMuon: A Decentralized Muon for Matrix Optimization over Graphs
by: He, Chuan, et al.
Published: (2025)
by: He, Chuan, et al.
Published: (2025)
Control of Medical Digital Twins with Artificial Neural Networks
by: Böttcher, Lucas, et al.
Published: (2024)
by: Böttcher, Lucas, et al.
Published: (2024)
Model-Based Reinforcement Learning Control of Reaction-Diffusion Problems
by: Schenk, Christina, et al.
Published: (2024)
by: Schenk, Christina, et al.
Published: (2024)
Nonlinear Model Order Reduction of Dynamical Systems in Process Engineering: Review and Comparison
by: Schulze, Jan C., et al.
Published: (2025)
by: Schulze, Jan C., et al.
Published: (2025)
Symmetric Hermite quadrature-based balanced truncation for learning linear dynamical systems from derivative data
by: Reiter, Sean, et al.
Published: (2026)
by: Reiter, Sean, et al.
Published: (2026)
A Riemannian Optimization Perspective of the Gauss-Newton Method for Feedforward Neural Networks
by: Cayci, Semih
Published: (2024)
by: Cayci, Semih
Published: (2024)
Stability of Primal-Dual Gradient Flow Dynamics for Multi-Block Convex Optimization Problems
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
by: Ozaslan, Ibrahim K., et al.
Published: (2024)
Optimal Control Operator Perspective and a Neural Adaptive Spectral Method
by: Feng, Mingquan, et al.
Published: (2024)
by: Feng, Mingquan, et al.
Published: (2024)
Similar Items
-
Gain Scheduling with a Neural Operator for a Transport PDE with Nonlinear Recirculation
by: Lamarque, Maxence, et al.
Published: (2024) -
Distributed Optimization via Energy Conservation Laws in Dilated Coordinates
by: Baranwal, Mayank, et al.
Published: (2024) -
Adaptive Neural-Operator Backstepping Control of a Benchmark Hyperbolic PDE
by: Lamarque, Maxence, et al.
Published: (2024) -
Analysis of Thompson Sampling for Controlling Unknown Linear Diffusion Processes
by: Faradonbeh, Mohamad Kazem Shirani, et al.
Published: (2022) -
Universal Approximation Power of Deep Residual Neural Networks via Nonlinear Control Theory
by: Tabuada, Paulo, et al.
Published: (2020)