Attention mechanisms in neural networks
Fuente:
arXiv
Saved in:
| Main Author: | Hays, Hasi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Selective Synchronization Attention
by: Hays, Hasi
Published: (2026)
by: Hays, Hasi
Published: (2026)
Resonant Sparse Geometry Networks
by: Hays, Hasi
Published: (2026)
by: Hays, Hasi
Published: (2026)
RAG-GNN: Integrating Retrieved Knowledge with Graph Neural Networks for Precision Medicine
by: Hays, Hasi, et al.
Published: (2026)
by: Hays, Hasi, et al.
Published: (2026)
Hebbian-Oscillatory Co-Learning
by: Hays, Hasi
Published: (2026)
by: Hays, Hasi
Published: (2026)
On permutation-invariant neural networks
by: Kimura, Masanari, et al.
Published: (2024)
by: Kimura, Masanari, et al.
Published: (2024)
Sobolev acceleration for neural networks
by: Oh, Jong Kwon, et al.
Published: (2025)
by: Oh, Jong Kwon, et al.
Published: (2025)
Principles of Lipschitz continuity in neural networks
by: Luo, Róisín
Published: (2026)
by: Luo, Róisín
Published: (2026)
Linearity-based neural network compression
by: Dobler, Silas, et al.
Published: (2025)
by: Dobler, Silas, et al.
Published: (2025)
Applying graph neural network to SupplyGraph for supply chain network
by: Han, Kihwan
Published: (2024)
by: Han, Kihwan
Published: (2024)
Understanding the dynamics of the frequency bias in neural networks
by: Molina, Juan, et al.
Published: (2024)
by: Molina, Juan, et al.
Published: (2024)
Graph neural networks informed locally by thermodynamics
by: Tierz, Alicia, et al.
Published: (2024)
by: Tierz, Alicia, et al.
Published: (2024)
Graph neural networks and non-commuting operators
by: Velasco, Mauricio, et al.
Published: (2024)
by: Velasco, Mauricio, et al.
Published: (2024)
LayerCollapse: Adaptive compression of neural networks
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2023)
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2023)
Kinematic analysis of structural mechanics based on convolutional neural network
by: Zhang, Leye, et al.
Published: (2024)
by: Zhang, Leye, et al.
Published: (2024)
Randomness and signal propagation in physics-informed neural networks (PINNs): A neural PDE perspective
by: Tucny, Jean-Michel, et al.
Published: (2025)
by: Tucny, Jean-Michel, et al.
Published: (2025)
Elimination-compensation pruning for fully-connected neural networks
by: Ballini, Enrico, et al.
Published: (2026)
by: Ballini, Enrico, et al.
Published: (2026)
Concealed Adversarial attacks on neural networks for sequential data
by: Sokerin, Petr, et al.
Published: (2025)
by: Sokerin, Petr, et al.
Published: (2025)
Planning in a recurrent neural network that plays Sokoban
by: Taufeeque, Mohammad, et al.
Published: (2024)
by: Taufeeque, Mohammad, et al.
Published: (2024)
Dopamine-driven synaptic credit assignment in neural networks
by: Nambusubramaniyan, Saranraj, et al.
Published: (2025)
by: Nambusubramaniyan, Saranraj, et al.
Published: (2025)
Understanding polysemanticity in neural networks through coding theory
by: Marshall, Simon C., et al.
Published: (2024)
by: Marshall, Simon C., et al.
Published: (2024)
Variational autoencoder-based neural network model compression
by: Cheng, Liang, et al.
Published: (2024)
by: Cheng, Liang, et al.
Published: (2024)
Conditional computation in neural networks: principles and research trends
by: Scardapane, Simone, et al.
Published: (2024)
by: Scardapane, Simone, et al.
Published: (2024)
Deep neural networks have an inbuilt Occam's razor
by: Mingard, Chris, et al.
Published: (2023)
by: Mingard, Chris, et al.
Published: (2023)
Utilizing Lyapunov Exponents in designing deep neural networks
by: Mittra, Tirthankar
Published: (2024)
by: Mittra, Tirthankar
Published: (2024)
Grey-informed neural network for time-series forecasting
by: Xie, Wanli, et al.
Published: (2024)
by: Xie, Wanli, et al.
Published: (2024)
A comparative analysis of a neural network with calculated weights and a neural network with random generation of weights based on the training dataset size
by: Geidarov, Polad
Published: (2025)
by: Geidarov, Polad
Published: (2025)
Tucker Attention: A generalization of approximate attention mechanisms
by: Klein, Timon, et al.
Published: (2026)
by: Klein, Timon, et al.
Published: (2026)
Understanding the learned look-ahead behavior of chess neural networks
by: Cruz, Diogo
Published: (2025)
by: Cruz, Diogo
Published: (2025)
Confidence-gated training for efficient early-exit neural networks
by: Mokssit, Saad, et al.
Published: (2025)
by: Mokssit, Saad, et al.
Published: (2025)
Adaptive multiple optimal learning factors for neural network training
by: Challagundla, Jeshwanth
Published: (2024)
by: Challagundla, Jeshwanth
Published: (2024)
Stochastic stem bucking using mixture density neural networks
by: Schmiedel, Simon
Published: (2024)
by: Schmiedel, Simon
Published: (2024)
StaQ it! Growing neural networks for Policy Mirror Descent
by: Shilova, Alena, et al.
Published: (2025)
by: Shilova, Alena, et al.
Published: (2025)
Topological derivative approach for deep neural network architecture adaptation
by: Krishnanunni, C G, et al.
Published: (2025)
by: Krishnanunni, C G, et al.
Published: (2025)
Do graph neural network states contain graph properties?
by: Pelletreau-Duris, Tom, et al.
Published: (2024)
by: Pelletreau-Duris, Tom, et al.
Published: (2024)
Robustness questions the interpretability of graph neural networks: what to do?
by: Lukyanov, Kirill, et al.
Published: (2025)
by: Lukyanov, Kirill, et al.
Published: (2025)
Do deep neural networks utilize the weight space efficiently?
by: Koyun, Onur Can, et al.
Published: (2024)
by: Koyun, Onur Can, et al.
Published: (2024)
sHGCN: Simplified hyperbolic graph convolutional neural networks
by: Arévalo, Pol, et al.
Published: (2025)
by: Arévalo, Pol, et al.
Published: (2025)
Simmering: Sufficient is better than optimal for training neural networks
by: Babayan, Irina, et al.
Published: (2024)
by: Babayan, Irina, et al.
Published: (2024)
Multi-layer random features and the approximation power of neural networks
by: Takhanov, Rustem
Published: (2024)
by: Takhanov, Rustem
Published: (2024)
Addressing divergent representations from causal interventions on neural networks
by: Grant, Satchel, et al.
Published: (2025)
by: Grant, Satchel, et al.
Published: (2025)
Similar Items
-
Selective Synchronization Attention
by: Hays, Hasi
Published: (2026) -
Resonant Sparse Geometry Networks
by: Hays, Hasi
Published: (2026) -
RAG-GNN: Integrating Retrieved Knowledge with Graph Neural Networks for Precision Medicine
by: Hays, Hasi, et al.
Published: (2026) -
Hebbian-Oscillatory Co-Learning
by: Hays, Hasi
Published: (2026) -
On permutation-invariant neural networks
by: Kimura, Masanari, et al.
Published: (2024)