Wave-Attractor-Tree: A Hierarchical Binary Tree Reduction Architecture for Efficient Sequence Modeling
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Berezkin, Igor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Generalization Bound for a Family of Implicit Networks
von: Fung, Samy Wu, et al.
Veröffentlicht: (2024)
von: Fung, Samy Wu, et al.
Veröffentlicht: (2024)
Deep Learning and Transfer Learning Architectures for English Premier League Player Performance Forecasting
von: Frees, Daniel, et al.
Veröffentlicht: (2024)
von: Frees, Daniel, et al.
Veröffentlicht: (2024)
Exact Sequence Interpolation with Transformers
von: Alcalde, Albert, et al.
Veröffentlicht: (2025)
von: Alcalde, Albert, et al.
Veröffentlicht: (2025)
The Inhibitor: ReLU and Addition-Based Attention for Efficient Transformers under Fully Homomorphic Encryption on the Torus
von: Brännvall, Rickard, et al.
Veröffentlicht: (2023)
von: Brännvall, Rickard, et al.
Veröffentlicht: (2023)
Attention to Mamba: A Recipe for Cross-Architecture Distillation
von: Moudgil, Abhinav, et al.
Veröffentlicht: (2026)
von: Moudgil, Abhinav, et al.
Veröffentlicht: (2026)
Improving Fairness and Mitigating MADness in Generative Models
von: Mayer, Paul, et al.
Veröffentlicht: (2024)
von: Mayer, Paul, et al.
Veröffentlicht: (2024)
Tadpole: Autoencoders as Foundation Models for 3D PDEs with Online Learning
von: Liu, Qiang, et al.
Veröffentlicht: (2026)
von: Liu, Qiang, et al.
Veröffentlicht: (2026)
Two-Stage Generative Model for Intracranial Aneurysm Meshes with Morphological Marker Conditioning
von: Ding, Wenhao, et al.
Veröffentlicht: (2025)
von: Ding, Wenhao, et al.
Veröffentlicht: (2025)
Influence-Inspired Spectral Rotations for Extreme Low-Bit LLM Quantization
von: Pavlov, Gorgi
Veröffentlicht: (2026)
von: Pavlov, Gorgi
Veröffentlicht: (2026)
From Features to Graphs: Exploring Graph Structures and Pairwise Interactions via GNNs
von: Yamchote, Phaphontee, et al.
Veröffentlicht: (2025)
von: Yamchote, Phaphontee, et al.
Veröffentlicht: (2025)
Beyond the Class Subspace: Teacher-Guided Training for Reliable Out-of-Distribution Detection in Single-Domain Models
von: Yang, Hong, et al.
Veröffentlicht: (2026)
von: Yang, Hong, et al.
Veröffentlicht: (2026)
A Teacher-Student Perspective on the Dynamics of Learning Near the Optimal Point
von: Couto, Carlos, et al.
Veröffentlicht: (2025)
von: Couto, Carlos, et al.
Veröffentlicht: (2025)
A Comprehensive View of Personalized Federated Learning on Heterogeneous Clinical Datasets
von: Tavakoli, Fatemeh, et al.
Veröffentlicht: (2023)
von: Tavakoli, Fatemeh, et al.
Veröffentlicht: (2023)
The Boundaries of Verifiable Accuracy, Robustness, and Generalisation in Deep Learning
von: Bastounis, Alexander, et al.
Veröffentlicht: (2023)
von: Bastounis, Alexander, et al.
Veröffentlicht: (2023)
Flow matching on homogeneous spaces
von: Ruscelli, Francesco
Veröffentlicht: (2026)
von: Ruscelli, Francesco
Veröffentlicht: (2026)
Training-Free Generative Sampling via Moment-Matched Score Smoothing
von: Yao, Zhenyu, et al.
Veröffentlicht: (2026)
von: Yao, Zhenyu, et al.
Veröffentlicht: (2026)
Generative Design of Ship Propellers using Conditional Flow Matching
von: Kruger, Patrick, et al.
Veröffentlicht: (2026)
von: Kruger, Patrick, et al.
Veröffentlicht: (2026)
TabPFN for Zero-shot Parametric Engineering Design Generation
von: Wang, Ke, et al.
Veröffentlicht: (2026)
von: Wang, Ke, et al.
Veröffentlicht: (2026)
Balanced LoRA: Removing Parameter Invariance to Accelerate Convergence
von: Castin, Valérie, et al.
Veröffentlicht: (2026)
von: Castin, Valérie, et al.
Veröffentlicht: (2026)
Adynamical systems view of training generativemodels and the memorization phenomenon
von: Athreya, Siva, et al.
Veröffentlicht: (2026)
von: Athreya, Siva, et al.
Veröffentlicht: (2026)
Graph-Conditional Flow Matching for Relational Data Generation
von: Scassola, Davide, et al.
Veröffentlicht: (2025)
von: Scassola, Davide, et al.
Veröffentlicht: (2025)
Feature Learning Beyond the Edge of Stability
von: Terjék, Dávid
Veröffentlicht: (2025)
von: Terjék, Dávid
Veröffentlicht: (2025)
ConFIG: Towards Conflict-free Training of Physics Informed Neural Networks
von: Liu, Qiang, et al.
Veröffentlicht: (2024)
von: Liu, Qiang, et al.
Veröffentlicht: (2024)
Optimizing Basis Function Selection in Constructive Wavelet Neural Networks and Its Applications
von: Huang, Dunsheng, et al.
Veröffentlicht: (2025)
von: Huang, Dunsheng, et al.
Veröffentlicht: (2025)
Closed-Form Feedback-Free Learning with Forward Projection
von: O'Shea, Robert, et al.
Veröffentlicht: (2025)
von: O'Shea, Robert, et al.
Veröffentlicht: (2025)
Markov Chain Estimation with In-Context Learning
von: Lepage, Simon, et al.
Veröffentlicht: (2025)
von: Lepage, Simon, et al.
Veröffentlicht: (2025)
Structured Knowledge Accumulation: The Principle of Entropic Least Action in Forward-Only Neural Learning
von: Quantiota, Bouarfa Mahi
Veröffentlicht: (2025)
von: Quantiota, Bouarfa Mahi
Veröffentlicht: (2025)
Cooperative Multi-Agent Deep Reinforcement Learning in Content Ranking Optimization
von: Qin, Zhou, et al.
Veröffentlicht: (2024)
von: Qin, Zhou, et al.
Veröffentlicht: (2024)
Stability Analysis of Equivariant Convolutional Representations Through The Lens of Equivariant Multi-layered CKNs
von: Chowdhury, Soutrik Roy
Veröffentlicht: (2024)
von: Chowdhury, Soutrik Roy
Veröffentlicht: (2024)
Iterative Orthogonalization Scaling Laws
von: Selvaraj, Devan
Veröffentlicht: (2025)
von: Selvaraj, Devan
Veröffentlicht: (2025)
MLPs at the EOC: Spectrum of the NTK
von: Terjék, Dávid, et al.
Veröffentlicht: (2025)
von: Terjék, Dávid, et al.
Veröffentlicht: (2025)
Segmentation of cracks in 3d images of fiber reinforced concrete using deep learning
von: Nowacka, Anna, et al.
Veröffentlicht: (2025)
von: Nowacka, Anna, et al.
Veröffentlicht: (2025)
The Feature Speed Formula: a flexible approach to scale hyper-parameters of deep neural networks
von: Chizat, Lénaïc, et al.
Veröffentlicht: (2023)
von: Chizat, Lénaïc, et al.
Veröffentlicht: (2023)
Overcoming Oversmoothness in Graph Convolutional Networks via Hybrid Scattering Networks
von: Wenkel, Frederik, et al.
Veröffentlicht: (2022)
von: Wenkel, Frederik, et al.
Veröffentlicht: (2022)
Deep generative models as the probability transformation functions
von: Bondar, Vitalii, et al.
Veröffentlicht: (2025)
von: Bondar, Vitalii, et al.
Veröffentlicht: (2025)
MLPs at the EOC: Concentration of the NTK
von: Terjék, Dávid, et al.
Veröffentlicht: (2025)
von: Terjék, Dávid, et al.
Veröffentlicht: (2025)
BP(λ): Online Learning via Synthetic Gradients
von: Pemberton, Joseph, et al.
Veröffentlicht: (2024)
von: Pemberton, Joseph, et al.
Veröffentlicht: (2024)
Mamba for Scalable and Efficient Personalized Recommendations
von: Starnes, Andrew, et al.
Veröffentlicht: (2024)
von: Starnes, Andrew, et al.
Veröffentlicht: (2024)
Multi-modal Transfer Learning between Biological Foundation Models
von: Garau-Luis, Juan Jose, et al.
Veröffentlicht: (2024)
von: Garau-Luis, Juan Jose, et al.
Veröffentlicht: (2024)
MutaPLM: Protein Language Modeling for Mutation Explanation and Engineering
von: Luo, Yizhen, et al.
Veröffentlicht: (2024)
von: Luo, Yizhen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Generalization Bound for a Family of Implicit Networks
von: Fung, Samy Wu, et al.
Veröffentlicht: (2024) -
Deep Learning and Transfer Learning Architectures for English Premier League Player Performance Forecasting
von: Frees, Daniel, et al.
Veröffentlicht: (2024) -
Exact Sequence Interpolation with Transformers
von: Alcalde, Albert, et al.
Veröffentlicht: (2025) -
The Inhibitor: ReLU and Addition-Based Attention for Efficient Transformers under Fully Homomorphic Encryption on the Torus
von: Brännvall, Rickard, et al.
Veröffentlicht: (2023) -
Attention to Mamba: A Recipe for Cross-Architecture Distillation
von: Moudgil, Abhinav, et al.
Veröffentlicht: (2026)