Saved in:
| Main Author: | Ji, Zhongping |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.07405 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IsoQuant: Hardware-Aligned SO(4) Isoclinic Rotations for LLM KV Cache Compression
by: Ji, Zhongping
Published: (2026)
by: Ji, Zhongping
Published: (2026)
CliffordNet: All You Need is Geometric Algebra
by: Ji, Zhongping
Published: (2026)
by: Ji, Zhongping
Published: (2026)
SFi-Former: Sparse Flow Induced Attention for Graph Transformer
by: Li, Zhonghao, et al.
Published: (2025)
by: Li, Zhonghao, et al.
Published: (2025)
EulerFormer: Sequential User Behavior Modeling with Complex Vector Attention
by: Tian, Zhen, et al.
Published: (2024)
by: Tian, Zhen, et al.
Published: (2024)
CacheFormer: High Attention-Based Segment Caching
by: Singh, Sushant, et al.
Published: (2025)
by: Singh, Sushant, et al.
Published: (2025)
NoiseFormer -- Noise Diffused Symmetric Attention Transformer
by: Kumar, Phani, et al.
Published: (2026)
by: Kumar, Phani, et al.
Published: (2026)
Online Decision MetaMorphFormer: A Casual Transformer-Based Reinforcement Learning Framework of Universal Embodied Intelligence
by: Ji, Luo, et al.
Published: (2024)
by: Ji, Luo, et al.
Published: (2024)
BasisFormer: Attention-based Time Series Forecasting with Learnable and Interpretable Basis
by: Ni, Zelin, et al.
Published: (2023)
by: Ni, Zelin, et al.
Published: (2023)
Hyperbolic Gaussian Blurring Mean Shift: A Statistical Mode-Seeking Framework for Clustering in Curved Spaces
by: Pratihar, Arghya, et al.
Published: (2025)
by: Pratihar, Arghya, et al.
Published: (2025)
TimeFormer: Transformer with Attention Modulation Empowered by Temporal Characteristics for Time Series Forecasting
by: Liu, Zhipeng, et al.
Published: (2025)
by: Liu, Zhipeng, et al.
Published: (2025)
S$^2$M-Former: Spiking Symmetric Mixing Branchformer for Brain Auditory Attention Detection
by: Wang, Jiaqi, et al.
Published: (2025)
by: Wang, Jiaqi, et al.
Published: (2025)
Rethinking Transformer Connectivity: TLinFormer, A Path to Exact, Full Context-Aware Linear Attention
by: Tang, Zhongpan
Published: (2025)
by: Tang, Zhongpan
Published: (2025)
SelectFormer: Private and Practical Data Selection for Transformers
by: Ouyang, Xu, et al.
Published: (2023)
by: Ouyang, Xu, et al.
Published: (2023)
AnchorFormer: Differentiable Anchor Attention for Efficient Vision Transformer
by: Shan, Jiquan, et al.
Published: (2025)
by: Shan, Jiquan, et al.
Published: (2025)
RecurFormer: Not All Transformer Heads Need Self-Attention
by: Yan, Ruiqing, et al.
Published: (2024)
by: Yan, Ruiqing, et al.
Published: (2024)
TouchFormer: A Robust Transformer-based Framework for Multimodal Material Perception
by: Lyu, Kailin, et al.
Published: (2025)
by: Lyu, Kailin, et al.
Published: (2025)
From TLinFormer to TConstFormer: The Leap to Constant-Time Transformer Attention: Achieving O(1) Computation and O(1) KV Cache during Autoregressive Inference
by: Tang, Zhongpan
Published: (2025)
by: Tang, Zhongpan
Published: (2025)
VecFormer: Towards Efficient and Generalizable Graph Transformer with Graph Token Attention
by: Zhou, Jingbo, et al.
Published: (2026)
by: Zhou, Jingbo, et al.
Published: (2026)
RingFormer: A Neural Vocoder with Ring Attention and Convolution-Augmented Transformer
by: Hong, Seongho, et al.
Published: (2025)
by: Hong, Seongho, et al.
Published: (2025)
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and $\mathcal{O}(T)$ Complexity
by: Zou, Shihao, et al.
Published: (2025)
by: Zou, Shihao, et al.
Published: (2025)
RTA-Former: Reverse Transformer Attention for Polyp Segmentation
by: Li, Zhikai, et al.
Published: (2024)
by: Li, Zhikai, et al.
Published: (2024)
E2Former-V2: On-the-Fly Equivariant Attention with Linear Activation Memory
by: Huang, Lin, et al.
Published: (2026)
by: Huang, Lin, et al.
Published: (2026)
Decision ConvFormer: Local Filtering in MetaFormer is Sufficient for Decision Making
by: Kim, Jeonghye, et al.
Published: (2023)
by: Kim, Jeonghye, et al.
Published: (2023)
EEG-FuseFormer: A Transformer-Driven Feature Fusion Framework for Seizure Onset Prediction
by: Hariharan, Vigneshwar, et al.
Published: (2026)
by: Hariharan, Vigneshwar, et al.
Published: (2026)
RouteFormer: A Transformer-Based Routing Framework for Autonomous Vehicles
by: Youssef, Yazan, et al.
Published: (2025)
by: Youssef, Yazan, et al.
Published: (2025)
AlgoFormer: An Efficient Transformer Framework with Algorithmic Structures
by: Gao, Yihang, et al.
Published: (2024)
by: Gao, Yihang, et al.
Published: (2024)
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding
by: Setyawan, Novendra, et al.
Published: (2024)
by: Setyawan, Novendra, et al.
Published: (2024)
CloudFormer: An Attention-based Performance Prediction for Public Clouds with Unknown Workload
by: Shahbazinia, Amirhossein, et al.
Published: (2025)
by: Shahbazinia, Amirhossein, et al.
Published: (2025)
Riemannian Time Warping: Multiple Sequence Alignment in Curved Spaces
by: Richter, Julian, et al.
Published: (2025)
by: Richter, Julian, et al.
Published: (2025)
CardioPatternFormer: Pattern-Guided Attention for Interpretable ECG Classification with Transformer Architecture
by: Uğraş, Berat Kutay, et al.
Published: (2025)
by: Uğraş, Berat Kutay, et al.
Published: (2025)
ComplexFormer: Disruptively Advancing Transformer Inference Ability via Head-Specific Complex Vector Attention
by: Shao, Jintian, et al.
Published: (2025)
by: Shao, Jintian, et al.
Published: (2025)
MultiResFormer: Transformer with Adaptive Multi-Resolution Modeling for General Time Series Forecasting
by: Du, Linfeng, et al.
Published: (2023)
by: Du, Linfeng, et al.
Published: (2023)
RiemannONets: Interpretable Neural Operators for Riemann Problems
by: Peyvan, Ahmad, et al.
Published: (2024)
by: Peyvan, Ahmad, et al.
Published: (2024)
ZETA: Leveraging Z-order Curves for Efficient Top-k Attention
by: Zeng, Qiuhao, et al.
Published: (2025)
by: Zeng, Qiuhao, et al.
Published: (2025)
HessFormer: Hessians at Foundation Scale
by: Granziol, Diego
Published: (2025)
by: Granziol, Diego
Published: (2025)
How Many Heads Make an SSM? A Unified Framework for Attention and State Space Models
by: Ghodsi, Ali
Published: (2025)
by: Ghodsi, Ali
Published: (2025)
PINNsFormer: A Transformer-Based Framework For Physics-Informed Neural Networks
by: Zhao, Zhiyuan, et al.
Published: (2023)
by: Zhao, Zhiyuan, et al.
Published: (2023)
UniMamba: A Unified Spatial-Temporal Modeling Framework with State-Space and Attention Integration
by: Chen, Xingsheng, et al.
Published: (2026)
by: Chen, Xingsheng, et al.
Published: (2026)
WeatherFormer: Empowering Global Numerical Weather Forecasting with Space-Time Transformer
by: Gong, Junchao, et al.
Published: (2024)
by: Gong, Junchao, et al.
Published: (2024)
Riemann-Lebesgue Forest for Regression
by: Qin, Tian, et al.
Published: (2024)
by: Qin, Tian, et al.
Published: (2024)
Similar Items
-
IsoQuant: Hardware-Aligned SO(4) Isoclinic Rotations for LLM KV Cache Compression
by: Ji, Zhongping
Published: (2026) -
CliffordNet: All You Need is Geometric Algebra
by: Ji, Zhongping
Published: (2026) -
SFi-Former: Sparse Flow Induced Attention for Graph Transformer
by: Li, Zhonghao, et al.
Published: (2025) -
EulerFormer: Sequential User Behavior Modeling with Complex Vector Attention
by: Tian, Zhen, et al.
Published: (2024) -
CacheFormer: High Attention-Based Segment Caching
by: Singh, Sushant, et al.
Published: (2025)