Cluster-wise Graph Transformer with Dual-granularity Kernelized Attention
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Siyuan, Song, Yunchong, Zhou, Jiayue, Lin, Zhouhan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Graph Parsing Networks
by: Song, Yunchong, et al.
Published: (2024)
by: Song, Yunchong, et al.
Published: (2024)
Towards Compressive and Scalable Recurrent Memory
by: Song, Yunchong, et al.
Published: (2026)
by: Song, Yunchong, et al.
Published: (2026)
Controlling Exploration-Exploitation in GFlowNets via Markov Chain Perspectives
by: Chen, Lin, et al.
Published: (2026)
by: Chen, Lin, et al.
Published: (2026)
Simplifying Graph Kernels for Efficient
by: Wang, Lin, et al.
Published: (2025)
by: Wang, Lin, et al.
Published: (2025)
VecFormer: Towards Efficient and Generalizable Graph Transformer with Graph Token Attention
by: Zhou, Jingbo, et al.
Published: (2026)
by: Zhou, Jingbo, et al.
Published: (2026)
Projection-Free Transformers via Gaussian Kernel Attention
by: Kundu, Debarshi, et al.
Published: (2026)
by: Kundu, Debarshi, et al.
Published: (2026)
Multi-Relation Graph-Kernel Strengthen Network for Graph-Level Clustering
by: Han, Renda, et al.
Published: (2025)
by: Han, Renda, et al.
Published: (2025)
Towards Federated Clustering: A Client-wise Private Graph Aggregation Framework
by: He, Guanxiong, et al.
Published: (2025)
by: He, Guanxiong, et al.
Published: (2025)
Attention Beyond Neighborhoods: Reviving Transformer for Graph Clustering
by: Xie, Xuanting, et al.
Published: (2025)
by: Xie, Xuanting, et al.
Published: (2025)
Kernelized Edge Attention: Addressing Semantic Attention Blurring in Temporal Graph Neural Networks
by: Waghmare, Govind, et al.
Published: (2026)
by: Waghmare, Govind, et al.
Published: (2026)
Dual-Kernel Graph Community Contrastive Learning
by: Chen, Xiang, et al.
Published: (2025)
by: Chen, Xiang, et al.
Published: (2025)
PolyFormer: Scalable Node-wise Filters via Polynomial Graph Transformer
by: Ma, Jiahong, et al.
Published: (2024)
by: Ma, Jiahong, et al.
Published: (2024)
AnchorGT: Efficient and Flexible Attention Architecture for Scalable Graph Transformers
by: Zhu, Wenhao, et al.
Published: (2024)
by: Zhu, Wenhao, et al.
Published: (2024)
CuBridge: An LLM-Based Framework for Understanding and Reconstructing High-Performance Attention Kernels
by: Ma, Xing, et al.
Published: (2026)
by: Ma, Xing, et al.
Published: (2026)
DAM-GT: Dual Positional Encoding-Based Attention Masking Graph Transformer for Node Classification
by: Li, Chenyang, et al.
Published: (2025)
by: Li, Chenyang, et al.
Published: (2025)
Mitigating Premature Discretization with Progressive Quantization for Robust Vector Tokenization
by: Zhao, Wenhao, et al.
Published: (2026)
by: Zhao, Wenhao, et al.
Published: (2026)
Solving Attention Kernel Regression Problem via Pre-conditioner
by: Song, Zhao, et al.
Published: (2023)
by: Song, Zhao, et al.
Published: (2023)
Critical Data Size of Language Models from a Grokking Perspective
by: Zhu, Xuekai, et al.
Published: (2024)
by: Zhu, Xuekai, et al.
Published: (2024)
Dual-Center Graph Clustering with Neighbor Distribution
by: Cheng, Enhao, et al.
Published: (2025)
by: Cheng, Enhao, et al.
Published: (2025)
Graph External Attention Enhanced Transformer
by: Liang, Jianqing, et al.
Published: (2024)
by: Liang, Jianqing, et al.
Published: (2024)
Cluster Attention for Graph Machine Learning
by: Platonov, Oleg, et al.
Published: (2026)
by: Platonov, Oleg, et al.
Published: (2026)
Transformer Based Linear Attention with Optimized GPU Kernel Implementation
by: Gerami, Armin, et al.
Published: (2025)
by: Gerami, Armin, et al.
Published: (2025)
Disentangling Recall and Reasoning in Transformer Models through Layer-wise Attention and Activation Analysis
by: Fartale, Harshwardhan, et al.
Published: (2025)
by: Fartale, Harshwardhan, et al.
Published: (2025)
Multiple Kernel Clustering via Local Regression Integration
by: Du, Liang, et al.
Published: (2024)
by: Du, Liang, et al.
Published: (2024)
Graph-based Clustering Revisited: A Relaxation of Kernel $k$-Means Perspective
by: Lyu, Wenlong, et al.
Published: (2025)
by: Lyu, Wenlong, et al.
Published: (2025)
Towards Subgraph Isomorphism Counting with Graph Kernels
by: Liu, Xin, et al.
Published: (2024)
by: Liu, Xin, et al.
Published: (2024)
Dual-Optimized Adaptive Graph Reconstruction for Multi-View Graph Clustering
by: Wen, Zichen, et al.
Published: (2024)
by: Wen, Zichen, et al.
Published: (2024)
Cross-Contrastive Clustering for Multimodal Attributed Graphs with Dual Graph Filtering
by: Zheng, Haoran, et al.
Published: (2025)
by: Zheng, Haoran, et al.
Published: (2025)
Dual Boost-Driven Graph-Level Clustering Network
by: Smith, John, et al.
Published: (2025)
by: Smith, John, et al.
Published: (2025)
Generative Kernel Spectral Clustering
by: Winant, David, et al.
Published: (2025)
by: Winant, David, et al.
Published: (2025)
Generating Directed Graphs with Dual Attention and Asymmetric Encoding
by: Carballo-Castro, Alba, et al.
Published: (2025)
by: Carballo-Castro, Alba, et al.
Published: (2025)
Stronger Graph Transformer with Regularized Attention Scores
by: Ku, Eugene
Published: (2023)
by: Ku, Eugene
Published: (2023)
LI-DSN: A Layer-wise Interactive Dual-Stream Network for EEG Decoding
by: Yue, Chenghao, et al.
Published: (2026)
by: Yue, Chenghao, et al.
Published: (2026)
Curse of Attention: A Kernel-Based Perspective for Why Transformers Fail to Generalize on Time Series Forecasting and Beyond
by: Ke, Yekun, et al.
Published: (2024)
by: Ke, Yekun, et al.
Published: (2024)
PACT: Peak-Aware Cross-Attention Graph Transformers for Efficient Storm-Surge Emulation
by: Liu, Zesheng, et al.
Published: (2026)
by: Liu, Zesheng, et al.
Published: (2026)
Element-wise Attention Is All You Need
by: Feng, Guoxin
Published: (2025)
by: Feng, Guoxin
Published: (2025)
DGTN: Graph-Enhanced Transformer with Diffusive Attention Gating Mechanism for Enzyme DDG Prediction
by: Lin, Abigail
Published: (2025)
by: Lin, Abigail
Published: (2025)
On-the-Fly Adaptive Distillation of Transformer to Dual-State Linear Attention
by: Ro, Yeonju, et al.
Published: (2025)
by: Ro, Yeonju, et al.
Published: (2025)
B-TGAT: A Bi-directional Temporal Graph Attention Transformer for Clustering Multivariate Spatiotemporal Data
by: Nji, Francis Ndikum, et al.
Published: (2025)
by: Nji, Francis Ndikum, et al.
Published: (2025)
CAST: Clustering Self-Attention using Surrogate Tokens for Efficient Transformers
by: van Engelenhoven, Adjorn, et al.
Published: (2024)
by: van Engelenhoven, Adjorn, et al.
Published: (2024)
Similar Items
-
Graph Parsing Networks
by: Song, Yunchong, et al.
Published: (2024) -
Towards Compressive and Scalable Recurrent Memory
by: Song, Yunchong, et al.
Published: (2026) -
Controlling Exploration-Exploitation in GFlowNets via Markov Chain Perspectives
by: Chen, Lin, et al.
Published: (2026) -
Simplifying Graph Kernels for Efficient
by: Wang, Lin, et al.
Published: (2025) -
VecFormer: Towards Efficient and Generalizable Graph Transformer with Graph Token Attention
by: Zhou, Jingbo, et al.
Published: (2026)