SP$^2$T: Sparse Proxy Attention for Dual-stream Point Transformer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wan, Jiaxu, Zhang, Hong, He, Ziqi, Deng, Yangyan, Wang, Qishu, Yuan, Ding, Yang, Yifan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LiteAttention: A Temporal Sparse Attention for Diffusion Transformers
von: Shmilovich, Dor, et al.
Veröffentlicht: (2025)
von: Shmilovich, Dor, et al.
Veröffentlicht: (2025)
MoVE: Mixture of Value Embeddings -- A New Axis for Scaling Parametric Memory in Autoregressive Models
von: Li, Yangyan
Veröffentlicht: (2026)
von: Li, Yangyan
Veröffentlicht: (2026)
Dual-stream Transformer-GCN Model with Contextualized Representations Learning for Monocular 3D Human Pose Estimation
von: Ye, Mingrui, et al.
Veröffentlicht: (2025)
von: Ye, Mingrui, et al.
Veröffentlicht: (2025)
PSReg: Prior-guided Sparse Mixture of Experts for Point Cloud Registration
von: Huang, Xiaoshui, et al.
Veröffentlicht: (2025)
von: Huang, Xiaoshui, et al.
Veröffentlicht: (2025)
ShapeShifter: 3D Variations Using Multiscale and Sparse Point-Voxel Diffusion
von: Maruani, Nissim, et al.
Veröffentlicht: (2025)
von: Maruani, Nissim, et al.
Veröffentlicht: (2025)
Symbolic Rule Extraction from Attention-Guided Sparse Representations in Vision Transformers
von: Padalkar, Parth, et al.
Veröffentlicht: (2025)
von: Padalkar, Parth, et al.
Veröffentlicht: (2025)
EMRA-proxy: Enhancing Multi-Class Region Semantic Segmentation in Remote Sensing Images with Attention Proxy
von: Yu, Yichun, et al.
Veröffentlicht: (2025)
von: Yu, Yichun, et al.
Veröffentlicht: (2025)
DualTeacher: Bridging Coexistence of Unlabelled Classes for Semi-supervised Incremental Object Detection
von: Yuan, Ziqi, et al.
Veröffentlicht: (2023)
von: Yuan, Ziqi, et al.
Veröffentlicht: (2023)
SuperiorGAT: Graph Attention Networks for Sparse LiDAR Point Cloud Reconstruction in Autonomous Systems
von: Awedat, Khalfalla, et al.
Veröffentlicht: (2025)
von: Awedat, Khalfalla, et al.
Veröffentlicht: (2025)
HIMOSA: Efficient Remote Sensing Image Super-Resolution with Hierarchical Mixture of Sparse Attention
von: Liu, Yi, et al.
Veröffentlicht: (2025)
von: Liu, Yi, et al.
Veröffentlicht: (2025)
SwinECAT: A Transformer-based fundus disease classification model with Shifted Window Attention and Efficient Channel Attention
von: Gu, Peiran, et al.
Veröffentlicht: (2025)
von: Gu, Peiran, et al.
Veröffentlicht: (2025)
Sparse-vDiT: Unleashing the Power of Sparse Attention to Accelerate Video Diffusion Transformers
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
Hilbert-Guided Sparse Local Attention
von: Li, Yunge, et al.
Veröffentlicht: (2025)
von: Li, Yunge, et al.
Veröffentlicht: (2025)
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
von: Li, Wenxi, et al.
Veröffentlicht: (2025)
von: Li, Wenxi, et al.
Veröffentlicht: (2025)
Dual-Channel Attention Guidance for Training-Free Image Editing Control in Diffusion Transformers
von: Li, Guandong
Veröffentlicht: (2026)
von: Li, Guandong
Veröffentlicht: (2026)
RegionPLC: Regional Point-Language Contrastive Learning for Open-World 3D Scene Understanding
von: Yang, Jihan, et al.
Veröffentlicht: (2023)
von: Yang, Jihan, et al.
Veröffentlicht: (2023)
Spiking Vision Transformer with Saccadic Attention
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
Unlocking Attributes' Contribution to Successful Camouflage: A Combined Textual and VisualAnalysis Strategy
von: Zhang, Hong, et al.
Veröffentlicht: (2024)
von: Zhang, Hong, et al.
Veröffentlicht: (2024)
Local-Global Attention: An Adaptive Mechanism for Multi-Scale Feature Integration
von: Shao, Yifan
Veröffentlicht: (2024)
von: Shao, Yifan
Veröffentlicht: (2024)
PPTArena: A Benchmark for Agentic PowerPoint Editing
von: Ofengenden, Michael, et al.
Veröffentlicht: (2025)
von: Ofengenden, Michael, et al.
Veröffentlicht: (2025)
DGSAN: Dual-Graph Spatiotemporal Attention Network for Pulmonary Nodule Malignancy Prediction
von: Yu, Xiao, et al.
Veröffentlicht: (2025)
von: Yu, Xiao, et al.
Veröffentlicht: (2025)
Exploiting Point-Language Models with Dual-Prompts for 3D Anomaly Detection
von: Wang, Jiaxiang, et al.
Veröffentlicht: (2025)
von: Wang, Jiaxiang, et al.
Veröffentlicht: (2025)
DSXFormer: Dual-Pooling Spectral Squeeze-Expansion and Dynamic Context Attention Transformer for Hyperspectral Image Classification
von: Ullah, Farhan, et al.
Veröffentlicht: (2026)
von: Ullah, Farhan, et al.
Veröffentlicht: (2026)
SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable Sparse-Linear Attention
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
von: Zhang, Jintao, et al.
Veröffentlicht: (2025)
Geometric Point Attention Transformer for 3D Shape Reassembly
von: Li, Jiahan, et al.
Veröffentlicht: (2024)
von: Li, Jiahan, et al.
Veröffentlicht: (2024)
Online Handwritten Signature Verification Based on Temporal-Spatial Graph Attention Transformer
von: Yuan, Hai-jie, et al.
Veröffentlicht: (2025)
von: Yuan, Hai-jie, et al.
Veröffentlicht: (2025)
VidBridge-R1: Bridging QA and Captioning for RL-based Video Understanding Models with Intermediate Proxy Tasks
von: Chen, Xinlong, et al.
Veröffentlicht: (2025)
von: Chen, Xinlong, et al.
Veröffentlicht: (2025)
Region-Transformer: Self-Attention Region Based Class-Agnostic Point Cloud Segmentation
von: Gyawali, Dipesh, et al.
Veröffentlicht: (2024)
von: Gyawali, Dipesh, et al.
Veröffentlicht: (2024)
ContextDrag: Precise Drag-Based Image Editing via Context-Preserving Token Injection and Position-Aligned Attention
von: He, Huiguo, et al.
Veröffentlicht: (2025)
von: He, Huiguo, et al.
Veröffentlicht: (2025)
GeoT: Geometry-guided Instance-dependent Transition Matrix for Semi-supervised Tooth Point Cloud Segmentation
von: Yu, Weihao, et al.
Veröffentlicht: (2025)
von: Yu, Weihao, et al.
Veröffentlicht: (2025)
Trans${^2}$-CBCT: A Dual-Transformer Framework for Sparse-View CBCT Reconstruction
von: Yang, Minmin, et al.
Veröffentlicht: (2025)
von: Yang, Minmin, et al.
Veröffentlicht: (2025)
D2SP: Dynamic Dual-Stage Purification Framework for Dual Noise Mitigation in Vision-based Affective Recognition
von: Wang, Haoran, et al.
Veröffentlicht: (2024)
von: Wang, Haoran, et al.
Veröffentlicht: (2024)
Make It Efficient: Dynamic Sparse Attention for Autoregressive Image Generation
von: Xiang, Xunzhi, et al.
Veröffentlicht: (2025)
von: Xiang, Xunzhi, et al.
Veröffentlicht: (2025)
SparseSwin: Swin Transformer with Sparse Transformer Block
von: Pinasthika, Krisna, et al.
Veröffentlicht: (2023)
von: Pinasthika, Krisna, et al.
Veröffentlicht: (2023)
Reciprocal Attention Mixing Transformer for Lightweight Image Restoration
von: Choi, Haram, et al.
Veröffentlicht: (2023)
von: Choi, Haram, et al.
Veröffentlicht: (2023)
Coarse-Guided Visual Generation via Weighted h-Transform Sampling
von: Wang, Yanghao, et al.
Veröffentlicht: (2026)
von: Wang, Yanghao, et al.
Veröffentlicht: (2026)
physfusion: A Transformer-based Dual-Stream Radar and Vision Fusion Framework for Open Water Surface Object Detection
von: Wan, Yuting, et al.
Veröffentlicht: (2026)
von: Wan, Yuting, et al.
Veröffentlicht: (2026)
Region-Adaptive Sampling for Diffusion Transformers
von: Liu, Ziming, et al.
Veröffentlicht: (2025)
von: Liu, Ziming, et al.
Veröffentlicht: (2025)
Ride the Wave: Precision-Allocated Sparse Attention for Smooth Video Generation
von: Zhang, Wentai, et al.
Veröffentlicht: (2026)
von: Zhang, Wentai, et al.
Veröffentlicht: (2026)
Attention-Guided Flow-Matching for Sparse 3D Geological Generation
von: Lu, Zhixiang, et al.
Veröffentlicht: (2026)
von: Lu, Zhixiang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LiteAttention: A Temporal Sparse Attention for Diffusion Transformers
von: Shmilovich, Dor, et al.
Veröffentlicht: (2025) -
MoVE: Mixture of Value Embeddings -- A New Axis for Scaling Parametric Memory in Autoregressive Models
von: Li, Yangyan
Veröffentlicht: (2026) -
Dual-stream Transformer-GCN Model with Contextualized Representations Learning for Monocular 3D Human Pose Estimation
von: Ye, Mingrui, et al.
Veröffentlicht: (2025) -
PSReg: Prior-guided Sparse Mixture of Experts for Point Cloud Registration
von: Huang, Xiaoshui, et al.
Veröffentlicht: (2025) -
ShapeShifter: 3D Variations Using Multiscale and Sparse Point-Voxel Diffusion
von: Maruani, Nissim, et al.
Veröffentlicht: (2025)