GTA: A Geometry-Aware Attention Mechanism for Multi-View Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Miyato, Takeru, Jaeger, Bernhard, Welling, Max, Geiger, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Artificial Kuramoto Oscillatory Neurons
by: Miyato, Takeru, et al.
Published: (2024)
by: Miyato, Takeru, et al.
Published: (2024)
Hidden Biases of End-to-End Driving Datasets
by: Zimmerlin, Julian, et al.
Published: (2024)
by: Zimmerlin, Julian, et al.
Published: (2024)
End-to-end Autonomous Driving: Challenges and Frontiers
by: Chen, Li, et al.
Published: (2023)
by: Chen, Li, et al.
Published: (2023)
Erwin: A Tree-based Hierarchical Transformer for Large-scale Physical Systems
by: Zhdanov, Maksim, et al.
Published: (2025)
by: Zhdanov, Maksim, et al.
Published: (2025)
LEAD: Minimizing Learner-Expert Asymmetry in End-to-End Driving
by: Nguyen, Long, et al.
Published: (2025)
by: Nguyen, Long, et al.
Published: (2025)
Kuramoto Orientation Diffusion Models
by: Song, Yue, et al.
Published: (2025)
by: Song, Yue, et al.
Published: (2025)
Understanding Multi-View Transformers
by: Stary, Michal, et al.
Published: (2025)
by: Stary, Michal, et al.
Published: (2025)
GLAM: Geometry-Guided Local Alignment for Multi-View VLP in Mammography
by: Du, Yuexi, et al.
Published: (2025)
by: Du, Yuexi, et al.
Published: (2025)
Binding Dynamics in Rotating Features
by: Löwe, Sindy, et al.
Published: (2024)
by: Löwe, Sindy, et al.
Published: (2024)
Towards Robust Uncertainty-Aware Incomplete Multi-View Classification
by: Chen, Mulin, et al.
Published: (2024)
by: Chen, Mulin, et al.
Published: (2024)
Geo-NVS-w: Geometry-Aware Novel View Synthesis In-the-Wild with an SDF Renderer
by: Tsalakopoulos, Anastasios, et al.
Published: (2026)
by: Tsalakopoulos, Anastasios, et al.
Published: (2026)
(Sparse) Attention to the Details: Preserving Spectral Fidelity in ML-based Weather Forecasting Models
by: Zhdanov, Maksim, et al.
Published: (2026)
by: Zhdanov, Maksim, et al.
Published: (2026)
Benchmarking Feature Upsampling Methods for Vision Foundation Models using Interactive Segmentation
by: Havrylov, Volodymyr, et al.
Published: (2025)
by: Havrylov, Volodymyr, et al.
Published: (2025)
Blending 3D Geometry and Machine Learning for Multi-View Stereopsis
by: Vats, Vibhas, et al.
Published: (2025)
by: Vats, Vibhas, et al.
Published: (2025)
SLEDGE: Synthesizing Driving Environments with Generative Models and Rule-Based Traffic
by: Chitta, Kashyap, et al.
Published: (2024)
by: Chitta, Kashyap, et al.
Published: (2024)
Robust Multi-View Learning via Representation Fusion of Sample-Level Attention and Alignment of Simulated Perturbation
by: Xu, Jie, et al.
Published: (2025)
by: Xu, Jie, et al.
Published: (2025)
KANGURA: Kolmogorov-Arnold Network-Based Geometry-Aware Learning with Unified Representation Attention for 3D Modeling of Complex Structures
by: Shafie, Mohammad Reza, et al.
Published: (2025)
by: Shafie, Mohammad Reza, et al.
Published: (2025)
Missing Data as Augmentation in the Earth Observation Domain: A Multi-View Learning Approach
by: Mena, Francisco, et al.
Published: (2025)
by: Mena, Francisco, et al.
Published: (2025)
BSA: Ball Sparse Attention for Large-scale Geometries
by: Brita, Catalin E., et al.
Published: (2025)
by: Brita, Catalin E., et al.
Published: (2025)
SGAT4PASS: Spherical Geometry-Aware Transformer for PAnoramic Semantic Segmentation
by: Li, Xuewei, et al.
Published: (2023)
by: Li, Xuewei, et al.
Published: (2023)
MFAF: An EVA02-Based Multi-scale Frequency Attention Fusion Method for Cross-View Geo-Localization
by: Liu, YiTong, et al.
Published: (2025)
by: Liu, YiTong, et al.
Published: (2025)
T-TAME: Trainable Attention Mechanism for Explaining Convolutional Networks and Vision Transformers
by: Ntrougkas, Mariano V., et al.
Published: (2024)
by: Ntrougkas, Mariano V., et al.
Published: (2024)
Streaming 4D Visual Geometry Transformer
by: Zhuo, Dong, et al.
Published: (2025)
by: Zhuo, Dong, et al.
Published: (2025)
AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
by: Achtibat, Reduan, et al.
Published: (2024)
by: Achtibat, Reduan, et al.
Published: (2024)
VG3T: Visual Geometry Grounded Gaussian Transformer
by: Kim, Junho, et al.
Published: (2025)
by: Kim, Junho, et al.
Published: (2025)
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics
by: Colagrande, Alex, et al.
Published: (2025)
by: Colagrande, Alex, et al.
Published: (2025)
Class-Discriminative Attention Maps for Vision Transformers
by: Brocki, Lennart, et al.
Published: (2023)
by: Brocki, Lennart, et al.
Published: (2023)
Scratching Visual Transformer's Back with Uniform Attention
by: Hyeon-Woo, Nam, et al.
Published: (2022)
by: Hyeon-Woo, Nam, et al.
Published: (2022)
Spontaneous symmetry breaking and Goldstone modes for deep information propagation
by: Iqbal, Nabil, et al.
Published: (2026)
by: Iqbal, Nabil, et al.
Published: (2026)
MoH: Multi-Head Attention as Mixture-of-Head Attention
by: Jin, Peng, et al.
Published: (2024)
by: Jin, Peng, et al.
Published: (2024)
FasterViT: Fast Vision Transformers with Hierarchical Attention
by: Hatamizadeh, Ali, et al.
Published: (2023)
by: Hatamizadeh, Ali, et al.
Published: (2023)
Precipitation Nowcasting Using Diffusion Transformer with Causal Attention
by: Li, ChaoRong, et al.
Published: (2024)
by: Li, ChaoRong, et al.
Published: (2024)
Generalized Neighborhood Attention: Multi-dimensional Sparse Attention at the Speed of Light
by: Hassani, Ali, et al.
Published: (2025)
by: Hassani, Ali, et al.
Published: (2025)
RETR: Multi-View Radar Detection Transformer for Indoor Perception
by: Yataka, Ryoma, et al.
Published: (2024)
by: Yataka, Ryoma, et al.
Published: (2024)
Fairness-aware Vision Transformer via Debiased Self-Attention
by: Qiang, Yao, et al.
Published: (2023)
by: Qiang, Yao, et al.
Published: (2023)
Multi-View Hypercomplex Learning for Breast Cancer Screening
by: Lopez, Eleonora, et al.
Published: (2022)
by: Lopez, Eleonora, et al.
Published: (2022)
Random Token Fusion for Multi-View Medical Diagnosis
by: Guo, Jingyu, et al.
Published: (2024)
by: Guo, Jingyu, et al.
Published: (2024)
Geometric Point Attention Transformer for 3D Shape Reassembly
by: Li, Jiahan, et al.
Published: (2024)
by: Li, Jiahan, et al.
Published: (2024)
AttentionDrop: A Novel Regularization Method for Transformer Models
by: Baig, Mirza Samad Ahmed, et al.
Published: (2025)
by: Baig, Mirza Samad Ahmed, et al.
Published: (2025)
Beyond Conventional Transformers: The Medical X-ray Attention (MXA) Block for Improved Multi-Label Diagnosis Using Knowledge Distillation
by: Rand, Amit, et al.
Published: (2025)
by: Rand, Amit, et al.
Published: (2025)
Similar Items
-
Artificial Kuramoto Oscillatory Neurons
by: Miyato, Takeru, et al.
Published: (2024) -
Hidden Biases of End-to-End Driving Datasets
by: Zimmerlin, Julian, et al.
Published: (2024) -
End-to-end Autonomous Driving: Challenges and Frontiers
by: Chen, Li, et al.
Published: (2023) -
Erwin: A Tree-based Hierarchical Transformer for Large-scale Physical Systems
by: Zhdanov, Maksim, et al.
Published: (2025) -
LEAD: Minimizing Learner-Expert Asymmetry in End-to-End Driving
by: Nguyen, Long, et al.
Published: (2025)