IBiT: Utilizing Inductive Biases to Create a More Data Efficient Attention Mechanism
Fuente:
arXiv
Saved in:
| Main Author: | Giri, Adithya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Leveraging Geometric Visual Illusions as Perceptual Inductive Biases for Vision Models
by: Yang, Haobo, et al.
Published: (2025)
by: Yang, Haobo, et al.
Published: (2025)
SageAttention2++: A More Efficient Implementation of SageAttention2
by: Zhang, Jintao, et al.
Published: (2025)
by: Zhang, Jintao, et al.
Published: (2025)
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics
by: Colagrande, Alex, et al.
Published: (2025)
by: Colagrande, Alex, et al.
Published: (2025)
ENA: Efficient N-dimensional Attention
by: Zhong, Yibo
Published: (2025)
by: Zhong, Yibo
Published: (2025)
Efficient Image Generation with Variadic Attention Heads
by: Walton, Steven, et al.
Published: (2022)
by: Walton, Steven, et al.
Published: (2022)
HierSum: A Global and Local Attention Mechanism for Video Summarization
by: Beedu, Apoorva, et al.
Published: (2025)
by: Beedu, Apoorva, et al.
Published: (2025)
GTA: A Geometry-Aware Attention Mechanism for Multi-View Transformers
by: Miyato, Takeru, et al.
Published: (2023)
by: Miyato, Takeru, et al.
Published: (2023)
Peer-Ranked Precision: Creating a Foundational Dataset for Fine-Tuning Vision Models from DataSeeds' Annotated Imagery
by: Abdoli, Sajjad, et al.
Published: (2025)
by: Abdoli, Sajjad, et al.
Published: (2025)
Can Biases in ImageNet Models Explain Generalization?
by: Gavrikov, Paul, et al.
Published: (2024)
by: Gavrikov, Paul, et al.
Published: (2024)
Biased Binary Attribute Classifiers Ignore the Majority Classes
by: Zhang, Xinyi, et al.
Published: (2024)
by: Zhang, Xinyi, et al.
Published: (2024)
TruKAN: Towards More Efficient Kolmogorov-Arnold Networks Using Truncated Power Functions
by: Bayeh, Ali, et al.
Published: (2026)
by: Bayeh, Ali, et al.
Published: (2026)
PSA: Pyramid Sparse Attention for Efficient Video Understanding and Generation
by: Li, Xiaolong, et al.
Published: (2025)
by: Li, Xiaolong, et al.
Published: (2025)
Attention-space Contrastive Guidance for Efficient Hallucination Mitigation in LVLMs
by: Jo, Yujin, et al.
Published: (2026)
by: Jo, Yujin, et al.
Published: (2026)
COMCAT: Towards Efficient Compression and Customization of Attention-Based Vision Models
by: Xiao, Jinqi, et al.
Published: (2023)
by: Xiao, Jinqi, et al.
Published: (2023)
Exploring the Efficacy of Meta-Learning: Unveiling Superior Data Diversity Utilization of MAML Over Pre-training
by: Selva, Kavita, et al.
Published: (2025)
by: Selva, Kavita, et al.
Published: (2025)
Prototype Guided Backdoor Defense
by: Amula, Venkat Adithya, et al.
Published: (2025)
by: Amula, Venkat Adithya, et al.
Published: (2025)
Fake or JPEG? Revealing Common Biases in Generated Image Detection Datasets
by: Grommelt, Patrick, et al.
Published: (2024)
by: Grommelt, Patrick, et al.
Published: (2024)
Embracing Biased Transition Matrices for Complementary-Label Learning with Many Classes
by: Mai, Tan-Ha, et al.
Published: (2026)
by: Mai, Tan-Ha, et al.
Published: (2026)
BLADE: Block-Sparse Attention Meets Step Distillation for Efficient Video Generation
by: Gu, Youping, et al.
Published: (2025)
by: Gu, Youping, et al.
Published: (2025)
Communication Efficient Split Learning of ViTs with Attention-based Double Compression
by: Alvetreti, Federico, et al.
Published: (2025)
by: Alvetreti, Federico, et al.
Published: (2025)
Three Creates All: You Only Sample 3 Steps
by: Cai, Yuren, et al.
Published: (2026)
by: Cai, Yuren, et al.
Published: (2026)
Data-Efficient Multimodal Fusion on a Single GPU
by: Vouitsis, Noël, et al.
Published: (2023)
by: Vouitsis, Noël, et al.
Published: (2023)
SuperLoRA: Parameter-Efficient Unified Adaptation of Multi-Layer Attention Modules
by: Chen, Xiangyu, et al.
Published: (2024)
by: Chen, Xiangyu, et al.
Published: (2024)
Human Activity Recognition from Wearable Sensor Data Using Self-Attention
by: Mahmud, Saif, et al.
Published: (2020)
by: Mahmud, Saif, et al.
Published: (2020)
Advanced Brain Tumor Segmentation Using EMCAD: Efficient Multi-scale Convolutional Attention Decoding
by: Uzor, GodsGift, et al.
Published: (2025)
by: Uzor, GodsGift, et al.
Published: (2025)
Hidden Biases of End-to-End Driving Datasets
by: Zimmerlin, Julian, et al.
Published: (2024)
by: Zimmerlin, Julian, et al.
Published: (2024)
T-TAME: Trainable Attention Mechanism for Explaining Convolutional Networks and Vision Transformers
by: Ntrougkas, Mariano V., et al.
Published: (2024)
by: Ntrougkas, Mariano V., et al.
Published: (2024)
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and $\mathcal{O}(T)$ Complexity
by: Zou, Shihao, et al.
Published: (2025)
by: Zou, Shihao, et al.
Published: (2025)
TRUST: Test-time Resource Utilization for Superior Trustworthiness
by: Harikumar, Haripriya, et al.
Published: (2025)
by: Harikumar, Haripriya, et al.
Published: (2025)
FedDistill: Global Model Distillation for Local Model De-Biasing in Non-IID Federated Learning
by: Song, Changlin, et al.
Published: (2024)
by: Song, Changlin, et al.
Published: (2024)
A More Word-like Image Tokenization for MLLMs
by: Lee, Hyun, et al.
Published: (2026)
by: Lee, Hyun, et al.
Published: (2026)
Make Your LVLM KV Cache More Lightweight
by: Chen, Xihao, et al.
Published: (2026)
by: Chen, Xihao, et al.
Published: (2026)
Not All Layers Are Created Equal: Adaptive LoRA Ranks for Personalized Image Generation
by: Shenaj, Donald, et al.
Published: (2026)
by: Shenaj, Donald, et al.
Published: (2026)
How Do Training Methods Influence the Utilization of Vision Models?
by: Gavrikov, Paul, et al.
Published: (2024)
by: Gavrikov, Paul, et al.
Published: (2024)
MoH: Multi-Head Attention as Mixture-of-Head Attention
by: Jin, Peng, et al.
Published: (2024)
by: Jin, Peng, et al.
Published: (2024)
What Variables Affect Out-of-Distribution Generalization in Pretrained Models?
by: Harun, Md Yousuf, et al.
Published: (2024)
by: Harun, Md Yousuf, et al.
Published: (2024)
Generalized Neighborhood Attention: Multi-dimensional Sparse Attention at the Speed of Light
by: Hassani, Ali, et al.
Published: (2025)
by: Hassani, Ali, et al.
Published: (2025)
MimiQ: Low-Bit Data-Free Quantization of Vision Transformers with Encouraging Inter-Head Attention Similarity
by: Choi, Kanghyun, et al.
Published: (2024)
by: Choi, Kanghyun, et al.
Published: (2024)
Investigation of Customized Medical Decision Algorithms Utilizing Graph Neural Networks
by: Yan, Yafeng, et al.
Published: (2024)
by: Yan, Yafeng, et al.
Published: (2024)
CUE-Net: Violence Detection Video Analytics with Spatial Cropping, Enhanced UniformerV2 and Modified Efficient Additive Attention
by: Senadeera, Damith Chamalke, et al.
Published: (2024)
by: Senadeera, Damith Chamalke, et al.
Published: (2024)
Similar Items
-
Leveraging Geometric Visual Illusions as Perceptual Inductive Biases for Vision Models
by: Yang, Haobo, et al.
Published: (2025) -
SageAttention2++: A More Efficient Implementation of SageAttention2
by: Zhang, Jintao, et al.
Published: (2025) -
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics
by: Colagrande, Alex, et al.
Published: (2025) -
ENA: Efficient N-dimensional Attention
by: Zhong, Yibo
Published: (2025) -
Efficient Image Generation with Variadic Attention Heads
by: Walton, Steven, et al.
Published: (2022)