NiNformer: A Network in Network Transformer with Token Mixing Generated Gating Function
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abdullah, Abdullah Nazhat, Aydin, Tarkan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Activator: GLU Activation Function as the Core Component of a Vision Transformer
von: Abdullah, Abdullah Nazhat, et al.
Veröffentlicht: (2024)
von: Abdullah, Abdullah Nazhat, et al.
Veröffentlicht: (2024)
VLM-KG: Multimodal Radiology Knowledge Graph Generation
von: Abdullah, Abdullah, et al.
Veröffentlicht: (2025)
von: Abdullah, Abdullah, et al.
Veröffentlicht: (2025)
Convolutional Differentiable Logic Gate Networks
von: Petersen, Felix, et al.
Veröffentlicht: (2024)
von: Petersen, Felix, et al.
Veröffentlicht: (2024)
Transformer-Based Contrastive Meta-Learning For Low-Resource Generalizable Activity Recognition
von: Wang, Junyao, et al.
Veröffentlicht: (2024)
von: Wang, Junyao, et al.
Veröffentlicht: (2024)
A Real-time Face Mask Detection and Social Distancing System for COVID-19 using Attention-InceptionV3 Model
von: Asif, Abdullah Al, et al.
Veröffentlicht: (2024)
von: Asif, Abdullah Al, et al.
Veröffentlicht: (2024)
Quantum Machine Learning for Image Classification: A Hybrid Model of Residual Network with Quantum Support Vector Machine
von: Shahriyar, Md. Farhan, et al.
Veröffentlicht: (2025)
von: Shahriyar, Md. Farhan, et al.
Veröffentlicht: (2025)
Quanvolutional Neural Networks for Pneumonia Detection: An Efficient Quantum-Assisted Feature Extraction Paradigm
von: Tanbhir, Gazi, et al.
Veröffentlicht: (2025)
von: Tanbhir, Gazi, et al.
Veröffentlicht: (2025)
Metric Unreliability in Multimodal Machine Unlearning: A Systematic Analysis and Principled Unified Score
von: Khan, Abdullah Ahmad, et al.
Veröffentlicht: (2026)
von: Khan, Abdullah Ahmad, et al.
Veröffentlicht: (2026)
TransformMix: Learning Transformation and Mixing Strategies from Data
von: Cheung, Tsz-Him, et al.
Veröffentlicht: (2024)
von: Cheung, Tsz-Him, et al.
Veröffentlicht: (2024)
UKBOB: One Billion MRI Labeled Masks for Generalizable 3D Medical Image Segmentation
von: Bourigault, Emmanuelle, et al.
Veröffentlicht: (2025)
von: Bourigault, Emmanuelle, et al.
Veröffentlicht: (2025)
BanglaMM-Disaster: A Multimodal Transformer-Based Deep Learning Framework for Multiclass Disaster Classification in Bangla
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
von: Islam, Ariful, et al.
Veröffentlicht: (2025)
Zero-Shot Action Generalization with Limited Observations
von: Alchihabi, Abdullah, et al.
Veröffentlicht: (2025)
von: Alchihabi, Abdullah, et al.
Veröffentlicht: (2025)
DualSwinFusionSeg: Multimodal Martian Landslide Segmentation via Dual Swin Transformer with Multi-Scale Fusion and UNet++
von: Kabir, Shahriar, et al.
Veröffentlicht: (2026)
von: Kabir, Shahriar, et al.
Veröffentlicht: (2026)
MADFormer: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation
von: Chen, Junhao, et al.
Veröffentlicht: (2025)
von: Chen, Junhao, et al.
Veröffentlicht: (2025)
Token Caching for Diffusion Transformer Acceleration
von: Lou, Jinming, et al.
Veröffentlicht: (2024)
von: Lou, Jinming, et al.
Veröffentlicht: (2024)
SFC-GAN: A Generative Adversarial Network for Brain Functional and Structural Connectome Translation
von: Tan, Yee-Fan, et al.
Veröffentlicht: (2025)
von: Tan, Yee-Fan, et al.
Veröffentlicht: (2025)
AR-GAN: Generative Adversarial Network-Based Defense Method Against Adversarial Attacks on the Traffic Sign Classification System of Autonomous Vehicles
von: Salek, M Sabbir, et al.
Veröffentlicht: (2023)
von: Salek, M Sabbir, et al.
Veröffentlicht: (2023)
Leveraging Transformers for Weakly Supervised Object Localization in Unconstrained Videos
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
von: Murtaza, Shakeeb, et al.
Veröffentlicht: (2024)
SugarcaneShuffleNet: A Very Fast, Lightweight Convolutional Neural Network for Diagnosis of 15 Sugarcane Leaf Diseases
von: Arman, Shifat E., et al.
Veröffentlicht: (2025)
von: Arman, Shifat E., et al.
Veröffentlicht: (2025)
Efficient Visual Transformer by Learnable Token Merging
von: Wang, Yancheng, et al.
Veröffentlicht: (2024)
von: Wang, Yancheng, et al.
Veröffentlicht: (2024)
SPoT: Subpixel Placement of Tokens in Vision Transformers
von: Hjelkrem-Tan, Martine, et al.
Veröffentlicht: (2025)
von: Hjelkrem-Tan, Martine, et al.
Veröffentlicht: (2025)
MosquitoFusion: A Multiclass Dataset for Real-Time Detection of Mosquitoes, Swarms, and Breeding Sites Using Deep Learning
von: Sayeedi, Md. Faiyaz Abdullah, et al.
Veröffentlicht: (2024)
von: Sayeedi, Md. Faiyaz Abdullah, et al.
Veröffentlicht: (2024)
ABFR-KAN: Kolmogorov-Arnold Networks for Functional Brain Analysis
von: Ward, Tyler, et al.
Veröffentlicht: (2026)
von: Ward, Tyler, et al.
Veröffentlicht: (2026)
Uncertainty-Aware Token Importance Estimation in Spiking Transformers
von: Liu, Wenxuan, et al.
Veröffentlicht: (2026)
von: Liu, Wenxuan, et al.
Veröffentlicht: (2026)
AdaPerceiver: Transformers with Adaptive Width, Depth, and Tokens
von: Jajal, Purvish, et al.
Veröffentlicht: (2025)
von: Jajal, Purvish, et al.
Veröffentlicht: (2025)
Tokenizing Buildings: A Transformer for Layout Synthesis
von: de Guevara, Manuel Ladron, et al.
Veröffentlicht: (2025)
von: de Guevara, Manuel Ladron, et al.
Veröffentlicht: (2025)
Capsule Endoscopy Multi-classification via Gated Attention and Wavelet Transformations
von: Panchananam, Lakshmi Srinivas, et al.
Veröffentlicht: (2024)
von: Panchananam, Lakshmi Srinivas, et al.
Veröffentlicht: (2024)
SPAQ-DL-SLAM: Towards Optimizing Deep Learning-based SLAM for Resource-Constrained Embedded Platforms
von: Pudasaini, Niraj, et al.
Veröffentlicht: (2024)
von: Pudasaini, Niraj, et al.
Veröffentlicht: (2024)
Fractional Concepts in Neural Networks: Enhancing Activation Functions
von: Alijani, Zahra, et al.
Veröffentlicht: (2023)
von: Alijani, Zahra, et al.
Veröffentlicht: (2023)
Calibrating Bayesian UNet++ for Sub-Seasonal Forecasting
von: Asan, Busra, et al.
Veröffentlicht: (2024)
von: Asan, Busra, et al.
Veröffentlicht: (2024)
ShrinkBox: Backdoor Attack on Object Detection to Disrupt Collision Avoidance in Machine Learning-based Advanced Driver Assistance Systems
von: Shahzad, Muhammad Zaeem, et al.
Veröffentlicht: (2025)
von: Shahzad, Muhammad Zaeem, et al.
Veröffentlicht: (2025)
SATA: Spatial Autocorrelation Token Analysis for Enhancing the Robustness of Vision Transformers
von: Nikzad, Nick, et al.
Veröffentlicht: (2024)
von: Nikzad, Nick, et al.
Veröffentlicht: (2024)
Don't Look Twice: Faster Video Transformers with Run-Length Tokenization
von: Choudhury, Rohan, et al.
Veröffentlicht: (2024)
von: Choudhury, Rohan, et al.
Veröffentlicht: (2024)
Hybrid Spiking Neural Network -- Transformer Video Classification Model
von: Bateni, Aaron
Veröffentlicht: (2024)
von: Bateni, Aaron
Veröffentlicht: (2024)
VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations
von: Patel, Maitreya, et al.
Veröffentlicht: (2026)
von: Patel, Maitreya, et al.
Veröffentlicht: (2026)
Efficient and Effective Methods for Mixed Precision Neural Network Quantization for Faster, Energy-efficient Inference
von: Bablani, Deepika, et al.
Veröffentlicht: (2023)
von: Bablani, Deepika, et al.
Veröffentlicht: (2023)
ArcGate: Adaptive Arctangent Gated Activation
von: Bhattacharya, Avik, et al.
Veröffentlicht: (2026)
von: Bhattacharya, Avik, et al.
Veröffentlicht: (2026)
Convolutional Neural Networks and Vision Transformers for Fashion MNIST Classification: A Literature Review
von: Bbouzidi, Sonia, et al.
Veröffentlicht: (2024)
von: Bbouzidi, Sonia, et al.
Veröffentlicht: (2024)
A Novel Shape Guided Transformer Network for Instance Segmentation in Remote Sensing Images
von: Yu, Dawen, et al.
Veröffentlicht: (2024)
von: Yu, Dawen, et al.
Veröffentlicht: (2024)
A General and Efficient Training for Transformer via Token Expansion
von: Huang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Activator: GLU Activation Function as the Core Component of a Vision Transformer
von: Abdullah, Abdullah Nazhat, et al.
Veröffentlicht: (2024) -
VLM-KG: Multimodal Radiology Knowledge Graph Generation
von: Abdullah, Abdullah, et al.
Veröffentlicht: (2025) -
Convolutional Differentiable Logic Gate Networks
von: Petersen, Felix, et al.
Veröffentlicht: (2024) -
Transformer-Based Contrastive Meta-Learning For Low-Resource Generalizable Activity Recognition
von: Wang, Junyao, et al.
Veröffentlicht: (2024) -
A Real-time Face Mask Detection and Social Distancing System for COVID-19 using Attention-InceptionV3 Model
von: Asif, Abdullah Al, et al.
Veröffentlicht: (2024)