iFormer: Integrating ConvNet and Transformer for Mobile Application
Fuente:
arXiv
Saved in:
| Main Author: | Zheng, Chuanyang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MixMask: Revisiting Masking Strategy for Siamese ConvNets
by: Vishniakov, Kirill, et al.
Published: (2022)
by: Vishniakov, Kirill, et al.
Published: (2022)
RepVGG-GELAN: Enhanced GELAN with VGG-STYLE ConvNets for Brain Tumour Detection
by: Balakrishnan, Thennarasi, et al.
Published: (2024)
by: Balakrishnan, Thennarasi, et al.
Published: (2024)
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
by: Chin, Zhi-Yi, et al.
Published: (2023)
by: Chin, Zhi-Yi, et al.
Published: (2023)
Scaling Up Your Kernels: Large Kernel Design in ConvNets towards Universal Representations
by: Zhang, Yiyuan, et al.
Published: (2024)
by: Zhang, Yiyuan, et al.
Published: (2024)
Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
by: Chen, Hao, et al.
Published: (2022)
by: Chen, Hao, et al.
Published: (2022)
UniRepLKNet: A Universal Perception Large-Kernel ConvNet for Audio, Video, Point Cloud, Time-Series and Image Recognition
by: Ding, Xiaohan, et al.
Published: (2023)
by: Ding, Xiaohan, et al.
Published: (2023)
ConvNet vs Transformer, Supervised vs CLIP: Beyond ImageNet Accuracy
by: Vishniakov, Kirill, et al.
Published: (2023)
by: Vishniakov, Kirill, et al.
Published: (2023)
Surface EMG-Based Inter-Session/Inter-Subject Gesture Recognition by Leveraging Lightweight All-ConvNet and Transfer Learning
by: Islam, Md. Rabiul, et al.
Published: (2023)
by: Islam, Md. Rabiul, et al.
Published: (2023)
Ultrafast-and-Ultralight ConvNet-Based Intelligent Monitoring System for Diagnosing Early-Stage Mpox Anytime and Anywhere
by: Yue, Yubiao, et al.
Published: (2023)
by: Yue, Yubiao, et al.
Published: (2023)
Human-annotated label noise and their impact on ConvNets for remote sensing image scene classification
by: Peng, Longkang, et al.
Published: (2023)
by: Peng, Longkang, et al.
Published: (2023)
MedNeXt: Transformer-driven Scaling of ConvNets for Medical Image Segmentation
by: Roy, Saikat, et al.
Published: (2023)
by: Roy, Saikat, et al.
Published: (2023)
BCFPL: Binary classification ConvNet based Fast Parking space recognition with Low resolution image
by: Zhang, Shuo, et al.
Published: (2024)
by: Zhang, Shuo, et al.
Published: (2024)
The Linear Attention Resurrection in Vision Transformer
by: Zheng, Chuanyang
Published: (2025)
by: Zheng, Chuanyang
Published: (2025)
PDiscoFormer: Relaxing Part Discovery Constraints with Vision Transformers
by: Aniraj, Ananthu, et al.
Published: (2024)
by: Aniraj, Ananthu, et al.
Published: (2024)
ScribFormer: Transformer Makes CNN Work Better for Scribble-based Medical Image Segmentation
by: Li, Zihan, et al.
Published: (2024)
by: Li, Zihan, et al.
Published: (2024)
HGTS-Former: Hierarchical HyperGraph Transformer for Multivariate Time Series Analysis
by: Si, Hao, et al.
Published: (2025)
by: Si, Hao, et al.
Published: (2025)
Reviving ConvNeXt for Efficient Convolutional Diffusion Models
by: Kwon, Taesung, et al.
Published: (2026)
by: Kwon, Taesung, et al.
Published: (2026)
InceptionNeXt: When Inception Meets ConvNeXt
by: Yu, Weihao, et al.
Published: (2023)
by: Yu, Weihao, et al.
Published: (2023)
Designing Concise ConvNets with Columnar Stages
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
Enhancing kelp forest detection in remote sensing images using crowdsourced labels with Mixed Vision Transformers and ConvNeXt segmentation models
by: Nasios, Ioannis
Published: (2025)
by: Nasios, Ioannis
Published: (2025)
MetaFormer Baselines for Vision
by: Yu, Weihao, et al.
Published: (2022)
by: Yu, Weihao, et al.
Published: (2022)
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and $\mathcal{O}(T)$ Complexity
by: Zou, Shihao, et al.
Published: (2025)
by: Zou, Shihao, et al.
Published: (2025)
Soft-TransFormers for Continual Learning
by: Kang, Haeyong, et al.
Published: (2024)
by: Kang, Haeyong, et al.
Published: (2024)
GaussianFormer-2: Probabilistic Gaussian Superposition for Efficient 3D Occupancy Prediction
by: Huang, Yuanhui, et al.
Published: (2024)
by: Huang, Yuanhui, et al.
Published: (2024)
ConvNets for Counting: Object Detection of Transient Phenomena in Steelpan Drums
by: Hawley, Scott H., et al.
Published: (2021)
by: Hawley, Scott H., et al.
Published: (2021)
RigidFormer: Learning Rigid Dynamics using Transformers
by: Dou, Zhiyang, et al.
Published: (2026)
by: Dou, Zhiyang, et al.
Published: (2026)
Residual-SwinCA-Net: A Channel-Aware Integrated Residual CNN-Swin Transformer for Malignant Lesion Segmentation in BUSI
by: Naz, Saeeda, et al.
Published: (2025)
by: Naz, Saeeda, et al.
Published: (2025)
VisTabNet: Adapting Vision Transformers for Tabular Data
by: Wydmański, Witold, et al.
Published: (2024)
by: Wydmański, Witold, et al.
Published: (2024)
Pick-or-Mix: Dynamic Channel Sampling for ConvNets
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
Universal Neural Architecture Space: Covering ConvNets, Transformers and Everything in Between
by: Týbl, Ondřej, et al.
Published: (2025)
by: Týbl, Ondřej, et al.
Published: (2025)
Convolutional Neural Nets vs Vision Transformers: A SpaceNet Case Study with Balanced vs Imbalanced Regimes
by: Gothi, Akshar
Published: (2025)
by: Gothi, Akshar
Published: (2025)
JetFormer: An Autoregressive Generative Model of Raw Images and Text
by: Tschannen, Michael, et al.
Published: (2024)
by: Tschannen, Michael, et al.
Published: (2024)
MobileIE: An Extremely Lightweight and Effective ConvNet for Real-Time Image Enhancement on Mobile Devices
by: Yan, Hailong, et al.
Published: (2025)
by: Yan, Hailong, et al.
Published: (2025)
Efficient Deformable ConvNets: Rethinking Dynamic and Sparse Operator for Vision Applications
by: Xiong, Yuwen, et al.
Published: (2024)
by: Xiong, Yuwen, et al.
Published: (2024)
EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided Reinforcement Learning
by: Jiang, Yue, et al.
Published: (2024)
by: Jiang, Yue, et al.
Published: (2024)
Multi-Slice Spatial Transcriptomics Data Integration Analysis with STG3Net
by: Fang, Donghai, et al.
Published: (2024)
by: Fang, Donghai, et al.
Published: (2024)
CaptionFormer: Unified Segmentation, Tracking, and Captioning for Spatio-Temporal Objects
by: Fiastre, Gabriel, et al.
Published: (2025)
by: Fiastre, Gabriel, et al.
Published: (2025)
MindFormer: Semantic Alignment of Multi-Subject fMRI for Brain Decoding
by: Han, Inhwa, et al.
Published: (2024)
by: Han, Inhwa, et al.
Published: (2024)
Learning to Generate Parameters of ConvNets for Unseen Image Data
by: Wang, Shiye, et al.
Published: (2023)
by: Wang, Shiye, et al.
Published: (2023)
VideoMAC: Video Masked Autoencoders Meet ConvNets
by: Pei, Gensheng, et al.
Published: (2024)
by: Pei, Gensheng, et al.
Published: (2024)
Similar Items
-
MixMask: Revisiting Masking Strategy for Siamese ConvNets
by: Vishniakov, Kirill, et al.
Published: (2022) -
RepVGG-GELAN: Enhanced GELAN with VGG-STYLE ConvNets for Brain Tumour Detection
by: Balakrishnan, Thennarasi, et al.
Published: (2024) -
Masking Improves Contrastive Self-Supervised Learning for ConvNets, and Saliency Tells You Where
by: Chin, Zhi-Yi, et al.
Published: (2023) -
Scaling Up Your Kernels: Large Kernel Design in ConvNets towards Universal Representations
by: Zhang, Yiyuan, et al.
Published: (2024) -
Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
by: Chen, Hao, et al.
Published: (2022)