RapidNet: Multi-Level Dilated Convolution Based Mobile Backbone
Fuente:
arXiv
Saved in:
| Main Authors: | Munir, Mustafa, Rahman, Md Mostafijur, Marculescu, Radu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AdaptViG: Adaptive Vision GNN with Exponential Decay Gating
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
EMCAD: Efficient Multi-scale Convolutional Attention Decoding for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2024)
by: Rahman, Md Mostafijur, et al.
Published: (2024)
GreedyViG: Dynamic Axial Graph Construction for Efficient Vision GNNs
by: Munir, Mustafa, et al.
Published: (2024)
by: Munir, Mustafa, et al.
Published: (2024)
PipeFlow: Pipelined Processing and Motion-Aware Frame Selection for Long-Form Video Editing
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
VCMamba: Bridging Convolutions with Multi-Directional Mamba for Efficient Visual Representation
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
MK-UNet: Multi-kernel Lightweight CNN for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2025)
by: Rahman, Md Mostafijur, et al.
Published: (2025)
LoMix: Learnable Weighted Multi-Scale Logits Mixing for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2025)
by: Rahman, Md Mostafijur, et al.
Published: (2025)
Ada-VE: Training-Free Consistent Video Editing Using Adaptive Motion Prior
by: Mahmud, Tanvir, et al.
Published: (2024)
by: Mahmud, Tanvir, et al.
Published: (2024)
Multi-Scale High-Resolution Logarithmic Grapher Module for Efficient Vision GNNs
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
PP-SAM: Perturbed Prompts for Robust Adaptation of Segment Anything Model for Polyp Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2024)
by: Rahman, Md Mostafijur, et al.
Published: (2024)
Scaling Graph Convolutions for Mobile Vision
by: Avery, William, et al.
Published: (2024)
by: Avery, William, et al.
Published: (2024)
An Efficient Dual-Line Decoder Network with Multi-Scale Convolutional Attention for Multi-organ Segmentation
by: Hassan, Riad, et al.
Published: (2025)
by: Hassan, Riad, et al.
Published: (2025)
BanglaNet: Bangla Handwritten Character Recognition using Ensembling of Convolutional Neural Network
by: Saha, Chandrika, et al.
Published: (2024)
by: Saha, Chandrika, et al.
Published: (2024)
AttentionViG: Cross-Attention-Based Dynamic Neighbor Aggregation in Vision GNNs
by: Gedik, Hakan Emre, et al.
Published: (2025)
by: Gedik, Hakan Emre, et al.
Published: (2025)
Fuel Gauge: Estimating Chain-of-Thought Length Ahead of Time in Large Multimodal Models
by: Yang, Yuedong, et al.
Published: (2026)
by: Yang, Yuedong, et al.
Published: (2026)
MobileAgeNet: Lightweight Facial Age Estimation for Mobile Deployment
by: Kumar, Arun, et al.
Published: (2026)
by: Kumar, Arun, et al.
Published: (2026)
A Mobile Application for Flower Recognition System Based on Convolutional Neural Networks
by: Yurdakul, Mustafa, et al.
Published: (2026)
by: Yurdakul, Mustafa, et al.
Published: (2026)
Revisiting the Integration of Convolution and Attention for Vision Backbone
by: Zhu, Lei, et al.
Published: (2024)
by: Zhu, Lei, et al.
Published: (2024)
DragonFruitQualityNet: A Lightweight Convolutional Neural Network for Real-Time Dragon Fruit Quality Inspection on Mobile Devices
by: Haquea, Md Zahurul, et al.
Published: (2025)
by: Haquea, Md Zahurul, et al.
Published: (2025)
Enhancing Satellite Object Localization with Dilated Convolutions and Attention-aided Spatial Pooling
by: Mostafa, Seraj Al Mahmud, et al.
Published: (2025)
by: Mostafa, Seraj Al Mahmud, et al.
Published: (2025)
Q-Sched: Pushing the Boundaries of Few-Step Diffusion Models with Quantization-Aware Scheduling
by: Frumkin, Natalia, et al.
Published: (2025)
by: Frumkin, Natalia, et al.
Published: (2025)
FUSED-Net: Detecting Traffic Signs with Limited Data
by: Rahman, Md. Atiqur, et al.
Published: (2024)
by: Rahman, Md. Atiqur, et al.
Published: (2024)
Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models
by: Rahman, Md Ashikur, et al.
Published: (2026)
by: Rahman, Md Ashikur, et al.
Published: (2026)
Dilated Convolution with Learnable Spacings makes visual models more aligned with humans: a Grad-CAM study
by: Chamas, Rabih, et al.
Published: (2024)
by: Chamas, Rabih, et al.
Published: (2024)
DilateQuant: Accurate and Efficient Diffusion Quantization via Weight Dilation
by: Liu, Xuewen, et al.
Published: (2024)
by: Liu, Xuewen, et al.
Published: (2024)
ObjectAlign: Neuro-Symbolic Object Consistency Verification and Correction
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
MRI-Based Brain Tumor Detection through an Explainable EfficientNetV2 and MLP-Mixer-Attention Architecture
by: Yurdakul, Mustafa, et al.
Published: (2025)
by: Yurdakul, Mustafa, et al.
Published: (2025)
ViscoNet: Bridging and Harmonizing Visual and Textual Conditioning for ControlNet
by: Cheong, Soon Yau, et al.
Published: (2023)
by: Cheong, Soon Yau, et al.
Published: (2023)
Malaria Detection from Blood Cell Images Using XceptionNet
by: Nusrat, Warisa, et al.
Published: (2025)
by: Nusrat, Warisa, et al.
Published: (2025)
A Cascaded Dilated Convolution Approach for Mpox Lesion Classification
by: Deshmukh, Ayush
Published: (2024)
by: Deshmukh, Ayush
Published: (2024)
Tighnari: Multi-modal Plant Species Prediction Based on Hierarchical Cross-Attention Using Graph-Based and Vision Backbone-Extracted Features
by: Liu, Haixu, et al.
Published: (2025)
by: Liu, Haixu, et al.
Published: (2025)
Video Compression Meets Video Generation: Latent Inter-Frame Pruning with Attention Recovery
by: Menn, Dennis, et al.
Published: (2026)
by: Menn, Dennis, et al.
Published: (2026)
Benchmarking ResNet Backbones in RT-DETR: Impact of Depth and Regularization under environmental conditions
by: Barboza, Pamela, et al.
Published: (2026)
by: Barboza, Pamela, et al.
Published: (2026)
HANS-Net: Hyperbolic Convolution and Adaptive Temporal Attention for Accurate and Generalizable Liver and Tumor Segmentation in CT Imaging
by: Abian, Arefin Ittesafun, et al.
Published: (2025)
by: Abian, Arefin Ittesafun, et al.
Published: (2025)
Efficiency Bottlenecks of Convolutional Kolmogorov-Arnold Networks: A Comprehensive Scrutiny with ImageNet, AlexNet, LeNet and Tabular Classification
by: Dahal, Ashim, et al.
Published: (2025)
by: Dahal, Ashim, et al.
Published: (2025)
MSA2-Net: Utilizing Self-Adaptive Convolution Module to Extract Multi-Scale Information in Medical Image Segmentation
by: Deng, Chao, et al.
Published: (2025)
by: Deng, Chao, et al.
Published: (2025)
PhytNet -- Tailored Convolutional Neural Networks for Custom Botanical Data
by: Sykes, Jamie R., et al.
Published: (2023)
by: Sykes, Jamie R., et al.
Published: (2023)
Multi-Level Feature Distillation of Joint Teachers Trained on Distinct Image Datasets
by: Iordache, Adrian, et al.
Published: (2024)
by: Iordache, Adrian, et al.
Published: (2024)
MobilePlantViT: A Mobile-friendly Hybrid ViT for Generalized Plant Disease Image Classification
by: Tonmoy, Moshiur Rahman, et al.
Published: (2025)
by: Tonmoy, Moshiur Rahman, et al.
Published: (2025)
DAUNet: A Lightweight UNet Variant with Deformable Convolutions and Parameter-Free Attention for Medical Image Segmentation
by: Munir, Adnan, et al.
Published: (2025)
by: Munir, Adnan, et al.
Published: (2025)
Similar Items
-
AdaptViG: Adaptive Vision GNN with Exponential Decay Gating
by: Munir, Mustafa, et al.
Published: (2025) -
EMCAD: Efficient Multi-scale Convolutional Attention Decoding for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2024) -
GreedyViG: Dynamic Axial Graph Construction for Efficient Vision GNNs
by: Munir, Mustafa, et al.
Published: (2024) -
PipeFlow: Pipelined Processing and Motion-Aware Frame Selection for Long-Form Video Editing
by: Munir, Mustafa, et al.
Published: (2025) -
VCMamba: Bridging Convolutions with Multi-Directional Mamba for Efficient Visual Representation
by: Munir, Mustafa, et al.
Published: (2025)