Hybrid CNN-ViT Framework for Motion-Blurred Scene Text Restoration
Fuente:
arXiv
Saved in:
| Main Authors: | Rashid, Umar, Arshad, Muhammad Arslan, Ahmad, Ghulam, Anjum, Muhammad Zeeshan, Khan, Rizwan, Akmal, Muhammad |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DC-ViT: Modulating Spatial and Channel Interactions for Multi-Channel Images
by: Marikkar, Umar, et al.
Published: (2026)
by: Marikkar, Umar, et al.
Published: (2026)
CNN and ViT Efficiency Study on Tiny ImageNet and DermaMNIST Datasets
by: Amangeldi, Aidar, et al.
Published: (2025)
by: Amangeldi, Aidar, et al.
Published: (2025)
A Hybrid Framework Bridging CNN and ViT based on Theory of Evidence for Diabetic Retinopathy Grading
by: Qiu, Junlai, et al.
Published: (2025)
by: Qiu, Junlai, et al.
Published: (2025)
RepViT: Revisiting Mobile CNN From ViT Perspective
by: Wang, Ao, et al.
Published: (2023)
by: Wang, Ao, et al.
Published: (2023)
CNN-ViT Fusion with Adaptive Attention Gate for Brain Tumor MRI Classification: A Hybrid Deep Learning Model
by: Hasnain, Syed Ibad, et al.
Published: (2026)
by: Hasnain, Syed Ibad, et al.
Published: (2026)
A Hybrid CNN-ViT-GNN Framework with GAN-Based Augmentation for Intelligent Weed Detection in Precision Agriculture
by: V, Pandiyaraju, et al.
Published: (2025)
by: V, Pandiyaraju, et al.
Published: (2025)
Language-Unlocked ViT (LUViT): Empowering Self-Supervised Vision Transformers with LLMs
by: Kuzucu, Selim, et al.
Published: (2025)
by: Kuzucu, Selim, et al.
Published: (2025)
Learning CNN on ViT: A Hybrid Model to Explicitly Class-specific Boundaries for Domain Adaptation
by: Ngo, Ba Hung, et al.
Published: (2024)
by: Ngo, Ba Hung, et al.
Published: (2024)
Surface Defect Detection with Gabor Filter Using Reconstruction-Based Blurring U-Net-ViT
by: Si, Jongwook, et al.
Published: (2025)
by: Si, Jongwook, et al.
Published: (2025)
Derm-T2IM: Harnessing Synthetic Skin Lesion Data via Stable Diffusion Models for Enhanced Skin Disease Classification using ViT and CNN
by: Farooq, Muhammad Ali, et al.
Published: (2024)
by: Farooq, Muhammad Ali, et al.
Published: (2024)
S-E Pipeline: A Vision Transformer (ViT) based Resilient Classification Pipeline for Medical Imaging Against Adversarial Attacks
by: S, Neha A, et al.
Published: (2024)
by: S, Neha A, et al.
Published: (2024)
CNN-ViT Hybrid for Pneumonia Detection: Theory and Empiric on Limited Data without Pretraining
by: Basnet, Prashant Singh, et al.
Published: (2025)
by: Basnet, Prashant Singh, et al.
Published: (2025)
Combined CNN and ViT features off-the-shelf: Another astounding baseline for recognition
by: Alonso-Fernandez, Fernando, et al.
Published: (2024)
by: Alonso-Fernandez, Fernando, et al.
Published: (2024)
YOLO-Former: YOLO Shakes Hand With ViT
by: Khoramdel, Javad, et al.
Published: (2024)
by: Khoramdel, Javad, et al.
Published: (2024)
TESL-Net: A Transformer-Enhanced CNN for Accurate Skin Lesion Segmentation
by: Iqbal, Shahzaib, et al.
Published: (2024)
by: Iqbal, Shahzaib, et al.
Published: (2024)
SIMSPINE: A Biomechanics-Aware Simulation Framework for 3D Spine Motion Annotation and Benchmarking
by: Khan, Muhammad Saif Ullah, et al.
Published: (2026)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2026)
Modulating CNN Features with Pre-Trained ViT Representations for Open-Vocabulary Object Detection
by: Gao, Xiangyu, et al.
Published: (2025)
by: Gao, Xiangyu, et al.
Published: (2025)
ReConText3D: Replay-based Continual Text-to-3D Generation
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
by: Khan, Muhammad Ahmed Ullah, et al.
Published: (2026)
Motion Aware ViT-based Framework for Monocular 6-DoF Spacecraft Pose Estimation
by: Sosa, Jose, et al.
Published: (2025)
by: Sosa, Jose, et al.
Published: (2025)
Hybrid State-Space and GRU-based Graph Tokenization Mamba for Hyperspectral Image Classification
by: Ahmad, Muhammad, et al.
Published: (2025)
by: Ahmad, Muhammad, et al.
Published: (2025)
Deeper Inside Deep ViT
by: Hong, Sungrae
Published: (2025)
by: Hong, Sungrae
Published: (2025)
A Deep Features-Based Approach Using Modified ResNet50 and Gradient Boosting for Visual Sentiments Classification
by: Arslan, Muhammad, et al.
Published: (2024)
by: Arslan, Muhammad, et al.
Published: (2024)
Ensemble Deep Learning and LLM-Assisted Reporting for Automated Skin Lesion Diagnosis
by: Khan, Sher, et al.
Published: (2025)
by: Khan, Sher, et al.
Published: (2025)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
by: Zhong, Yunshan, et al.
Published: (2023)
by: Zhong, Yunshan, et al.
Published: (2023)
Sharpend Cosine Similarity based Neural Network for Hyperspectral Image Classification
by: Ahmad, Muhammad
Published: (2023)
by: Ahmad, Muhammad
Published: (2023)
Dynamic Memory Transformer for Hyperspectral Image Classification
by: Ahmad, Muhammad
Published: (2025)
by: Ahmad, Muhammad
Published: (2025)
3D Fourier-based Global Feature Extraction for Hyperspectral Image Classification
by: Ahmad, Muhammad
Published: (2026)
by: Ahmad, Muhammad
Published: (2026)
Context-Aware Detection of Mixed Critical Events using Video Classification
by: Akhlaq, Filza, et al.
Published: (2024)
by: Akhlaq, Filza, et al.
Published: (2024)
Generative Adversarial Network on Motion-Blur Image Restoration
by: Li, Zhengdong
Published: (2024)
by: Li, Zhengdong
Published: (2024)
A Lightweight and Interpretable Deepfakes Detection Framework
by: Farooq, Muhammad Umar, et al.
Published: (2025)
by: Farooq, Muhammad Umar, et al.
Published: (2025)
NURBGen: High-Fidelity Text-to-CAD Generation through LLM-Driven NURBS Modeling
by: Usama, Muhammad, et al.
Published: (2025)
by: Usama, Muhammad, et al.
Published: (2025)
Robust and Label-Efficient Deep Waste Detection
by: Abid, Hassan, et al.
Published: (2025)
by: Abid, Hassan, et al.
Published: (2025)
Hyb-KAN ViT: Hybrid Kolmogorov-Arnold Networks Augmented Vision Transformer
by: Dey, Sainath, et al.
Published: (2025)
by: Dey, Sainath, et al.
Published: (2025)
DVFL-Net: A Lightweight Distilled Video Focal Modulation Network for Spatio-Temporal Action Recognition
by: Ullah, Hayat, et al.
Published: (2025)
by: Ullah, Hayat, et al.
Published: (2025)
3D Reconstruction via Incremental Structure From Motion
by: Zeeshan, Muhammad, et al.
Published: (2025)
by: Zeeshan, Muhammad, et al.
Published: (2025)
Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review
by: Khan, Ufaq, et al.
Published: (2025)
by: Khan, Ufaq, et al.
Published: (2025)
Towards Automated Solar Panel Integrity: Hybrid Deep Feature Extraction for Advanced Surface Defect Identification
by: Asif, Muhammad Junaid, et al.
Published: (2026)
by: Asif, Muhammad Junaid, et al.
Published: (2026)
TransResNet: Integrating the Strengths of ViTs and CNNs for High Resolution Medical Image Segmentation via Feature Grafting
by: Sharif, Muhammad Hamza, et al.
Published: (2024)
by: Sharif, Muhammad Hamza, et al.
Published: (2024)
CountZES: Counting via Zero-Shot Exemplar Selection
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
Domain Adaptation Without the Compute Burden for Efficient Whole Slide Image Analysis
by: Marikkar, Umar, et al.
Published: (2026)
by: Marikkar, Umar, et al.
Published: (2026)
Similar Items
-
DC-ViT: Modulating Spatial and Channel Interactions for Multi-Channel Images
by: Marikkar, Umar, et al.
Published: (2026) -
CNN and ViT Efficiency Study on Tiny ImageNet and DermaMNIST Datasets
by: Amangeldi, Aidar, et al.
Published: (2025) -
A Hybrid Framework Bridging CNN and ViT based on Theory of Evidence for Diabetic Retinopathy Grading
by: Qiu, Junlai, et al.
Published: (2025) -
RepViT: Revisiting Mobile CNN From ViT Perspective
by: Wang, Ao, et al.
Published: (2023) -
CNN-ViT Fusion with Adaptive Attention Gate for Brain Tumor MRI Classification: A Hybrid Deep Learning Model
by: Hasnain, Syed Ibad, et al.
Published: (2026)