A comprehensive overview of deep learning models for object detection from videos/images
Fuente:
arXiv
Saved in:
| Main Authors: | Zulfqar, Sukana, Saeed, Sadia, Zia, M. Azam, Ali, Anjum, Mehmood, Faisal, Ali, Abid |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Geometry-Aware Semantic Reasoning for Training Free Video Anomaly Detection
by: Zia, Ali, et al.
Published: (2026)
by: Zia, Ali, et al.
Published: (2026)
Leveraging counterfactual concepts for debugging and improving CNN model performance
by: Tariq, Syed Ali, et al.
Published: (2025)
by: Tariq, Syed Ali, et al.
Published: (2025)
Deep learning for automated detection of breast cancer in deep ultraviolet fluorescence images with diffusion probabilistic model
by: Ghahfarokhi, Sepehr Salem, et al.
Published: (2024)
by: Ghahfarokhi, Sepehr Salem, et al.
Published: (2024)
Generative deep learning for foundational video translation in ultrasound
by: Tomic, Nikolina, et al.
Published: (2025)
by: Tomic, Nikolina, et al.
Published: (2025)
Addressing a fundamental limitation in deep vision models: lack of spatial attention
by: Borji, Ali
Published: (2024)
by: Borji, Ali
Published: (2024)
Violence detection in videos using deep recurrent and convolutional neural networks
by: Traoré, Abdarahmane, et al.
Published: (2024)
by: Traoré, Abdarahmane, et al.
Published: (2024)
A deep learning pipeline for PAM50 subtype classification using histopathology images and multi-objective patch selection
by: Borji, Arezoo, et al.
Published: (2026)
by: Borji, Arezoo, et al.
Published: (2026)
Determining Mosaic Resilience in Sugarcane Plants using Hyperspectral Images
by: Zia, Ali, et al.
Published: (2025)
by: Zia, Ali, et al.
Published: (2025)
Towards Counterfactual and Contrastive Explainability and Transparency of DCNN Image Classifiers
by: Tariq, Syed Ali, et al.
Published: (2025)
by: Tariq, Syed Ali, et al.
Published: (2025)
How well are open sourced AI-generated image detection models out-of-the-box: A comprehensive benchmark study
by: Ren, Simiao, et al.
Published: (2026)
by: Ren, Simiao, et al.
Published: (2026)
A detection-task-specific deep-learning method to improve the quality of sparse-view myocardial perfusion SPECT images
by: Yang, Zezhang, et al.
Published: (2025)
by: Yang, Zezhang, et al.
Published: (2025)
STCL:Curriculum learning Strategies for deep learning image steganography models
by: Liu, Fengchun, et al.
Published: (2025)
by: Liu, Fengchun, et al.
Published: (2025)
A benchmark dataset for deep learning-based airplane detection: HRPlanes
by: Bakirman, Tolga, et al.
Published: (2022)
by: Bakirman, Tolga, et al.
Published: (2022)
Hybrid deep convolution model for lung cancer detection with transfer learning
by: Saxena, Sugandha, et al.
Published: (2025)
by: Saxena, Sugandha, et al.
Published: (2025)
SSTFB: Leveraging self-supervised pretext learning and temporal self-attention with feature branching for real-time video polyp segmentation
by: Xu, Ziang, et al.
Published: (2024)
by: Xu, Ziang, et al.
Published: (2024)
WS-Net: Weak-Signal Representation Learning and Gated Abundance Reconstruction for Hyperspectral Unmixing via State-Space and Weak Signal Attention Fusion
by: Long, Zekun, et al.
Published: (2026)
by: Long, Zekun, et al.
Published: (2026)
KAN-Mixers: a new deep learning architecture for image classification
by: Canuto, Jorge Luiz dos Santos, et al.
Published: (2025)
by: Canuto, Jorge Luiz dos Santos, et al.
Published: (2025)
A multimodal deep learning architecture for smoking detection with a small data approach
by: Lakatos, Robert, et al.
Published: (2023)
by: Lakatos, Robert, et al.
Published: (2023)
Prompt-Based Continual Compositional Zero-Shot Learning
by: Maryam, Sauda, et al.
Published: (2025)
by: Maryam, Sauda, et al.
Published: (2025)
Leaf diseases detection using deep learning methods
by: Fatimi, El Houcine El
Published: (2024)
by: Fatimi, El Houcine El
Published: (2024)
Using deep learning for predicting cleansing quality of colon capsule endoscopy images
by: Sharma, Puneet, et al.
Published: (2026)
by: Sharma, Puneet, et al.
Published: (2026)
IGAN: A New Inception-based Model for Stable and High-Fidelity Image Synthesis Using Generative Adversarial Networks
by: Hashim, Ahmed A., et al.
Published: (2026)
by: Hashim, Ahmed A., et al.
Published: (2026)
Deep learning approaches to surgical video segmentation and object detection: A Scoping Review
by: Kamtam, Devanish N., et al.
Published: (2025)
by: Kamtam, Devanish N., et al.
Published: (2025)
Real-Time Object Detection in Occluded Environment with Background Cluttering Effects Using Deep Learning
by: Aamir, Syed Muhammad, et al.
Published: (2024)
by: Aamir, Syed Muhammad, et al.
Published: (2024)
EAGLE: Egocentric AGgregated Language-video Engine
by: Bi, Jing, et al.
Published: (2024)
by: Bi, Jing, et al.
Published: (2024)
Training deep learning based dynamic MR image reconstruction using synthetic fractals
by: Raman, Anirudh, et al.
Published: (2026)
by: Raman, Anirudh, et al.
Published: (2026)
FusionVision: A comprehensive approach of 3D object reconstruction and segmentation from RGB-D cameras using YOLO and fast segment anything
by: Ghazouali, Safouane El, et al.
Published: (2024)
by: Ghazouali, Safouane El, et al.
Published: (2024)
A review of deep learning-based information fusion techniques for multimodal medical image classification
by: Li, Yihao, et al.
Published: (2024)
by: Li, Yihao, et al.
Published: (2024)
Individual mapping of large polymorphic shrubs in high mountains using satellite images and deep learning
by: Khaldi, Rohaifa, et al.
Published: (2024)
by: Khaldi, Rohaifa, et al.
Published: (2024)
Precise localization of corneal reflections in eye images using deep learning trained on synthetic data
by: Byrne, Sean Anthony, et al.
Published: (2023)
by: Byrne, Sean Anthony, et al.
Published: (2023)
Study of detecting behavioral signatures within DeepFake videos
by: Miao, Qiaomu, et al.
Published: (2022)
by: Miao, Qiaomu, et al.
Published: (2022)
Multi-model approach for autonomous driving: A comprehensive study on traffic sign-, vehicle- and lane detection and behavioral cloning
by: Jaisankar, Kanishkha, et al.
Published: (2026)
by: Jaisankar, Kanishkha, et al.
Published: (2026)
A Review on Coarse to Fine-Grained Animal Action Recognition
by: Zia, Ali, et al.
Published: (2025)
by: Zia, Ali, et al.
Published: (2025)
Multi-objective hybrid knowledge distillation for efficient deep learning in smart agriculture
by: Hoang, Phi-Hung, et al.
Published: (2025)
by: Hoang, Phi-Hung, et al.
Published: (2025)
Model-agnostic explainable artificial intelligence for object detection in image data
by: Moradi, Milad, et al.
Published: (2023)
by: Moradi, Milad, et al.
Published: (2023)
Latent Video Prediction Learns Better World Models
by: Alrasheed, Ali J, et al.
Published: (2026)
by: Alrasheed, Ali J, et al.
Published: (2026)
Deep learning-based automated damage detection in concrete structures using images from earthquake events
by: Turer, Abdullah, et al.
Published: (2025)
by: Turer, Abdullah, et al.
Published: (2025)
AI Driven Water Segmentation with deep learning models for Enhanced Flood Monitoring
by: Mou, Sanjida Afrin, et al.
Published: (2025)
by: Mou, Sanjida Afrin, et al.
Published: (2025)
GlitchBench: Can large multimodal models detect video game glitches?
by: Taesiri, Mohammad Reza, et al.
Published: (2023)
by: Taesiri, Mohammad Reza, et al.
Published: (2023)
Deep learning in computed tomography pulmonary angiography imaging: a dual-pronged approach for pulmonary embolism detection
by: Bushra, Fabiha, et al.
Published: (2023)
by: Bushra, Fabiha, et al.
Published: (2023)
Similar Items
-
Geometry-Aware Semantic Reasoning for Training Free Video Anomaly Detection
by: Zia, Ali, et al.
Published: (2026) -
Leveraging counterfactual concepts for debugging and improving CNN model performance
by: Tariq, Syed Ali, et al.
Published: (2025) -
Deep learning for automated detection of breast cancer in deep ultraviolet fluorescence images with diffusion probabilistic model
by: Ghahfarokhi, Sepehr Salem, et al.
Published: (2024) -
Generative deep learning for foundational video translation in ultrasound
by: Tomic, Nikolina, et al.
Published: (2025) -
Addressing a fundamental limitation in deep vision models: lack of spatial attention
by: Borji, Ali
Published: (2024)