A Comprehensive Review of YOLO Architectures in Computer Vision: From YOLOv1 to YOLOv8 and YOLO-NAS
Fuente:
arXiv
Saved in:
| Main Authors: | Terven, Juan, Cordova-Esparza, Diana |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiffYOLO: Object Detection for Anti-Noise via YOLO and Diffusion Models
by: Liu, Yichen, et al.
Published: (2024)
by: Liu, Yichen, et al.
Published: (2024)
Real-Time Flying Object Detection with YOLOv8
by: Reis, Dillon, et al.
Published: (2023)
by: Reis, Dillon, et al.
Published: (2023)
Land Cover Image Classification
by: Rangel, Antonio, et al.
Published: (2024)
by: Rangel, Antonio, et al.
Published: (2024)
Next-Generation License Plate Detection and Recognition System using YOLOv8
by: Amin, Arslan, et al.
Published: (2025)
by: Amin, Arslan, et al.
Published: (2025)
Enhancing Road Crack Detection Accuracy with BsS-YOLO: Optimizing Feature Fusion and Attention Mechanisms
by: Tang, Jiaze, et al.
Published: (2024)
by: Tang, Jiaze, et al.
Published: (2024)
Application of YOLOv8 in monocular downward multiple Car Target detection
by: Lyu, Shijie
Published: (2025)
by: Lyu, Shijie
Published: (2025)
An Analysis of Layer-Freezing Strategies for Enhanced Transfer Learning in YOLO Architectures
by: Dobrzycki, Andrzej D., et al.
Published: (2025)
by: Dobrzycki, Andrzej D., et al.
Published: (2025)
Quantizing YOLOv7: A Comprehensive Study
by: Baghbanbashi, Mohammadamin, et al.
Published: (2024)
by: Baghbanbashi, Mohammadamin, et al.
Published: (2024)
SPMamba-YOLO: An Underwater Object Detection Network Based on Multi-Scale Feature Enhancement and Global Context Modeling
by: Liao, Guanghao, et al.
Published: (2026)
by: Liao, Guanghao, et al.
Published: (2026)
Computer Vision for Clinical Gait Analysis: A Gait Abnormality Video Dataset
by: Ranjan, Rahm, et al.
Published: (2024)
by: Ranjan, Rahm, et al.
Published: (2024)
Hardware-Aware YOLO Compression for Low-Power Edge AI on STM32U5 for Weeds Detection in Digital Agriculture
by: Kouzinopoulos, Charalampos S., et al.
Published: (2025)
by: Kouzinopoulos, Charalampos S., et al.
Published: (2025)
Human-Centric Anomaly Detection in Surveillance Videos Using YOLO-World and Spatio-Temporal Deep Learning
by: Naeen, Mohammad Ali Etemadi, et al.
Published: (2025)
by: Naeen, Mohammad Ali Etemadi, et al.
Published: (2025)
BlindSight: Harnessing Sparsity for Efficient Vision-Language Models
by: Srikrishnan, Tharun Adithya, et al.
Published: (2025)
by: Srikrishnan, Tharun Adithya, et al.
Published: (2025)
U-Net-Like Spiking Neural Networks for Single Image Dehazing
by: Li, Huibin, et al.
Published: (2025)
by: Li, Huibin, et al.
Published: (2025)
MINT: Mitigating Hallucinations in Large Vision-Language Models via Token Reduction
by: Wang, Chao, et al.
Published: (2025)
by: Wang, Chao, et al.
Published: (2025)
VidNum-1.4K: A Comprehensive Benchmark for Video-based Numerical Reasoning
by: Cui, Shaoyang, et al.
Published: (2026)
by: Cui, Shaoyang, et al.
Published: (2026)
ViG-LRGC: Vision Graph Neural Networks with Learnable Reparameterized Graph Construction
by: Elsharkawi, Ismael, et al.
Published: (2025)
by: Elsharkawi, Ismael, et al.
Published: (2025)
Towards Hard and Soft Shadow Removal via Dual-Branch Separation Network and Vision Transformer
by: Liang, Jiajia
Published: (2025)
by: Liang, Jiajia
Published: (2025)
Do Generative Metrics Predict YOLO Performance? An Evaluation Across Models, Augmentation Ratios, and Dataset Complexity
by: Marian, Vasile, et al.
Published: (2026)
by: Marian, Vasile, et al.
Published: (2026)
IMASHRIMP: Automatic White Shrimp (Penaeus vannamei) Biometrical Analysis from Laboratory Images Using Computer Vision and Deep Learning
by: González, Abiam Remache, et al.
Published: (2025)
by: González, Abiam Remache, et al.
Published: (2025)
Boundary-Protection W8A8 HiFloat8 Quantization for Large-Scale Text-to-Video Diffusion Transformers
by: Zhao, Yiming
Published: (2026)
by: Zhao, Yiming
Published: (2026)
YOLO-APD: Enhancing YOLOv8 for Robust Pedestrian Detection on Complex Road Geometries
by: Joctum, Aquino, et al.
Published: (2025)
by: Joctum, Aquino, et al.
Published: (2025)
Pedestrian Detection in Low-Light Conditions: A Comprehensive Survey
by: Ghari, Bahareh, et al.
Published: (2024)
by: Ghari, Bahareh, et al.
Published: (2024)
Towards a Generalizable Fusion Architecture for Multimodal Object Detection
by: Berjawi, Jad, et al.
Published: (2025)
by: Berjawi, Jad, et al.
Published: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
by: Raoufi, Behnam, et al.
Published: (2025)
by: Raoufi, Behnam, et al.
Published: (2025)
Video Event Reasoning and Prediction by Fusing World Knowledge from LLMs with Vision Foundation Models
by: Dubois, L'ea, et al.
Published: (2025)
by: Dubois, L'ea, et al.
Published: (2025)
Evaluation of Attention Mechanisms in U-Net Architectures for Semantic Segmentation of Brazilian Rock Art Petroglyphs
by: Melo, Leonardi, et al.
Published: (2025)
by: Melo, Leonardi, et al.
Published: (2025)
A Vision-Language Model for Focal Liver Lesion Classification
by: Jian, Song, et al.
Published: (2025)
by: Jian, Song, et al.
Published: (2025)
NAC-TCN: Temporal Convolutional Networks with Causal Dilated Neighborhood Attention for Emotion Understanding
by: Mehta, Alexander, et al.
Published: (2023)
by: Mehta, Alexander, et al.
Published: (2023)
Context-Aware Indoor Point Cloud Object Generation through User Instructions
by: Luo, Yiyang, et al.
Published: (2023)
by: Luo, Yiyang, et al.
Published: (2023)
An Evaluation of a Visual Question Answering Strategy for Zero-shot Facial Expression Recognition in Still Images
by: Castrillón-Santana, Modesto, et al.
Published: (2025)
by: Castrillón-Santana, Modesto, et al.
Published: (2025)
A Challenging Benchmark of Anime Style Recognition
by: Li, Haotang, et al.
Published: (2022)
by: Li, Haotang, et al.
Published: (2022)
Textual and Visual Guided Task Adaptation for Source-Free Cross-Domain Few-Shot Segmentation
by: Liu, Jianming, et al.
Published: (2025)
by: Liu, Jianming, et al.
Published: (2025)
FUSE-Flow: Scalable Real-Time Multi-View Point Cloud Reconstruction Using Confidence
by: Sun, Chentian
Published: (2026)
by: Sun, Chentian
Published: (2026)
YotoR-You Only Transform One Representation
by: Villa, José Ignacio Díaz, et al.
Published: (2024)
by: Villa, José Ignacio Díaz, et al.
Published: (2024)
MSTA3D: Multi-scale Twin-attention for 3D Instance Segmentation
by: Tran, Duc Dang Trung, et al.
Published: (2024)
by: Tran, Duc Dang Trung, et al.
Published: (2024)
ERNet: Efficient Non-Rigid Registration Network for Point Sequences
by: He, Guangzhao, et al.
Published: (2025)
by: He, Guangzhao, et al.
Published: (2025)
GMAC: Global Multi-View Constraint for Automatic Multi-Camera Extrinsic Calibration
by: Sun, Chentian
Published: (2026)
by: Sun, Chentian
Published: (2026)
Escaping The Big Data Paradigm in Self-Supervised Representation Learning
by: García, Carlos Vélez, et al.
Published: (2025)
by: García, Carlos Vélez, et al.
Published: (2025)
A self-supervised cyclic neural-analytic approach for novel view synthesis and 3D reconstruction
by: Costea, Dragos, et al.
Published: (2025)
by: Costea, Dragos, et al.
Published: (2025)
Similar Items
-
DiffYOLO: Object Detection for Anti-Noise via YOLO and Diffusion Models
by: Liu, Yichen, et al.
Published: (2024) -
Real-Time Flying Object Detection with YOLOv8
by: Reis, Dillon, et al.
Published: (2023) -
Land Cover Image Classification
by: Rangel, Antonio, et al.
Published: (2024) -
Next-Generation License Plate Detection and Recognition System using YOLOv8
by: Amin, Arslan, et al.
Published: (2025) -
Enhancing Road Crack Detection Accuracy with BsS-YOLO: Optimizing Feature Fusion and Attention Mechanisms
by: Tang, Jiaze, et al.
Published: (2024)