YOLO advances to its genesis: a decadal and comprehensive review of the You Only Look Once (YOLO) series
Fuente:
arXiv
Saved in:
| Main Authors: | Sapkota, Ranjan, Calero, Marco Flores, Qureshi, Rizwan, Badgujar, Chetan, Nepal, Upesh, Poulose, Alwin, Zeno, Peter, Vaddevolu, Uday Bhanu Prakash, Khan, Sheheryar, Shoman, Maged, Yan, Hong, Karkee, Manoj |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agricultural Object Detection with You Look Only Once (YOLO) Algorithm: A Bibliometric and Systematic Literature Review
by: Badgujar, Chetan M, et al.
Published: (2024)
by: Badgujar, Chetan M, et al.
Published: (2024)
Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
YOLOE-26: Integrating YOLO26 with YOLOE for Real-Time Open-Vocabulary Instance Segmentation
by: Sapkota, Ranjan, et al.
Published: (2026)
by: Sapkota, Ranjan, et al.
Published: (2026)
Integrating YOLO11 and Convolution Block Attention Module for Multi-Season Segmentation of Tree Trunks and Branches in Commercial Apple Orchards
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
YOLO11 and Vision Transformers based 3D Pose Estimation of Immature Green Fruits in Commercial Apple Orchards for Robotic Thinning
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
YOLO26: Key Architectural Enhancements and Performance Benchmarking for Real-Time Object Detection
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Generalization vs. Specialization: Evaluating Segment Anything Model (SAM3) Zero-Shot Segmentation Against Fine-Tuned YOLO Detectors
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities
by: Sapkota, Ranjan, et al.
Published: (2023)
by: Sapkota, Ranjan, et al.
Published: (2023)
Comparing YOLOv11 and YOLOv8 for instance segmentation of occluded and non-occluded immature green fruits in complex orchard environment
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
Object Detection with Multimodal Large Vision-Language Models: An In-depth Review
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Improved YOLOv12 with LLM-Generated Synthetic Data for Enhanced Apple Detection and Benchmarking Against YOLOv11 and YOLOv10
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Comprehensive Performance Evaluation of YOLOv12, YOLO11, YOLOv10, YOLOv9 and YOLOv8 on Detecting and Counting Fruitlet in Complex Orchard Environments
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Comparing YOLOv8 and Mask R-CNN for instance segmentation in complex orchard environments
by: Sapkota, Ranjan, et al.
Published: (2023)
by: Sapkota, Ranjan, et al.
Published: (2023)
Zero-Shot Automatic Annotation and Instance Segmentation using LLM-Generated Datasets: Eliminating Field Imaging and Manual Annotation for Deep Learning Model Development
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
AI Agents vs. Agentic AI: A Conceptual Taxonomy, Applications and Challenges
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
The SAM2-to-SAM3 Gap in the Segment Anything Model Family: Why Prompt-Based Expertise Fails in Concept-Driven Image Segmentation
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Vibe Coding vs. Agentic Coding: Fundamentals and Practical Implications of Agentic AI
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
UAVs Meet Agentic AI: A Multidomain Survey of Autonomous Aerial Intelligence and Agentic UAVs
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
YOLO-MARL: You Only LLM Once for Multi-Agent Reinforcement Learning
by: Zhuang, Yuan, et al.
Published: (2024)
by: Zhuang, Yuan, et al.
Published: (2024)
TRiSM for Agentic AI: A Review of Trust, Risk, and Security Management in LLM-based Agentic Multi-Agent Systems
by: Raza, Shaina, et al.
Published: (2025)
by: Raza, Shaina, et al.
Published: (2025)
Immature Green Apple Detection and Sizing in Commercial Orchards using YOLOv8 and Shape Fitting Techniques
by: Sapkota, Ranjan, et al.
Published: (2023)
by: Sapkota, Ranjan, et al.
Published: (2023)
A Review of 3D Object Detection with Vision-Language Models
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
A Decade of You Only Look Once (YOLO) for Object Detection: A Review
by: Ramos, Leo Thomas, et al.
Published: (2025)
by: Ramos, Leo Thomas, et al.
Published: (2025)
Vision-Language-Action (VLA) Models: Concepts, Progress, Applications and Challenges
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
RF-DETR Object Detection vs YOLOv12 : A Study of Transformer-based and CNN-based Architectures for Single-Class and Multi-Class Greenfruit Detection in Complex Orchard Environments Under Label Ambiguity
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Surveying You Only Look Once (YOLO) Multispectral Object Detection Advancements, Applications And Challenges
by: Gallagher, James E., et al.
Published: (2024)
by: Gallagher, James E., et al.
Published: (2024)
A Layered Self-Supervised Knowledge Distillation Framework for Efficient Multimodal Learning on the Edge
by: Dahri, Tarique, et al.
Published: (2025)
by: Dahri, Tarique, et al.
Published: (2025)
Agentic AI with Orchestrator-Agent Trust: A Modular Visual Classification Framework with Trust-Aware Orchestration and RAG-Based Reasoning
by: Roumeliotis, Konstantinos I., et al.
Published: (2025)
by: Roumeliotis, Konstantinos I., et al.
Published: (2025)
You Only Look Once at Anytime (AnytimeYOLO): Analysis and Optimization of Early-Exits for Object-Detection
by: Kuhse, Daniel, et al.
Published: (2025)
by: Kuhse, Daniel, et al.
Published: (2025)
You Only Landmark Once: Lightweight U-Net Face Super Resolution with YOLO-World Landmark Heatmaps
by: Carraro, Riccardo, et al.
Published: (2026)
by: Carraro, Riccardo, et al.
Published: (2026)
RMP-YOLO: A Robust Motion Predictor for Partially Observable Scenarios even if You Only Look Once
by: Sun, Jiawei, et al.
Published: (2024)
by: Sun, Jiawei, et al.
Published: (2024)
3D Reconstruction and Information Fusion between Dormant and Canopy Seasons in Commercial Orchards Using Deep Learning and Fast GICP
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
YOLO-Former: YOLO Shakes Hand With ViT
by: Khoramdel, Javad, et al.
Published: (2024)
by: Khoramdel, Javad, et al.
Published: (2024)
Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact
by: Qureshi, Rizwan, et al.
Published: (2025)
by: Qureshi, Rizwan, et al.
Published: (2025)
Plant Disease Detection through Multimodal Large Language Models and Convolutional Neural Networks
by: Roumeliotis, Konstantinos I., et al.
Published: (2025)
by: Roumeliotis, Konstantinos I., et al.
Published: (2025)
Ultralytics YOLO
by: Jocher, Glenn, et al.
Published: (2026)
by: Jocher, Glenn, et al.
Published: (2026)
CA-YOLO: Cross Attention Empowered YOLO for Biomimetic Localization
by: Zhang, Zhen, et al.
Published: (2026)
by: Zhang, Zhen, et al.
Published: (2026)
FA ‐ YOLO : An Enhanced YOLO Network for Pedestrian and Vehicle Detection
by: Qiong Song, et al.
Published: (2026)
by: Qiong Song, et al.
Published: (2026)
Similar Items
-
Agricultural Object Detection with You Look Only Once (YOLO) Algorithm: A Bibliometric and Systematic Literature Review
by: Badgujar, Chetan M, et al.
Published: (2024) -
Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition
by: Sapkota, Ranjan, et al.
Published: (2025) -
YOLOE-26: Integrating YOLO26 with YOLOE for Real-Time Open-Vocabulary Instance Segmentation
by: Sapkota, Ranjan, et al.
Published: (2026) -
Integrating YOLO11 and Convolution Block Attention Module for Multi-Season Segmentation of Tree Trunks and Branches in Commercial Apple Orchards
by: Sapkota, Ranjan, et al.
Published: (2024) -
YOLO11 and Vision Transformers based 3D Pose Estimation of Immature Green Fruits in Commercial Apple Orchards for Robotic Thinning
by: Sapkota, Ranjan, et al.
Published: (2024)