Plant Disease Detection through Multimodal Large Language Models and Convolutional Neural Networks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Roumeliotis, Konstantinos I., Sapkota, Ranjan, Karkee, Manoj, Tselikas, Nikolaos D., Nasiopoulos, Dimitrios K. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generalization vs. Specialization: Evaluating Segment Anything Model (SAM3) Zero-Shot Segmentation Against Fine-Tuned YOLO Detectors
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
The SAM2-to-SAM3 Gap in the Segment Anything Model Family: Why Prompt-Based Expertise Fails in Concept-Driven Image Segmentation
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Vision-Language-Action (VLA) Models: Concepts, Progress, Applications and Challenges
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Agentic AI with Orchestrator-Agent Trust: A Modular Visual Classification Framework with Trust-Aware Orchestration and RAG-Based Reasoning
von: Roumeliotis, Konstantinos I., et al.
Veröffentlicht: (2025)
von: Roumeliotis, Konstantinos I., et al.
Veröffentlicht: (2025)
Object Detection with Multimodal Large Vision-Language Models: An In-depth Review
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
A Review of 3D Object Detection with Vision-Language Models
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Integrating YOLO11 and Convolution Block Attention Module for Multi-Season Segmentation of Tree Trunks and Branches in Commercial Apple Orchards
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2023)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2023)
Improved YOLOv12 with LLM-Generated Synthetic Data for Enhanced Apple Detection and Benchmarking Against YOLOv11 and YOLOv10
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
YOLO11 and Vision Transformers based 3D Pose Estimation of Immature Green Fruits in Commercial Apple Orchards for Robotic Thinning
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
Comparing YOLOv11 and YOLOv8 for instance segmentation of occluded and non-occluded immature green fruits in complex orchard environment
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
YOLOE-26: Integrating YOLO26 with YOLOE for Real-Time Open-Vocabulary Instance Segmentation
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2026)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2026)
Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
AI Agents vs. Agentic AI: A Conceptual Taxonomy, Applications and Challenges
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Vibe Coding vs. Agentic Coding: Fundamentals and Practical Implications of Agentic AI
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
UAVs Meet Agentic AI: A Multidomain Survey of Autonomous Aerial Intelligence and Agentic UAVs
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Comparing YOLOv8 and Mask R-CNN for instance segmentation in complex orchard environments
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2023)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2023)
Zero-Shot Automatic Annotation and Instance Segmentation using LLM-Generated Datasets: Eliminating Field Imaging and Manual Annotation for Deep Learning Model Development
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
Immature Green Apple Detection and Sizing in Commercial Orchards using YOLOv8 and Shape Fitting Techniques
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2023)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2023)
YOLO26: Key Architectural Enhancements and Performance Benchmarking for Real-Time Object Detection
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
RF-DETR Object Detection vs YOLOv12 : A Study of Transformer-based and CNN-based Architectures for Single-Class and Multi-Class Greenfruit Detection in Complex Orchard Environments Under Label Ambiguity
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Comprehensive Performance Evaluation of YOLOv12, YOLO11, YOLOv10, YOLOv9 and YOLOv8 on Detecting and Counting Fruitlet in Complex Orchard Environments
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
3D Reconstruction and Information Fusion between Dormant and Canopy Seasons in Commercial Orchards Using Deep Learning and Fast GICP
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Comprehensive Analysis of Transparency and Accessibility of ChatGPT, DeepSeek, And other SoTA Large Language Models
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)
Shallow- and Deep-fake Image Manipulation Localization Using Vision Mamba and Guided Graph Neural Network
von: Zhang, Junbin, et al.
Veröffentlicht: (2026)
von: Zhang, Junbin, et al.
Veröffentlicht: (2026)
AgRegNet: A Deep Regression Network for Flower and Fruit Density Estimation, Localization, and Counting in Orchards
von: Bhattarai, Uddhav, et al.
Veröffentlicht: (2024)
von: Bhattarai, Uddhav, et al.
Veröffentlicht: (2024)
A Light Weight Multi-Features-View Convolution Neural Network For Plant Disease Identification
von: Khan, Muhammad Kaleem Ullah
Veröffentlicht: (2026)
von: Khan, Muhammad Kaleem Ullah
Veröffentlicht: (2026)
TAME: Attention Mechanism Based Feature Fusion for Generating Explanation Maps of Convolutional Neural Networks
von: Ntrougkas, Mariano, et al.
Veröffentlicht: (2023)
von: Ntrougkas, Mariano, et al.
Veröffentlicht: (2023)
Metric as Transform: Exploring beyond Affine Transform for Interpretable Neural Network
von: Sapkota, Suman
Veröffentlicht: (2024)
von: Sapkota, Suman
Veröffentlicht: (2024)
YOLO advances to its genesis: a decadal and comprehensive review of the You Only Look Once (YOLO) series
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2024)
Weed Detection using Convolutional Neural Network
von: Tripathi, Santosh Kumar, et al.
Veröffentlicht: (2025)
von: Tripathi, Santosh Kumar, et al.
Veröffentlicht: (2025)
PND-Net: Plant Nutrition Deficiency and Disease Classification using Graph Convolutional Network
von: Bera, Asish, et al.
Veröffentlicht: (2024)
von: Bera, Asish, et al.
Veröffentlicht: (2024)
CLAP Convolutional Lightweight Autoencoder for Plant Disease Classification
von: Bera, Asish, et al.
Veröffentlicht: (2026)
von: Bera, Asish, et al.
Veröffentlicht: (2026)
Enhancing Road Safety: Real-Time Detection of Driver Distraction through Convolutional Neural Networks
von: Sheikh, Amaan Aijaz, et al.
Veröffentlicht: (2024)
von: Sheikh, Amaan Aijaz, et al.
Veröffentlicht: (2024)
Optimizing Violence Detection in Video Classification Accuracy through 3D Convolutional Neural Networks
von: Kavathia, Aarjav, et al.
Veröffentlicht: (2024)
von: Kavathia, Aarjav, et al.
Veröffentlicht: (2024)
TRiSM for Agentic AI: A Review of Trust, Risk, and Security Management in LLM-based Agentic Multi-Agent Systems
von: Raza, Shaina, et al.
Veröffentlicht: (2025)
von: Raza, Shaina, et al.
Veröffentlicht: (2025)
Object Detection Based on Distributed Convolutional Neural Networks
von: Sun, Liang
Veröffentlicht: (2026)
von: Sun, Liang
Veröffentlicht: (2026)
The Potential of Convolutional Neural Networks for Cancer Detection
von: Molaeian, Hossein, et al.
Veröffentlicht: (2024)
von: Molaeian, Hossein, et al.
Veröffentlicht: (2024)
Site-specific weed management in corn using UAS imagery analysis and computer vision techniques
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2022)
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Generalization vs. Specialization: Evaluating Segment Anything Model (SAM3) Zero-Shot Segmentation Against Fine-Tuned YOLO Detectors
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025) -
The SAM2-to-SAM3 Gap in the Segment Anything Model Family: Why Prompt-Based Expertise Fails in Concept-Driven Image Segmentation
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025) -
Vision-Language-Action (VLA) Models: Concepts, Progress, Applications and Challenges
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025) -
Agentic AI with Orchestrator-Agent Trust: A Modular Visual Classification Framework with Trust-Aware Orchestration and RAG-Based Reasoning
von: Roumeliotis, Konstantinos I., et al.
Veröffentlicht: (2025) -
Object Detection with Multimodal Large Vision-Language Models: An In-depth Review
von: Sapkota, Ranjan, et al.
Veröffentlicht: (2025)