YOLO11 and Vision Transformers based 3D Pose Estimation of Immature Green Fruits in Commercial Apple Orchards for Robotic Thinning
Fuente:
arXiv
Saved in:
| Main Authors: | Sapkota, Ranjan, Karkee, Manoj |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Integrating YOLO11 and Convolution Block Attention Module for Multi-Season Segmentation of Tree Trunks and Branches in Commercial Apple Orchards
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
Immature Green Apple Detection and Sizing in Commercial Orchards using YOLOv8 and Shape Fitting Techniques
by: Sapkota, Ranjan, et al.
Published: (2023)
by: Sapkota, Ranjan, et al.
Published: (2023)
Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Improved YOLOv12 with LLM-Generated Synthetic Data for Enhanced Apple Detection and Benchmarking Against YOLOv11 and YOLOv10
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
YOLOE-26: Integrating YOLO26 with YOLOE for Real-Time Open-Vocabulary Instance Segmentation
by: Sapkota, Ranjan, et al.
Published: (2026)
by: Sapkota, Ranjan, et al.
Published: (2026)
3D Reconstruction and Information Fusion between Dormant and Canopy Seasons in Commercial Orchards Using Deep Learning and Fast GICP
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Comparing YOLOv11 and YOLOv8 for instance segmentation of occluded and non-occluded immature green fruits in complex orchard environment
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
Object Detection with Multimodal Large Vision-Language Models: An In-depth Review
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
RF-DETR Object Detection vs YOLOv12 : A Study of Transformer-based and CNN-based Architectures for Single-Class and Multi-Class Greenfruit Detection in Complex Orchard Environments Under Label Ambiguity
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Comprehensive Performance Evaluation of YOLOv12, YOLO11, YOLOv10, YOLOv9 and YOLOv8 on Detecting and Counting Fruitlet in Complex Orchard Environments
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities
by: Sapkota, Ranjan, et al.
Published: (2023)
by: Sapkota, Ranjan, et al.
Published: (2023)
YOLO26: Key Architectural Enhancements and Performance Benchmarking for Real-Time Object Detection
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Generalization vs. Specialization: Evaluating Segment Anything Model (SAM3) Zero-Shot Segmentation Against Fine-Tuned YOLO Detectors
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Comparing YOLOv8 and Mask R-CNN for instance segmentation in complex orchard environments
by: Sapkota, Ranjan, et al.
Published: (2023)
by: Sapkota, Ranjan, et al.
Published: (2023)
Vision-Language-Action (VLA) Models: Concepts, Progress, Applications and Challenges
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
AgRegNet: A Deep Regression Network for Flower and Fruit Density Estimation, Localization, and Counting in Orchards
by: Bhattarai, Uddhav, et al.
Published: (2024)
by: Bhattarai, Uddhav, et al.
Published: (2024)
Zero-Shot Automatic Annotation and Instance Segmentation using LLM-Generated Datasets: Eliminating Field Imaging and Manual Annotation for Deep Learning Model Development
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
The SAM2-to-SAM3 Gap in the Segment Anything Model Family: Why Prompt-Based Expertise Fails in Concept-Driven Image Segmentation
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
A Review of 3D Object Detection with Vision-Language Models
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
Multimodal Large Language Models for Image, Text, and Speech Data Augmentation: A Survey
by: Sapkota, Ranjan, et al.
Published: (2025)
by: Sapkota, Ranjan, et al.
Published: (2025)
YOLO advances to its genesis: a decadal and comprehensive review of the You Only Look Once (YOLO) series
by: Sapkota, Ranjan, et al.
Published: (2024)
by: Sapkota, Ranjan, et al.
Published: (2024)
Plant Disease Detection through Multimodal Large Language Models and Convolutional Neural Networks
by: Roumeliotis, Konstantinos I., et al.
Published: (2025)
by: Roumeliotis, Konstantinos I., et al.
Published: (2025)
S$^3$AD: Semi-supervised Small Apple Detection in Orchard Environments
by: Johanson, Robert, et al.
Published: (2023)
by: Johanson, Robert, et al.
Published: (2023)
GDA-YOLO11: Amodal Instance Segmentation for Occlusion-Robust Robotic Fruit Harvesting
by: Beldek, Caner, et al.
Published: (2026)
by: Beldek, Caner, et al.
Published: (2026)
Waterfall Transformer for Multi-person Pose Estimation
by: Ranjan, Navin, et al.
Published: (2024)
by: Ranjan, Navin, et al.
Published: (2024)
Precise Apple Detection and Localization in Orchards using YOLOv5 for Robotic Harvesting Systems
by: Ziyue, Jiang, et al.
Published: (2024)
by: Ziyue, Jiang, et al.
Published: (2024)
A Comparative Study of Modern Object Detectors for Robust Apple Detection in Orchard Imagery
by: Asad, Mohammed, et al.
Published: (2026)
by: Asad, Mohammed, et al.
Published: (2026)
Robotic Pollination of Apples in Commercial Orchards
by: Sapkota, Ranjan, et al.
Published: (2023)
by: Sapkota, Ranjan, et al.
Published: (2023)
FastPose-ViT: A Vision Transformer for Real-Time Spacecraft Pose Estimation
by: Ancey, Pierre, et al.
Published: (2025)
by: Ancey, Pierre, et al.
Published: (2025)
OrchardDepth: Precise Metric Depth Estimation of Orchard Scene from Monocular Camera Images
by: Zheng, Zhichao, et al.
Published: (2025)
by: Zheng, Zhichao, et al.
Published: (2025)
Spectral Compression Transformer with Line Pose Graph for Monocular 3D Human Pose Estimation
by: Zheng, Zenghao, et al.
Published: (2025)
by: Zheng, Zenghao, et al.
Published: (2025)
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer
by: Lin, Xiao, et al.
Published: (2023)
by: Lin, Xiao, et al.
Published: (2023)
GS-Pose: Generalizable Segmentation-based 6D Object Pose Estimation with 3D Gaussian Splatting
by: Cai, Dingding, et al.
Published: (2024)
by: Cai, Dingding, et al.
Published: (2024)
SE(3)-PoseFlow: Estimating 6D Pose Distributions for Uncertainty-Aware Robotic Manipulation
by: Jin, Yufeng, et al.
Published: (2025)
by: Jin, Yufeng, et al.
Published: (2025)
Diffusion Model is a Good Pose Estimator from 3D RF-Vision
by: Fan, Junqiao, et al.
Published: (2024)
by: Fan, Junqiao, et al.
Published: (2024)
MVTOP: Multi-View Transformer-based Object Pose-Estimation
by: Ranftl, Lukas, et al.
Published: (2025)
by: Ranftl, Lukas, et al.
Published: (2025)
SkelFormer: Markerless 3D Pose and Shape Estimation using Skeletal Transformers
by: Davoodnia, Vandad, et al.
Published: (2024)
by: Davoodnia, Vandad, et al.
Published: (2024)
Object Pose Transformer: Unifying Unseen Object Pose Estimation
by: Li, Weihang, et al.
Published: (2026)
by: Li, Weihang, et al.
Published: (2026)
JRDB-Pose3D: A Multi-person 3D Human Pose and Shape Estimation Dataset for Robotics
by: Biswas, Sandika, et al.
Published: (2026)
by: Biswas, Sandika, et al.
Published: (2026)
Beyond Static Frames: Temporal Aggregate-and-Restore Vision Transformer for Human Pose Estimation
by: Fang, Hongwei, et al.
Published: (2026)
by: Fang, Hongwei, et al.
Published: (2026)
Similar Items
-
Integrating YOLO11 and Convolution Block Attention Module for Multi-Season Segmentation of Tree Trunks and Branches in Commercial Apple Orchards
by: Sapkota, Ranjan, et al.
Published: (2024) -
Immature Green Apple Detection and Sizing in Commercial Orchards using YOLOv8 and Shape Fitting Techniques
by: Sapkota, Ranjan, et al.
Published: (2023) -
Ultralytics YOLO Evolution: An Overview of YOLO26, YOLO11, YOLOv8 and YOLOv5 Object Detectors for Computer Vision and Pattern Recognition
by: Sapkota, Ranjan, et al.
Published: (2025) -
Improved YOLOv12 with LLM-Generated Synthetic Data for Enhanced Apple Detection and Benchmarking Against YOLOv11 and YOLOv10
by: Sapkota, Ranjan, et al.
Published: (2025) -
YOLOE-26: Integrating YOLO26 with YOLOE for Real-Time Open-Vocabulary Instance Segmentation
by: Sapkota, Ranjan, et al.
Published: (2026)