YOLOv10-Based Multi-Task Framework for Hand Localization and Laterality Classification in Surgical Videos
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Sun, Kedi, Zhang, Le |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Hi-GaTA: Hierarchical Gated Temporal Aggregation Adapter for Surgical Video Report Generation
par: Sun, Kedi, et autres
Publié: (2026)
par: Sun, Kedi, et autres
Publié: (2026)
Temporally Guided Articulated Hand Pose Tracking in Surgical Videos
par: Louis, Nathan, et autres
Publié: (2021)
par: Louis, Nathan, et autres
Publié: (2021)
Comprehensive Performance Evaluation of YOLOv11, YOLOv10, YOLOv9, YOLOv8 and YOLOv5 on Object Detection of Power Equipment
par: He, Zijian, et autres
Publié: (2024)
par: He, Zijian, et autres
Publié: (2024)
YOLOv5, YOLOv8 and YOLOv10: The Go-To Detectors for Real-time Vision
par: Hussain, Muhammad
Publié: (2024)
par: Hussain, Muhammad
Publié: (2024)
A Comparative Analysis of YOLOv5, YOLOv8, and YOLOv10 in Kitchen Safety
par: Geetha, Athulya Sundaresan, et autres
Publié: (2024)
par: Geetha, Athulya Sundaresan, et autres
Publié: (2024)
Optimizing YOLO Architectures for Optimal Road Damage Detection and Classification: A Comparative Study from YOLOv7 to YOLOv10
par: Pham, Vung, et autres
Publié: (2024)
par: Pham, Vung, et autres
Publié: (2024)
SHANDS: A Multi-View Dataset and Benchmark for Surgical Hand-Gesture and Error Recognition Toward Medical Training
par: Ma, Le, et autres
Publié: (2026)
par: Ma, Le, et autres
Publié: (2026)
SurgFed: Language-guided Multi-Task Federated Learning for Surgical Video Understanding
par: Fang, Zheng, et autres
Publié: (2026)
par: Fang, Zheng, et autres
Publié: (2026)
Enhanced Self-Checkout System for Retail Based on Improved YOLOv10
par: Tan, Lianghao, et autres
Publié: (2024)
par: Tan, Lianghao, et autres
Publié: (2024)
YOLOv1 to YOLOv10: A comprehensive review of YOLO variants and their application in the agricultural domain
par: Alif, Mujadded Al Rabbani, et autres
Publié: (2024)
par: Alif, Mujadded Al Rabbani, et autres
Publié: (2024)
Comparative Analysis of YOLOv9, YOLOv10 and RT-DETR for Real-Time Weed Detection
par: Saltık, Ahmet Oğuz, et autres
Publié: (2024)
par: Saltık, Ahmet Oğuz, et autres
Publié: (2024)
YOLOv1 to YOLOv10: The fastest and most accurate real-time object detection systems
par: Wang, Chien-Yao, et autres
Publié: (2024)
par: Wang, Chien-Yao, et autres
Publié: (2024)
A Comparative Study of YOLOv8 to YOLOv11 Performance in Underwater Vision Tasks
par: Hung, Gordon, et autres
Publié: (2025)
par: Hung, Gordon, et autres
Publié: (2025)
Improved YOLOv12 with LLM-Generated Synthetic Data for Enhanced Apple Detection and Benchmarking Against YOLOv11 and YOLOv10
par: Sapkota, Ranjan, et autres
Publié: (2025)
par: Sapkota, Ranjan, et autres
Publié: (2025)
Lightweight Shrimp Disease Detection Research Based on YOLOv8n
par: Yuhuan, Fei, et autres
Publié: (2025)
par: Yuhuan, Fei, et autres
Publié: (2025)
Smart Parking with Pixel-Wise ROI Selection for Vehicle Detection Using YOLOv8, YOLOv9, YOLOv10, and YOLOv11
par: da Luz, Gustavo P. C. P., et autres
Publié: (2024)
par: da Luz, Gustavo P. C. P., et autres
Publié: (2024)
YOLOv8-AM: YOLOv8 Based on Effective Attention Mechanisms for Pediatric Wrist Fracture Detection
par: Chien, Chun-Tse, et autres
Publié: (2024)
par: Chien, Chun-Tse, et autres
Publié: (2024)
Instrument-tissue Interaction Detection Framework for Surgical Video Understanding
par: Lin, Wenjun, et autres
Publié: (2024)
par: Lin, Wenjun, et autres
Publié: (2024)
YOLOv8-ResCBAM: YOLOv8 Based on An Effective Attention Module for Pediatric Wrist Fracture Detection
par: Ju, Rui-Yang, et autres
Publié: (2024)
par: Ju, Rui-Yang, et autres
Publié: (2024)
hYOLO Model: Enhancing Object Classification with Hierarchical Context in YOLOv8
par: Tsenkova, Veska, et autres
Publié: (2025)
par: Tsenkova, Veska, et autres
Publié: (2025)
Vehicle Detection and Classification for Toll collection using YOLOv11 and Ensemble OCR
par: Sivakoti, Karthik
Publié: (2024)
par: Sivakoti, Karthik
Publié: (2024)
EGFormer: Towards Efficient and Generalizable Multimodal Semantic Segmentation
par: Zhang, Zelin, et autres
Publié: (2025)
par: Zhang, Zelin, et autres
Publié: (2025)
Comprehensive Performance Evaluation of YOLOv12, YOLO11, YOLOv10, YOLOv9 and YOLOv8 on Detecting and Counting Fruitlet in Complex Orchard Environments
par: Sapkota, Ranjan, et autres
Publié: (2024)
par: Sapkota, Ranjan, et autres
Publié: (2024)
Seeing More with Less: Video Capsule Endoscopy with Multi-Task Learning
par: Werner, Julia, et autres
Publié: (2025)
par: Werner, Julia, et autres
Publié: (2025)
Novel Human Machine Interface via Robust Hand Gesture Recognition System using Channel Pruned YOLOv5s Model
par: Sen, Abir, et autres
Publié: (2024)
par: Sen, Abir, et autres
Publié: (2024)
YOLOv10: Real-Time End-to-End Object Detection
par: Wang, Ao, et autres
Publié: (2024)
par: Wang, Ao, et autres
Publié: (2024)
ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping
par: Pang, Youxin, et autres
Publié: (2024)
par: Pang, Youxin, et autres
Publié: (2024)
Local Spherical Harmonics Improve Skeleton-Based Hand Action Recognition
par: Prasse, Katharina, et autres
Publié: (2023)
par: Prasse, Katharina, et autres
Publié: (2023)
Facial Expression Recognition with YOLOv11 and YOLOv12: A Comparative Study
par: Aymon, Umma, et autres
Publié: (2025)
par: Aymon, Umma, et autres
Publié: (2025)
Mask-to-Height: A YOLOv11-Based Architecture for Joint Building Instance Segmentation and Height Classification from Satellite Imagery
par: Hussieni, Mahmoud El, et autres
Publié: (2025)
par: Hussieni, Mahmoud El, et autres
Publié: (2025)
Scaling Video Pretraining for Surgical Foundation Models
par: Lu, Sicheng, et autres
Publié: (2026)
par: Lu, Sicheng, et autres
Publié: (2026)
SurgiTrack: Fine-Grained Multi-Class Multi-Tool Tracking in Surgical Videos
par: Nwoye, Chinedu Innocent, et autres
Publié: (2024)
par: Nwoye, Chinedu Innocent, et autres
Publié: (2024)
mmWalk: Towards Multi-modal Multi-view Walking Assistance
par: Ying, Kedi, et autres
Publié: (2025)
par: Ying, Kedi, et autres
Publié: (2025)
Learning to Localize Actions in Instructional Videos with LLM-Based Multi-Pathway Text-Video Alignment
par: Chen, Yuxiao, et autres
Publié: (2024)
par: Chen, Yuxiao, et autres
Publié: (2024)
Multi-Class Abnormality Classification Task in Video Capsule Endoscopy
par: Verma, Dev Rishi, et autres
Publié: (2024)
par: Verma, Dev Rishi, et autres
Publié: (2024)
Automating Coral Reef Fish Family Identification on Video Transects Using a YOLOv8-Based Deep Learning Pipeline
par: Gerard, Jules, et autres
Publié: (2025)
par: Gerard, Jules, et autres
Publié: (2025)
SOD-YOLOv8 -- Enhancing YOLOv8 for Small Object Detection in Traffic Scenes
par: Khalili, Boshra, et autres
Publié: (2024)
par: Khalili, Boshra, et autres
Publié: (2024)
UnityVideo: Unified Multi-Modal Multi-Task Learning for Enhancing World-Aware Video Generation
par: Huang, Jiehui, et autres
Publié: (2025)
par: Huang, Jiehui, et autres
Publié: (2025)
Multi-Granularity Hand Action Detection
par: Zhe, Ting, et autres
Publié: (2023)
par: Zhe, Ting, et autres
Publié: (2023)
Recognizing Hand Use and Hand Role at Home After Stroke from Egocentric Video
par: Tsai, Meng-Fen, et autres
Publié: (2022)
par: Tsai, Meng-Fen, et autres
Publié: (2022)
Documents similaires
-
Hi-GaTA: Hierarchical Gated Temporal Aggregation Adapter for Surgical Video Report Generation
par: Sun, Kedi, et autres
Publié: (2026) -
Temporally Guided Articulated Hand Pose Tracking in Surgical Videos
par: Louis, Nathan, et autres
Publié: (2021) -
Comprehensive Performance Evaluation of YOLOv11, YOLOv10, YOLOv9, YOLOv8 and YOLOv5 on Object Detection of Power Equipment
par: He, Zijian, et autres
Publié: (2024) -
YOLOv5, YOLOv8 and YOLOv10: The Go-To Detectors for Real-time Vision
par: Hussain, Muhammad
Publié: (2024) -
A Comparative Analysis of YOLOv5, YOLOv8, and YOLOv10 in Kitchen Safety
par: Geetha, Athulya Sundaresan, et autres
Publié: (2024)