Multi-Modal Gesture Recognition from Video and Surgical Tool Pose Information via Motion Invariants
Fuente:
arXiv
Salvato in:
| Autori principali: | Atoum, Jumanh, Johnston, Garrison L. H., Simaan, Nabil, Wu, Jie Ying |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Zero-shot Prompt-based Video Encoder for Surgical Gesture Recognition
di: Rao, Mingxing, et al.
Pubblicazione: (2024)
di: Rao, Mingxing, et al.
Pubblicazione: (2024)
Fluoroscopic Shape and Pose Tracking of Catheters with Custom Radiopaque Markers
di: Lawson, Jared, et al.
Pubblicazione: (2025)
di: Lawson, Jared, et al.
Pubblicazione: (2025)
Task and Configuration Space Compliance of Continuum Robots via Lie Group and Modal Shape Formulations
di: Orekhov, Andrew L., et al.
Pubblicazione: (2023)
di: Orekhov, Andrew L., et al.
Pubblicazione: (2023)
A Modal-Space Formulation for Momentum Observer Contact Estimation and Effects of Uncertainty for Continuum Robots
di: Johnston, Garrison L. H., et al.
Pubblicazione: (2025)
di: Johnston, Garrison L. H., et al.
Pubblicazione: (2025)
SurgFormer: Scalable Learning of Organ Deformation with Resection Support and Real-Time Inference
di: Shahbazi, Ashkan, et al.
Pubblicazione: (2026)
di: Shahbazi, Ashkan, et al.
Pubblicazione: (2026)
Multi-view Video-Pose Pretraining for Operating Room Surgical Activity Recognition
di: Hamoud, Idris, et al.
Pubblicazione: (2025)
di: Hamoud, Idris, et al.
Pubblicazione: (2025)
Design Considerations and Robustness to Parameter Uncertainty in Wire-Wrapped Cam Mechanisms
di: Johnston, Garrison L. H., et al.
Pubblicazione: (2023)
di: Johnston, Garrison L. H., et al.
Pubblicazione: (2023)
Neural-Augmented Kelvinlet for Real-Time Soft Tissue Deformation Modeling
di: Shahbazi, Ashkan, et al.
Pubblicazione: (2025)
di: Shahbazi, Ashkan, et al.
Pubblicazione: (2025)
Design Considerations for 3RRR Parallel Robots with Lightweight, Approximate Static-Balancing
di: Del Giudice, Giuseppe, et al.
Pubblicazione: (2023)
di: Del Giudice, Giuseppe, et al.
Pubblicazione: (2023)
Focus on the Experts: Co-designing an Augmented Reality Eye-Gaze Tracking System with Surgical Trainees to Improve Endoscopic Instruction
di: Atoum, Jumanh, et al.
Pubblicazione: (2025)
di: Atoum, Jumanh, et al.
Pubblicazione: (2025)
SurgPose: a Dataset for Articulated Robotic Surgical Tool Pose Estimation and Tracking
di: Wu, Zijian, et al.
Pubblicazione: (2025)
di: Wu, Zijian, et al.
Pubblicazione: (2025)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
di: Song, Yafei, et al.
Pubblicazione: (2025)
di: Song, Yafei, et al.
Pubblicazione: (2025)
SHANDS: A Multi-View Dataset and Benchmark for Surgical Hand-Gesture and Error Recognition Toward Medical Training
di: Ma, Le, et al.
Pubblicazione: (2026)
di: Ma, Le, et al.
Pubblicazione: (2026)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
di: He, Xu, et al.
Pubblicazione: (2024)
di: He, Xu, et al.
Pubblicazione: (2024)
Dual Conditioned Motion Diffusion for Pose-Based Video Anomaly Detection
di: Wang, Hongsong, et al.
Pubblicazione: (2024)
di: Wang, Hongsong, et al.
Pubblicazione: (2024)
Joint-Motion Mutual Learning for Pose Estimation in Videos
di: Wu, Sifan, et al.
Pubblicazione: (2024)
di: Wu, Sifan, et al.
Pubblicazione: (2024)
A Feasibility Study of a Soft, Low-Cost, 6-Axis Load Cell for Haptics
di: Veliky, Madison, et al.
Pubblicazione: (2024)
di: Veliky, Madison, et al.
Pubblicazione: (2024)
MeViS: A Multi-Modal Dataset for Referring Motion Expression Video Segmentation
di: Ding, Henghui, et al.
Pubblicazione: (2025)
di: Ding, Henghui, et al.
Pubblicazione: (2025)
Gesture Matters: Pedestrian Gesture Recognition for AVs Through Skeleton Pose Evaluation
di: Mahdi, Alif Rizqullah, et al.
Pubblicazione: (2026)
di: Mahdi, Alif Rizqullah, et al.
Pubblicazione: (2026)
Adaptive Physical-Facial Representation Fusion via Subject-Invariant Cross-Modal Prompt Tuning for Video-Based Emotion Recognition
di: Luo, Xiwen, et al.
Pubblicazione: (2026)
di: Luo, Xiwen, et al.
Pubblicazione: (2026)
Recognition of Daily Activities through Multi-Modal Deep Learning: A Video, Pose, and Object-Aware Approach for Ambient Assisted Living
di: Hashemifard, Kooshan, et al.
Pubblicazione: (2026)
di: Hashemifard, Kooshan, et al.
Pubblicazione: (2026)
Efficient Surgical Tool Recognition via HMM-Stabilized Deep Learning
di: Wang, Haifeng, et al.
Pubblicazione: (2024)
di: Wang, Haifeng, et al.
Pubblicazione: (2024)
Multi-task Learning For Joint Action and Gesture Recognition
di: Spathis, Konstantinos, et al.
Pubblicazione: (2025)
di: Spathis, Konstantinos, et al.
Pubblicazione: (2025)
SurgiTrack: Fine-Grained Multi-Class Multi-Tool Tracking in Surgical Videos
di: Nwoye, Chinedu Innocent, et al.
Pubblicazione: (2024)
di: Nwoye, Chinedu Innocent, et al.
Pubblicazione: (2024)
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
di: Gu, Jihao, et al.
Pubblicazione: (2025)
di: Gu, Jihao, et al.
Pubblicazione: (2025)
Boosting Gesture Recognition with an Automatic Gesture Annotation Framework
di: Shen, Junxiao, et al.
Pubblicazione: (2024)
di: Shen, Junxiao, et al.
Pubblicazione: (2024)
Temporally Guided Articulated Hand Pose Tracking in Surgical Videos
di: Louis, Nathan, et al.
Pubblicazione: (2021)
di: Louis, Nathan, et al.
Pubblicazione: (2021)
FG-SGL: Fine-Grained Semantic Guidance Learning via Motion Process Decomposition for Micro-Gesture Recognition
di: Wei, Jinsheng, et al.
Pubblicazione: (2026)
di: Wei, Jinsheng, et al.
Pubblicazione: (2026)
A Methodological and Structural Review of Hand Gesture Recognition Across Diverse Data Modalities
di: Shin, Jungpil, et al.
Pubblicazione: (2024)
di: Shin, Jungpil, et al.
Pubblicazione: (2024)
MSF-Mamba: Motion-aware State Fusion Mamba for Efficient Micro-Gesture Recognition
di: Li, Deng, et al.
Pubblicazione: (2025)
di: Li, Deng, et al.
Pubblicazione: (2025)
Large Motion Model for Unified Multi-Modal Motion Generation
di: Zhang, Mingyuan, et al.
Pubblicazione: (2024)
di: Zhang, Mingyuan, et al.
Pubblicazione: (2024)
Towards Large-Scale Pose-Invariant Face Recognition Using Face Defrontalization
di: Mesec, Patrik, et al.
Pubblicazione: (2025)
di: Mesec, Patrik, et al.
Pubblicazione: (2025)
End to End AI System for Surgical Gesture Sequence Recognition and Clinical Outcome Prediction
di: Li, Xi, et al.
Pubblicazione: (2025)
di: Li, Xi, et al.
Pubblicazione: (2025)
MultiMotion: Multi Subject Video Motion Transfer via Video Diffusion Transformer
di: Liu, Penghui, et al.
Pubblicazione: (2025)
di: Liu, Penghui, et al.
Pubblicazione: (2025)
EndoPBR: Material and Lighting Estimation for Photorealistic Surgical Simulations via Physically-based Rendering
di: Han, John J., et al.
Pubblicazione: (2025)
di: Han, John J., et al.
Pubblicazione: (2025)
Multiscaled Multi-Head Attention-based Video Transformer Network for Hand Gesture Recognition
di: Garg, Mallika, et al.
Pubblicazione: (2025)
di: Garg, Mallika, et al.
Pubblicazione: (2025)
An Evaluation of Large Pre-Trained Models for Gesture Recognition using Synthetic Videos
di: Reddy, Arun, et al.
Pubblicazione: (2024)
di: Reddy, Arun, et al.
Pubblicazione: (2024)
Action Recognition with Multi-stream Motion Modeling and Mutual Information Maximization
di: Yang, Yuheng, et al.
Pubblicazione: (2023)
di: Yang, Yuheng, et al.
Pubblicazione: (2023)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
Follow Your Pose: Pose-Guided Text-to-Video Generation using Pose-Free Videos
di: Ma, Yue, et al.
Pubblicazione: (2023)
di: Ma, Yue, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Zero-shot Prompt-based Video Encoder for Surgical Gesture Recognition
di: Rao, Mingxing, et al.
Pubblicazione: (2024) -
Fluoroscopic Shape and Pose Tracking of Catheters with Custom Radiopaque Markers
di: Lawson, Jared, et al.
Pubblicazione: (2025) -
Task and Configuration Space Compliance of Continuum Robots via Lie Group and Modal Shape Formulations
di: Orekhov, Andrew L., et al.
Pubblicazione: (2023) -
A Modal-Space Formulation for Momentum Observer Contact Estimation and Effects of Uncertainty for Continuum Robots
di: Johnston, Garrison L. H., et al.
Pubblicazione: (2025) -
SurgFormer: Scalable Learning of Organ Deformation with Resection Support and Real-Time Inference
di: Shahbazi, Ashkan, et al.
Pubblicazione: (2026)