EHWGesture -- A dataset for multimodal understanding of clinical gestures
Fuente:
arXiv
Saved in:
| Main Authors: | Amprimo, Gianluca, Ancilotto, Alberto, Savino, Alessandro, Quazzolo, Fabio, Ferraris, Claudia, Olmo, Gabriella, Farella, Elisabetta, Di Carlo, Stefano |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Elastic Spiking Transformers for Efficient Gesture Understanding
by: Ancilotto, Alberto, et al.
Published: (2026)
by: Ancilotto, Alberto, et al.
Published: (2026)
A multimodal gesture recognition dataset for desktop human-computer interaction
by: Wang, Qi, et al.
Published: (2024)
by: Wang, Qi, et al.
Published: (2024)
Digital Monitoring of Motor Function in Parkinson's Disease Using Markerless Motion Analysis and Exergaming
by: Ferraris, Claudia, et al.
Published: (2026)
by: Ferraris, Claudia, et al.
Published: (2026)
A probabilistic framework for dynamic quantization
by: Santini, Gabriele, et al.
Published: (2025)
by: Santini, Gabriele, et al.
Published: (2025)
SFATTI: Spiking FPGA Accelerator for Temporal Task-driven Inference -- A Case Study on MNIST
by: Caviglia, Alessio, et al.
Published: (2025)
by: Caviglia, Alessio, et al.
Published: (2025)
R-CONV: An Analytical Approach for Efficient Data Reconstruction via Convolutional Gradients
by: Eltaras, Tamer Ahmed, et al.
Published: (2024)
by: Eltaras, Tamer Ahmed, et al.
Published: (2024)
A multimodal dataset for understanding the impact of mobile phones on remote online virtual education
by: Daza, Roberto, et al.
Published: (2024)
by: Daza, Roberto, et al.
Published: (2024)
Prototype Learning for Micro-gesture Classification
by: Chen, Guoliang, et al.
Published: (2024)
by: Chen, Guoliang, et al.
Published: (2024)
Investigating the impact of 2D gesture representation on co-speech gesture generation
by: Guichoux, Teo, et al.
Published: (2024)
by: Guichoux, Teo, et al.
Published: (2024)
MULTIAQUA: A multimodal maritime dataset and robust training strategies for multimodal semantic segmentation
by: Muhovič, Jon, et al.
Published: (2025)
by: Muhovič, Jon, et al.
Published: (2025)
Research on gesture recognition method based on SEDCNN-SVM
by: Zhang, Mingjin, et al.
Published: (2024)
by: Zhang, Mingjin, et al.
Published: (2024)
Headset: Human emotion awareness under partial occlusions multimodal dataset
by: Lohesara, Fatemeh Ghorbani, et al.
Published: (2024)
by: Lohesara, Fatemeh Ghorbani, et al.
Published: (2024)
PaSTe: Improving the Efficiency of Visual Anomaly Detection at the Edge
by: Barusco, Manuel, et al.
Published: (2024)
by: Barusco, Manuel, et al.
Published: (2024)
Latent Distillation for Continual Object Detection at the Edge
by: Pasti, Francesco, et al.
Published: (2024)
by: Pasti, Francesco, et al.
Published: (2024)
ArtSeek: Deep artwork understanding via multimodal in-context reasoning and late interaction retrieval
by: Fanelli, Nicola, et al.
Published: (2025)
by: Fanelli, Nicola, et al.
Published: (2025)
CheXthought: A global multimodal dataset of clinical chain-of-thought reasoning and visual attention for chest X-ray interpretation
by: Sharma, Sonali, et al.
Published: (2026)
by: Sharma, Sonali, et al.
Published: (2026)
Micro-gesture Online Recognition using Learnable Query Points
by: Liu, Pengyu, et al.
Published: (2024)
by: Liu, Pengyu, et al.
Published: (2024)
Interpreting Hand gestures using Object Detection and Digits Classification
by: K, Sangeetha, et al.
Published: (2024)
by: K, Sangeetha, et al.
Published: (2024)
Bias-constrained multimodal intelligence for equitable and reliable clinical AI
by: Li, Cheng, et al.
Published: (2026)
by: Li, Cheng, et al.
Published: (2026)
MObyGaze: a film dataset of multimodal objectification densely annotated by experts
by: Tores, Julie, et al.
Published: (2025)
by: Tores, Julie, et al.
Published: (2025)
GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding
by: Wu, Yiqi, et al.
Published: (2024)
by: Wu, Yiqi, et al.
Published: (2024)
From Vision to Sound: Advancing Audio Anomaly Detection with Vision-Based Algorithms
by: Barusco, Manuel, et al.
Published: (2025)
by: Barusco, Manuel, et al.
Published: (2025)
SideSeeing: A multimodal dataset and collection of tools for sidewalk assessment
by: Damaceno, R. J. P., et al.
Published: (2024)
by: Damaceno, R. J. P., et al.
Published: (2024)
Time Distributed Deep Learning Models for Purely Exogenous Forecasting: Application to Water Table Depth Prediction using Weather Image Time Series
by: Salis, Matteo, et al.
Published: (2024)
by: Salis, Matteo, et al.
Published: (2024)
Hybrid-supervised Hypergraph-enhanced Transformer for Micro-gesture Based Emotion Recognition
by: Xia, Zhaoqiang, et al.
Published: (2025)
by: Xia, Zhaoqiang, et al.
Published: (2025)
Online Micro-gesture Recognition Using Data Augmentation and Spatial-Temporal Attention
by: Liu, Pengyu, et al.
Published: (2025)
by: Liu, Pengyu, et al.
Published: (2025)
Online hand gesture recognition using Continual Graph Transformers
by: Slama, Rim, et al.
Published: (2025)
by: Slama, Rim, et al.
Published: (2025)
Replay Consolidation with Label Propagation for Continual Object Detection
by: De Monte, Riccardo, et al.
Published: (2024)
by: De Monte, Riccardo, et al.
Published: (2024)
A benchmark multimodal oro-dental dataset for large vision-language models
by: Lv, Haoxin, et al.
Published: (2025)
by: Lv, Haoxin, et al.
Published: (2025)
Functionality understanding and segmentation in 3D scenes
by: Corsetti, Jaime, et al.
Published: (2024)
by: Corsetti, Jaime, et al.
Published: (2024)
MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation
by: Xu, Lijian, et al.
Published: (2024)
by: Xu, Lijian, et al.
Published: (2024)
Deep self-supervised learning with visualisation for automatic gesture recognition
by: Allemand, Fabien, et al.
Published: (2024)
by: Allemand, Fabien, et al.
Published: (2024)
A comprehensive multimodal dataset and benchmark for ulcerative colitis scoring in endoscopy
by: Ghatwary, Noha, et al.
Published: (2026)
by: Ghatwary, Noha, et al.
Published: (2026)
AquaMonitor: A multimodal multi-view image sequence dataset for real-life aquatic invertebrate biodiversity monitoring
by: Impiö, Mikko, et al.
Published: (2025)
by: Impiö, Mikko, et al.
Published: (2025)
Zero123-6D: Zero-shot Novel View Synthesis for RGB Category-level 6D Pose Estimation
by: Di Felice, Francesco, et al.
Published: (2024)
by: Di Felice, Francesco, et al.
Published: (2024)
MAN TruckScenes: A multimodal dataset for autonomous trucking in diverse conditions
by: Fent, Felix, et al.
Published: (2024)
by: Fent, Felix, et al.
Published: (2024)
Planktonzilla: Multimodal dataset and models for understanding plankton ecosystems
by: Montanares, Alan Gerson Contreras, et al.
Published: (2026)
by: Montanares, Alan Gerson Contreras, et al.
Published: (2026)
OmDet: Large-scale vision-language multi-dataset pre-training with multimodal detection network
by: Zhao, Tiancheng, et al.
Published: (2022)
by: Zhao, Tiancheng, et al.
Published: (2022)
Attacks on multimodal models
by: Iablochnikov, Viacheslav, et al.
Published: (2024)
by: Iablochnikov, Viacheslav, et al.
Published: (2024)
HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos
by: Peirone, Simone Alberto, et al.
Published: (2025)
by: Peirone, Simone Alberto, et al.
Published: (2025)
Similar Items
-
Elastic Spiking Transformers for Efficient Gesture Understanding
by: Ancilotto, Alberto, et al.
Published: (2026) -
A multimodal gesture recognition dataset for desktop human-computer interaction
by: Wang, Qi, et al.
Published: (2024) -
Digital Monitoring of Motor Function in Parkinson's Disease Using Markerless Motion Analysis and Exergaming
by: Ferraris, Claudia, et al.
Published: (2026) -
A probabilistic framework for dynamic quantization
by: Santini, Gabriele, et al.
Published: (2025) -
SFATTI: Spiking FPGA Accelerator for Temporal Task-driven Inference -- A Case Study on MNIST
by: Caviglia, Alessio, et al.
Published: (2025)