Multi-task Learning For Joint Action and Gesture Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Spathis, Konstantinos, Kardaris, Nikolaos, Maragos, Petros |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pre-training for Action Recognition with Automatically Generated Fractal Datasets
by: Svyezhentsev, Davyd, et al.
Published: (2024)
by: Svyezhentsev, Davyd, et al.
Published: (2024)
TropNNC: Structured Neural Network Compression Using Tropical Geometry
by: Fotopoulos, Konstantinos, et al.
Published: (2024)
by: Fotopoulos, Konstantinos, et al.
Published: (2024)
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning
by: Giannakakis, Nikos, et al.
Published: (2025)
by: Giannakakis, Nikos, et al.
Published: (2025)
Optimal Transport for Handwritten Text Recognition in a Low-Resource Regime
by: Wraight, Petros Georgoulas, et al.
Published: (2025)
by: Wraight, Petros Georgoulas, et al.
Published: (2025)
Age-Inclusive 3D Human Mesh Recovery for Action-Preserving Data Anonymization
by: Chatzichristodoulou, Georgios, et al.
Published: (2025)
by: Chatzichristodoulou, Georgios, et al.
Published: (2025)
Mushroom Segmentation and 3D Pose Estimation from Point Clouds using Fully Convolutional Geometric Features and Implicit Pose Encoding
by: Retsinas, George, et al.
Published: (2024)
by: Retsinas, George, et al.
Published: (2024)
Only-Style: Stylistic Consistency in Image Generation without Content Leakage
by: Aravanis, Tilemachos, et al.
Published: (2025)
by: Aravanis, Tilemachos, et al.
Published: (2025)
Category-Level 6D Object Pose Estimation in Agricultural Settings Using a Lattice-Deformation Framework and Diffusion-Augmented Synthetic Data
by: Glytsos, Marios, et al.
Published: (2025)
by: Glytsos, Marios, et al.
Published: (2025)
GHR-VQA: Graph-guided Hierarchical Relational Reasoning for Video Question Answering
by: Brilli, Dionysia Danai, et al.
Published: (2025)
by: Brilli, Dionysia Danai, et al.
Published: (2025)
VGGT-HPE: Reframing Head Pose Estimation as Relative Pose Prediction
by: Vasileiou, Vasiliki, et al.
Published: (2026)
by: Vasileiou, Vasiliki, et al.
Published: (2026)
Registration-Free Learnable Multi-View Capture of Faces in Dense Semantic Correspondence
by: Filntisis, Panagiotis P., et al.
Published: (2026)
by: Filntisis, Panagiotis P., et al.
Published: (2026)
From Detection to Action Recognition: An Edge-Based Pipeline for Robot Human Perception
by: Toupas, Petros, et al.
Published: (2023)
by: Toupas, Petros, et al.
Published: (2023)
Boosting Gesture Recognition with an Automatic Gesture Annotation Framework
by: Shen, Junxiao, et al.
Published: (2024)
by: Shen, Junxiao, et al.
Published: (2024)
M2-CLIP: A Multimodal, Multi-task Adapting Framework for Video Action Recognition
by: Wang, Mengmeng, et al.
Published: (2024)
by: Wang, Mengmeng, et al.
Published: (2024)
Action Selection Learning for Multi-label Multi-view Action Recognition
by: Nguyen, Trung Thanh, et al.
Published: (2024)
by: Nguyen, Trung Thanh, et al.
Published: (2024)
Multi-Track Multimodal Learning on iMiGUE: Micro-Gesture and Emotion Recognition
by: Martirosyan, Arman, et al.
Published: (2025)
by: Martirosyan, Arman, et al.
Published: (2025)
Visual Hand Gesture Recognition with Deep Learning: A Comprehensive Review of Methods, Datasets, Challenges and Future Research Directions
by: Foteinos, Konstantinos, et al.
Published: (2025)
by: Foteinos, Konstantinos, et al.
Published: (2025)
Spatio-Temporal Joint Density Driven Learning for Skeleton-Based Action Recognition
by: Gunasekara, Shanaka Ramesh, et al.
Published: (2025)
by: Gunasekara, Shanaka Ramesh, et al.
Published: (2025)
Multi-task Learning with Extended Temporal Shift Module for Temporal Action Localization
by: Duong, Anh-Kiet, et al.
Published: (2025)
by: Duong, Anh-Kiet, et al.
Published: (2025)
Interpretable Underwater Diver Gesture Recognition
by: Mangalvedhekar, Sudeep, et al.
Published: (2023)
by: Mangalvedhekar, Sudeep, et al.
Published: (2023)
Zero-Shot Underwater Gesture Recognition
by: Sarma, Sandipan, et al.
Published: (2024)
by: Sarma, Sandipan, et al.
Published: (2024)
Towards Open-World Gesture Recognition
by: Shen, Junxiao, et al.
Published: (2024)
by: Shen, Junxiao, et al.
Published: (2024)
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
by: Gu, Jihao, et al.
Published: (2025)
by: Gu, Jihao, et al.
Published: (2025)
Flatten: Video Action Recognition is an Image Classification task
by: Chen, Junlin, et al.
Published: (2024)
by: Chen, Junlin, et al.
Published: (2024)
Joint Temporal Pooling for Improving Skeleton-based Action Recognition
by: Gunasekara, Shanaka Ramesh, et al.
Published: (2024)
by: Gunasekara, Shanaka Ramesh, et al.
Published: (2024)
SocialGesture: Delving into Multi-person Gesture Understanding
by: Cao, Xu, et al.
Published: (2025)
by: Cao, Xu, et al.
Published: (2025)
MMHMER:Multi-viewer and Multi-task for Handwritten Mathematical Expression Recognition
by: Chen, Kehua, et al.
Published: (2025)
by: Chen, Kehua, et al.
Published: (2025)
Instance-Level Composed Image Retrieval
by: Psomas, Bill, et al.
Published: (2025)
by: Psomas, Bill, et al.
Published: (2025)
SMIRK: 3D Facial Expressions through Analysis-by-Neural-Synthesis
by: Retsinas, George, et al.
Published: (2024)
by: Retsinas, George, et al.
Published: (2024)
Towards End-to-End Explainable Facial Action Unit Recognition via Vision-Language Joint Learning
by: Ge, Xuri, et al.
Published: (2024)
by: Ge, Xuri, et al.
Published: (2024)
Duo Streamers: A Streaming Gesture Recognition Framework
by: Zhu, Boxuan, et al.
Published: (2025)
by: Zhu, Boxuan, et al.
Published: (2025)
HaGRID - HAnd Gesture Recognition Image Dataset
by: Kapitanov, Alexander, et al.
Published: (2022)
by: Kapitanov, Alexander, et al.
Published: (2022)
MS-CLR: Multi-Skeleton Contrastive Learning for Human Action Recognition
by: Kiray, Mert, et al.
Published: (2025)
by: Kiray, Mert, et al.
Published: (2025)
Multi-Modality Co-Learning for Efficient Skeleton-based Action Recognition
by: Liu, Jinfu, et al.
Published: (2024)
by: Liu, Jinfu, et al.
Published: (2024)
Active Inference for Micro-Gesture Recognition: EFE-Guided Temporal Sampling and Adaptive Learning
by: Feng, Weijia, et al.
Published: (2026)
by: Feng, Weijia, et al.
Published: (2026)
Joint Image-Instance Spatial-Temporal Attention for Few-shot Action Recognition
by: Qian, Zefeng, et al.
Published: (2025)
by: Qian, Zefeng, et al.
Published: (2025)
Text to Image for Multi-Label Image Recognition with Joint Prompt-Adapter Learning
by: Feng, Chun-Mei, et al.
Published: (2025)
by: Feng, Chun-Mei, et al.
Published: (2025)
Forearm Ultrasound based Gesture Recognition on Edge
by: Bimbraw, Keshav, et al.
Published: (2024)
by: Bimbraw, Keshav, et al.
Published: (2024)
An Advanced Deep Learning Based Three-Stream Hybrid Model for Dynamic Hand Gesture Recognition
by: Rahim, Md Abdur, et al.
Published: (2024)
by: Rahim, Md Abdur, et al.
Published: (2024)
Learning using privileged information for segmenting tumors on digital mammograms
by: Tzortzis, Ioannis N., et al.
Published: (2024)
by: Tzortzis, Ioannis N., et al.
Published: (2024)
Similar Items
-
Pre-training for Action Recognition with Automatically Generated Fractal Datasets
by: Svyezhentsev, Davyd, et al.
Published: (2024) -
TropNNC: Structured Neural Network Compression Using Tropical Geometry
by: Fotopoulos, Konstantinos, et al.
Published: (2024) -
Object-Centric Action-Enhanced Representations for Robot Visuo-Motor Policy Learning
by: Giannakakis, Nikos, et al.
Published: (2025) -
Optimal Transport for Handwritten Text Recognition in a Low-Resource Regime
by: Wraight, Petros Georgoulas, et al.
Published: (2025) -
Age-Inclusive 3D Human Mesh Recovery for Action-Preserving Data Anonymization
by: Chatzichristodoulou, Georgios, et al.
Published: (2025)