Deep self-supervised learning with visualisation for automatic gesture recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Allemand, Fabien, Mazzela, Alessio, Villette, Jun, Aspandi, Decky, Zaharia, Titus |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-Modal interpretable automatic video captioning
di: Hanna-Asaad, Antoine, et al.
Pubblicazione: (2024)
di: Hanna-Asaad, Antoine, et al.
Pubblicazione: (2024)
Variational Contrastive Learning for Skeleton-based Action Recognition
di: Nguyen, Dang Dinh, et al.
Pubblicazione: (2026)
di: Nguyen, Dang Dinh, et al.
Pubblicazione: (2026)
Accurate online action and gesture recognition system using detectors and Deep SPD Siamese Networks
di: Akremi, Mohamed Sanim, et al.
Pubblicazione: (2025)
di: Akremi, Mohamed Sanim, et al.
Pubblicazione: (2025)
Helios: An extremely low power event-based gesture recognition for always-on smart eyewear
di: Bhattacharyya, Prarthana, et al.
Pubblicazione: (2024)
di: Bhattacharyya, Prarthana, et al.
Pubblicazione: (2024)
UST-Hand: An Uncertainty-aware Spatiotemporal Point Cloud Interaction Network for 3D Self-supervised Hand Pose Estimation
di: Han, Tianhao, et al.
Pubblicazione: (2026)
di: Han, Tianhao, et al.
Pubblicazione: (2026)
SVFAP: Self-supervised Video Facial Affect Perceiver
di: Sun, Licai, et al.
Pubblicazione: (2023)
di: Sun, Licai, et al.
Pubblicazione: (2023)
Deep Learning for Virtual Reality User Identification: A Benchmark
di: Frizzo, Davide, et al.
Pubblicazione: (2026)
di: Frizzo, Davide, et al.
Pubblicazione: (2026)
Unsupervised learning of Data-driven Facial Expression Coding System (DFECS) using keypoint tracking
di: Tripathi, Shivansh Chandra, et al.
Pubblicazione: (2024)
di: Tripathi, Shivansh Chandra, et al.
Pubblicazione: (2024)
A Deep Learning Framework for Visual Attention Prediction and Analysis of News Interfaces
di: Kenely, Matthew, et al.
Pubblicazione: (2025)
di: Kenely, Matthew, et al.
Pubblicazione: (2025)
DeepSORT-Driven Visual Tracking Approach for Gesture Recognition in Interactive Systems
di: Zhang, Tong, et al.
Pubblicazione: (2025)
di: Zhang, Tong, et al.
Pubblicazione: (2025)
Deep Learning-based Lightweight RGB Object Tracking for Augmented Reality Devices
di: Smith, Alice, et al.
Pubblicazione: (2025)
di: Smith, Alice, et al.
Pubblicazione: (2025)
Facial recognition technology and human raters can predict political orientation from images of expressionless faces even when controlling for demographics and self-presentation
di: Kosinski, Michal, et al.
Pubblicazione: (2023)
di: Kosinski, Michal, et al.
Pubblicazione: (2023)
DeepFace-Attention: Multimodal Face Biometrics for Attention Estimation with Application to e-Learning
di: Daza, Roberto, et al.
Pubblicazione: (2024)
di: Daza, Roberto, et al.
Pubblicazione: (2024)
Deep Neural Encoder-Decoder Model to Relate fMRI Brain Activity with Naturalistic Stimuli
di: David, Florian, et al.
Pubblicazione: (2025)
di: David, Florian, et al.
Pubblicazione: (2025)
Deep Learning in Mild Cognitive Impairment Diagnosis using Eye Movements and Image Content in Visual Memory Tasks
di: Rocha, Tomás Silva Santos, et al.
Pubblicazione: (2025)
di: Rocha, Tomás Silva Santos, et al.
Pubblicazione: (2025)
ChatStitch: Visualizing Through Structures via Surround-View Unsupervised Deep Image Stitching with Collaborative LLM-Agents
di: Liang, Hao, et al.
Pubblicazione: (2025)
di: Liang, Hao, et al.
Pubblicazione: (2025)
UF-AMA: A unified framework for cross-domain emotion recognition via adaptive multimodal alignment
di: Wang, Zheng, et al.
Pubblicazione: (2026)
di: Wang, Zheng, et al.
Pubblicazione: (2026)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
di: Liu, Zhi-Song, et al.
Pubblicazione: (2024)
di: Liu, Zhi-Song, et al.
Pubblicazione: (2024)
Adaptive Modality Balanced Online Knowledge Distillation for Brain-Eye-Computer based Dim Object Detection
di: Li, Zixing, et al.
Pubblicazione: (2024)
di: Li, Zixing, et al.
Pubblicazione: (2024)
Dreamcrafter: Immersive Editing of 3D Radiance Fields Through Flexible, Generative Inputs and Outputs
di: Vachha, Cyrus, et al.
Pubblicazione: (2025)
di: Vachha, Cyrus, et al.
Pubblicazione: (2025)
YOLOA: Real-Time Affordance Detection via LLM Adapter
di: Ji, Yuqi, et al.
Pubblicazione: (2025)
di: Ji, Yuqi, et al.
Pubblicazione: (2025)
Deep Sketch-Based 3D Modeling: A Survey
di: Tono, Alberto, et al.
Pubblicazione: (2026)
di: Tono, Alberto, et al.
Pubblicazione: (2026)
Analyzing Swimming Performance Using Drone Captured Aerial Videos
di: Tran, Thu, et al.
Pubblicazione: (2025)
di: Tran, Thu, et al.
Pubblicazione: (2025)
Motion Generation Review: Exploring Deep Learning for Lifelike Animation with Manifold
di: Zhao, Jiayi, et al.
Pubblicazione: (2024)
di: Zhao, Jiayi, et al.
Pubblicazione: (2024)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
di: Lin, Yukang, et al.
Pubblicazione: (2025)
di: Lin, Yukang, et al.
Pubblicazione: (2025)
Topology of surface electromyogram signals: hand gesture decoding on Riemannian manifolds
di: Gowda, Harshavardhana T., et al.
Pubblicazione: (2023)
di: Gowda, Harshavardhana T., et al.
Pubblicazione: (2023)
User Experience Estimation in Human-Robot Interaction Via Multi-Instance Learning of Multimodal Social Signals
di: Miyoshi, Ryo, et al.
Pubblicazione: (2025)
di: Miyoshi, Ryo, et al.
Pubblicazione: (2025)
Unsupervised visualization of image datasets using contrastive learning
di: Böhm, Jan Niklas, et al.
Pubblicazione: (2022)
di: Böhm, Jan Niklas, et al.
Pubblicazione: (2022)
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
di: Rekimoto, Jun
Pubblicazione: (2025)
di: Rekimoto, Jun
Pubblicazione: (2025)
A Dataset for Crucial Object Recognition in Blind and Low-Vision Individuals' Navigation
di: Islam, Md Touhidul, et al.
Pubblicazione: (2024)
di: Islam, Md Touhidul, et al.
Pubblicazione: (2024)
Resource-Efficient Gesture Recognition using Low-Resolution Thermal Camera via Spiking Neural Networks and Sparse Segmentation
di: Safa, Ali, et al.
Pubblicazione: (2024)
di: Safa, Ali, et al.
Pubblicazione: (2024)
Towards Context-aware Support for Color Vision Deficiency: An Approach Integrating LLM and AR
di: Morita, Shogo, et al.
Pubblicazione: (2024)
di: Morita, Shogo, et al.
Pubblicazione: (2024)
Stratified Avatar Generation from Sparse Observations
di: Feng, Han, et al.
Pubblicazione: (2024)
di: Feng, Han, et al.
Pubblicazione: (2024)
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
di: Jiao, Chuhan, et al.
Pubblicazione: (2024)
di: Jiao, Chuhan, et al.
Pubblicazione: (2024)
Generalized Pose Space Embeddings for Training In-the-Wild using Anaylis-by-Synthesis
di: Borer, Dominik, et al.
Pubblicazione: (2024)
di: Borer, Dominik, et al.
Pubblicazione: (2024)
Don't Look at the Camera: Achieving Perceived Eye Contact
di: Gao, Alice, et al.
Pubblicazione: (2024)
di: Gao, Alice, et al.
Pubblicazione: (2024)
Real-Time Hand Gesture Recognition: Integrating Skeleton-Based Data Fusion and Multi-Stream CNN
di: Yusuf, Oluwaleke, et al.
Pubblicazione: (2024)
di: Yusuf, Oluwaleke, et al.
Pubblicazione: (2024)
Enhanced Automated Quality Assessment Network for Interactive Building Segmentation in High-Resolution Remote Sensing Imagery
di: Zhang, Zhili, et al.
Pubblicazione: (2024)
di: Zhang, Zhili, et al.
Pubblicazione: (2024)
AccessLens: Auto-detecting Inaccessibility of Everyday Objects
di: Kwon, Nahyun, et al.
Pubblicazione: (2024)
di: Kwon, Nahyun, et al.
Pubblicazione: (2024)
Leveraging Digital Perceptual Technologies for Remote Perception and Analysis of Human Biomechanical Processes: A Contactless Approach for Workload and Joint Force Assessment
di: Omidokun, Jesudara, et al.
Pubblicazione: (2024)
di: Omidokun, Jesudara, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Multi-Modal interpretable automatic video captioning
di: Hanna-Asaad, Antoine, et al.
Pubblicazione: (2024) -
Variational Contrastive Learning for Skeleton-based Action Recognition
di: Nguyen, Dang Dinh, et al.
Pubblicazione: (2026) -
Accurate online action and gesture recognition system using detectors and Deep SPD Siamese Networks
di: Akremi, Mohamed Sanim, et al.
Pubblicazione: (2025) -
Helios: An extremely low power event-based gesture recognition for always-on smart eyewear
di: Bhattacharyya, Prarthana, et al.
Pubblicazione: (2024) -
UST-Hand: An Uncertainty-aware Spatiotemporal Point Cloud Interaction Network for 3D Self-supervised Hand Pose Estimation
di: Han, Tianhao, et al.
Pubblicazione: (2026)