Advanced Gesture Recognition for Autism Spectrum Disorder Detection: Integrating YOLOv7, Video Augmentation, and VideoMAE for Naturalistic Video Analysis
Fuente:
arXiv
Salvato in:
| Autori principali: | Singh, Amit Kumar, Singh, Vrijendra |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FedVideoMAE: Efficient Privacy-Preserving Federated Video Moderation
di: Tao, Ziyuan, et al.
Pubblicazione: (2025)
di: Tao, Ziyuan, et al.
Pubblicazione: (2025)
RAVU: Retrieval Augmented Video Understanding with Compositional Reasoning over Graph
di: Malik, Sameer, et al.
Pubblicazione: (2025)
di: Malik, Sameer, et al.
Pubblicazione: (2025)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
di: Liu, Pinxin, et al.
Pubblicazione: (2025)
di: Liu, Pinxin, et al.
Pubblicazione: (2025)
Enhancing Bidirectional Sign Language Communication: Integrating YOLOv8 and NLP for Real-Time Gesture Recognition & Translation
di: Bhuiyan, Hasnat Jamil, et al.
Pubblicazione: (2024)
di: Bhuiyan, Hasnat Jamil, et al.
Pubblicazione: (2024)
Surveillance Video-Based Traffic Accident Detection Using Transformer Architecture
di: Singh, Tanu, et al.
Pubblicazione: (2025)
di: Singh, Tanu, et al.
Pubblicazione: (2025)
Augmenting End-to-End Steering Angle Prediction with CAN Bus Data
di: Singh, Amit
Pubblicazione: (2023)
di: Singh, Amit
Pubblicazione: (2023)
Exploring Explainability in Video Action Recognition
di: Saha, Avinab, et al.
Pubblicazione: (2024)
di: Saha, Avinab, et al.
Pubblicazione: (2024)
Democratizing High-Fidelity Co-Speech Gesture Video Generation
di: Yang, Xu, et al.
Pubblicazione: (2025)
di: Yang, Xu, et al.
Pubblicazione: (2025)
This&That: Language-Gesture Controlled Video Generation for Robot Planning
di: Wang, Boyang, et al.
Pubblicazione: (2024)
di: Wang, Boyang, et al.
Pubblicazione: (2024)
Detecting Children with Autism Spectrum Disorder based on Script-Centric Behavior Understanding with Emotional Enhancement
di: Liu, Wenxing, et al.
Pubblicazione: (2024)
di: Liu, Wenxing, et al.
Pubblicazione: (2024)
Think Step by Step: Chain-of-Gesture Prompting for Error Detection in Robotic Surgical Videos
di: Shao, Zhimin, et al.
Pubblicazione: (2024)
di: Shao, Zhimin, et al.
Pubblicazione: (2024)
SALOVA: Segment-Augmented Long Video Assistant for Targeted Retrieval and Routing in Long-Form Video Analysis
di: Kim, Junho, et al.
Pubblicazione: (2024)
di: Kim, Junho, et al.
Pubblicazione: (2024)
Video-RAG: Visually-aligned Retrieval-Augmented Long Video Comprehension
di: Luo, Yongdong, et al.
Pubblicazione: (2024)
di: Luo, Yongdong, et al.
Pubblicazione: (2024)
On Occlusions in Video Action Detection: Benchmark Datasets And Training Recipes
di: Modi, Rajat, et al.
Pubblicazione: (2024)
di: Modi, Rajat, et al.
Pubblicazione: (2024)
VideoWeave: A Data-Centric Approach for Efficient Video Understanding
di: Durante, Zane, et al.
Pubblicazione: (2026)
di: Durante, Zane, et al.
Pubblicazione: (2026)
On Denoising Walking Videos for Gait Recognition
di: Jin, Dongyang, et al.
Pubblicazione: (2025)
di: Jin, Dongyang, et al.
Pubblicazione: (2025)
Agentic Video Intelligence: A Flexible Framework for Advanced Video Exploration and Understanding
di: Gao, Hong, et al.
Pubblicazione: (2025)
di: Gao, Hong, et al.
Pubblicazione: (2025)
Ensemble Modeling of Multiple Physical Indicators to Dynamically Phenotype Autism Spectrum Disorder
di: Huynh, Marie, et al.
Pubblicazione: (2024)
di: Huynh, Marie, et al.
Pubblicazione: (2024)
Advance Fake Video Detection via Vision Transformers
di: Battocchio, Joy, et al.
Pubblicazione: (2025)
di: Battocchio, Joy, et al.
Pubblicazione: (2025)
Continuous Patient Monitoring with AI: Real-Time Analysis of Video in Hospital Care Settings
di: Gabriel, Paolo, et al.
Pubblicazione: (2024)
di: Gabriel, Paolo, et al.
Pubblicazione: (2024)
ReTool-Video: Recursive Tool-Using Video Agents with Meta-Augmented Tool Grounding
di: Liu, Xiao, et al.
Pubblicazione: (2026)
di: Liu, Xiao, et al.
Pubblicazione: (2026)
Reangle-A-Video: 4D Video Generation as Video-to-Video Translation
di: Jeong, Hyeonho, et al.
Pubblicazione: (2025)
di: Jeong, Hyeonho, et al.
Pubblicazione: (2025)
Automated Wicket-Taking Delivery Segmentation and Trajectory-Based Dismissal-Zone Analysis in Cricket Videos Using OCR-Guided YOLOv8
di: Karmoker, Joy, et al.
Pubblicazione: (2025)
di: Karmoker, Joy, et al.
Pubblicazione: (2025)
Spatiotemporal Learning with Context-aware Video Tubelets for Ultrasound Video Analysis
di: Li, Gary Y., et al.
Pubblicazione: (2025)
di: Li, Gary Y., et al.
Pubblicazione: (2025)
GV-VAD : Exploring Video Generation for Weakly-Supervised Video Anomaly Detection
di: Cai, Suhang, et al.
Pubblicazione: (2025)
di: Cai, Suhang, et al.
Pubblicazione: (2025)
StreamMem: Query-Agnostic KV Cache Memory for Streaming Video Understanding
di: Yang, Yanlai, et al.
Pubblicazione: (2025)
di: Yang, Yanlai, et al.
Pubblicazione: (2025)
Artificial Intelligence in Gastrointestinal Bleeding Analysis for Video Capsule Endoscopy: Insights, Innovations, and Prospects (2008-2023)
di: Singh, Tanisha, et al.
Pubblicazione: (2024)
di: Singh, Tanisha, et al.
Pubblicazione: (2024)
Multimodal Video Emotion Recognition with Reliable Reasoning Priors
di: Wang, Zhepeng, et al.
Pubblicazione: (2025)
di: Wang, Zhepeng, et al.
Pubblicazione: (2025)
SkateboardAI: The Coolest Video Action Recognition for Skateboarding
di: Chen, Hanxiao
Pubblicazione: (2023)
di: Chen, Hanxiao
Pubblicazione: (2023)
A Survey on Backbones for Deep Video Action Recognition
di: Tang, Zixuan, et al.
Pubblicazione: (2024)
di: Tang, Zixuan, et al.
Pubblicazione: (2024)
Flatten: Video Action Recognition is an Image Classification task
di: Chen, Junlin, et al.
Pubblicazione: (2024)
di: Chen, Junlin, et al.
Pubblicazione: (2024)
Exploring Ordinal Bias in Action Recognition for Instructional Videos
di: Kim, Joochan, et al.
Pubblicazione: (2025)
di: Kim, Joochan, et al.
Pubblicazione: (2025)
Revealing Temporal Label Noise in Multimodal Hateful Video Classification
di: Yang, Shuonan, et al.
Pubblicazione: (2025)
di: Yang, Shuonan, et al.
Pubblicazione: (2025)
Comparative Analysis of Image, Video, and Audio Classifiers for Automated News Video Segmentation
di: Attard, Jonathan, et al.
Pubblicazione: (2025)
di: Attard, Jonathan, et al.
Pubblicazione: (2025)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
di: Kumar, Yogesh, et al.
Pubblicazione: (2025)
di: Kumar, Yogesh, et al.
Pubblicazione: (2025)
Video Enriched Retrieval Augmented Generation Using Aligned Video Captions
di: Rosa, Kevin Dela
Pubblicazione: (2024)
di: Rosa, Kevin Dela
Pubblicazione: (2024)
VideoRAG: Retrieval-Augmented Generation with Extreme Long-Context Videos
di: Ren, Xubin, et al.
Pubblicazione: (2025)
di: Ren, Xubin, et al.
Pubblicazione: (2025)
Towards Video Anomaly Retrieval from Video Anomaly Detection: New Benchmarks and Model
di: Wu, Peng, et al.
Pubblicazione: (2023)
di: Wu, Peng, et al.
Pubblicazione: (2023)
MMASD+: A Novel Dataset for Privacy-Preserving Behavior Analysis of Children with Autism Spectrum Disorder
di: Ravva, Pavan Uttej, et al.
Pubblicazione: (2024)
di: Ravva, Pavan Uttej, et al.
Pubblicazione: (2024)
Dissecting Multimodality in VideoQA Transformer Models by Impairing Modality Fusion
di: Rawal, Ishaan Singh, et al.
Pubblicazione: (2023)
di: Rawal, Ishaan Singh, et al.
Pubblicazione: (2023)
Documenti analoghi
-
FedVideoMAE: Efficient Privacy-Preserving Federated Video Moderation
di: Tao, Ziyuan, et al.
Pubblicazione: (2025) -
RAVU: Retrieval Augmented Video Understanding with Compositional Reasoning over Graph
di: Malik, Sameer, et al.
Pubblicazione: (2025) -
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
di: Liu, Pinxin, et al.
Pubblicazione: (2025) -
Enhancing Bidirectional Sign Language Communication: Integrating YOLOv8 and NLP for Real-Time Gesture Recognition & Translation
di: Bhuiyan, Hasnat Jamil, et al.
Pubblicazione: (2024) -
Surveillance Video-Based Traffic Accident Detection Using Transformer Architecture
di: Singh, Tanu, et al.
Pubblicazione: (2025)