Quantitative Gait Analysis from Single RGB Videos Using a Dual-Input Transformer-Based Network
Fuente:
arXiv
Salvato in:
| Autori principali: | Dinh, Hiep, Le, Son, Than, My, Ho, Minh, Vuillerme, Nicolas, Pham, Hieu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning to Estimate Critical Gait Parameters from Single-View RGB Videos with Transformer-Based Attention Network
di: Le, Quoc Hung T., et al.
Pubblicazione: (2023)
di: Le, Quoc Hung T., et al.
Pubblicazione: (2023)
ConstStyle: Robust Domain Generalization with Unified Style Transformation
di: Tran, Nam Duong, et al.
Pubblicazione: (2025)
di: Tran, Nam Duong, et al.
Pubblicazione: (2025)
Fourier-Attentive Representation Learning: A Fourier-Guided Framework for Few-Shot Generalization in Vision-Language Models
di: Pham, Hieu Dinh Trung, et al.
Pubblicazione: (2025)
di: Pham, Hieu Dinh Trung, et al.
Pubblicazione: (2025)
NeurFlow: Interpreting Neural Networks through Neuron Groups and Functional Interactions
di: Cao, Tue M., et al.
Pubblicazione: (2025)
di: Cao, Tue M., et al.
Pubblicazione: (2025)
DuFal: Dual-Frequency-Aware Learning for High-Fidelity Extremely Sparse-view CBCT Reconstruction
di: Van, Cuong Tran, et al.
Pubblicazione: (2026)
di: Van, Cuong Tran, et al.
Pubblicazione: (2026)
D-SarcNet: A Dual-stream Deep Learning Framework for Automatic Analysis of Sarcomere Structures in Fluorescently Labeled hiPSC-CMs
di: Le, Huyen, et al.
Pubblicazione: (2024)
di: Le, Huyen, et al.
Pubblicazione: (2024)
SEMT: Static-Expansion-Mesh Transformer Network Architecture for Remote Sensing Image Captioning
di: Truong, Khang, et al.
Pubblicazione: (2025)
di: Truong, Khang, et al.
Pubblicazione: (2025)
Advancing Monocular Video-Based Gait Analysis Using Motion Imitation with Physics-Based Simulation
di: Smyrnakis, Nikolaos, et al.
Pubblicazione: (2024)
di: Smyrnakis, Nikolaos, et al.
Pubblicazione: (2024)
Using Machine Learning and Nationwide Population‐Based Data to Unravel Predictors of Treated Depression in Farmers
di: Pascal Petit, et al.
Pubblicazione: (2025)
di: Pascal Petit, et al.
Pubblicazione: (2025)
KGAlign: Joint Semantic-Structural Knowledge Encoding for Multimodal Fake News Detection
di: La, Tuan-Vinh, et al.
Pubblicazione: (2025)
di: La, Tuan-Vinh, et al.
Pubblicazione: (2025)
Structural Attention: Rethinking Transformer for Unpaired Medical Image Synthesis
di: Phan, Vu Minh Hieu, et al.
Pubblicazione: (2024)
di: Phan, Vu Minh Hieu, et al.
Pubblicazione: (2024)
Diffusion-based RGB-D Semantic Segmentation with Deformable Attention Transformer
di: Bui, Minh, et al.
Pubblicazione: (2024)
di: Bui, Minh, et al.
Pubblicazione: (2024)
Improving the Robustness of 3D Human Pose Estimation: A Benchmark and Learning from Noisy Input
di: Hoang, Trung-Hieu, et al.
Pubblicazione: (2023)
di: Hoang, Trung-Hieu, et al.
Pubblicazione: (2023)
VinDr-CXR-VQA: A Visual Question Answering Dataset for Explainable Chest X-Ray Analysis with Multi-Task Learning
di: Nguyen, Dang H., et al.
Pubblicazione: (2025)
di: Nguyen, Dang H., et al.
Pubblicazione: (2025)
TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning
di: Dinh, Quang Minh, et al.
Pubblicazione: (2024)
di: Dinh, Quang Minh, et al.
Pubblicazione: (2024)
Brain Network Analysis Based on Fine-tuned Self-supervised Model for Brain Disease Diagnosis
di: Tang, Yifei, et al.
Pubblicazione: (2025)
di: Tang, Yifei, et al.
Pubblicazione: (2025)
Importance-Based Token Merging for Efficient Image and Video Generation
di: Wu, Haoyu, et al.
Pubblicazione: (2024)
di: Wu, Haoyu, et al.
Pubblicazione: (2024)
High Resolution UDF Meshing via Iterative Networks
di: Stella, Federico, et al.
Pubblicazione: (2025)
di: Stella, Federico, et al.
Pubblicazione: (2025)
RobustGait: Robustness Analysis for Appearance Based Gait Recognition
di: Sayera, Reeshoon, et al.
Pubblicazione: (2025)
di: Sayera, Reeshoon, et al.
Pubblicazione: (2025)
Benchmarking Jetson Edge Devices with an End-to-end Video-based Anomaly Detection System
di: Pham, Hoang Viet, et al.
Pubblicazione: (2023)
di: Pham, Hoang Viet, et al.
Pubblicazione: (2023)
Enhancing Gait Video Analysis in Neurodegenerative Diseases by Knowledge Augmentation in Vision Language Model
di: Wang, Diwei, et al.
Pubblicazione: (2024)
di: Wang, Diwei, et al.
Pubblicazione: (2024)
Leveraging knowledge distillation for partial multi-task learning from multiple remote sensing datasets
di: Lê, Hoàng-Ân, et al.
Pubblicazione: (2024)
di: Lê, Hoàng-Ân, et al.
Pubblicazione: (2024)
Sampling Foundational Transformer: A Theoretical Perspective
di: Nguyen, Viet Anh, et al.
Pubblicazione: (2024)
di: Nguyen, Viet Anh, et al.
Pubblicazione: (2024)
Language-driven Grasp Detection
di: Vuong, An Dinh, et al.
Pubblicazione: (2024)
di: Vuong, An Dinh, et al.
Pubblicazione: (2024)
Exploring the Application of Visual Question Answering (VQA) for Classroom Activity Monitoring
di: Vu, Sinh Trong, et al.
Pubblicazione: (2025)
di: Vu, Sinh Trong, et al.
Pubblicazione: (2025)
Computer Vision for Clinical Gait Analysis: A Gait Abnormality Video Dataset
di: Ranjan, Rahm, et al.
Pubblicazione: (2024)
di: Ranjan, Rahm, et al.
Pubblicazione: (2024)
A2VIS: Amodal-Aware Approach to Video Instance Segmentation
di: Tran, Minh, et al.
Pubblicazione: (2024)
di: Tran, Minh, et al.
Pubblicazione: (2024)
Technology Mapping with Large Language Models
di: Nguyen, Minh Hieu, et al.
Pubblicazione: (2025)
di: Nguyen, Minh Hieu, et al.
Pubblicazione: (2025)
Comparing Deep Neural Network for Multi-Label ECG Diagnosis From Scanned ECG
di: Nguyen, Cuong V., et al.
Pubblicazione: (2025)
di: Nguyen, Cuong V., et al.
Pubblicazione: (2025)
Explainable Parkinsons Disease Gait Recognition Using Multimodal RGB-D Fusion and Large Language Models
di: Alnaasan, Manar, et al.
Pubblicazione: (2025)
di: Alnaasan, Manar, et al.
Pubblicazione: (2025)
GaitPoint+: A Gait Recognition Network Incorporating Point Cloud Analysis and Recycling
di: Ren, Huantao, et al.
Pubblicazione: (2024)
di: Ren, Huantao, et al.
Pubblicazione: (2024)
Emotional Vietnamese Speech-Based Depression Diagnosis Using Dynamic Attention Mechanism
di: D., Quang-Anh N., et al.
Pubblicazione: (2024)
di: D., Quang-Anh N., et al.
Pubblicazione: (2024)
Combo-Gait: Unified Transformer Framework for Multi-Modal Gait Recognition and Attribute Analysis
di: Wang, Zhao-Yang, et al.
Pubblicazione: (2025)
di: Wang, Zhao-Yang, et al.
Pubblicazione: (2025)
Enhancing Rotated Object Detection via Anisotropic Gaussian Bounding Box and Bhattacharyya Distance
di: Thai, Chien, et al.
Pubblicazione: (2025)
di: Thai, Chien, et al.
Pubblicazione: (2025)
Smooth-Distill: A Self-distillation Framework for Multitask Learning with Wearable Sensor Data
di: Vu, Hoang-Dieu, et al.
Pubblicazione: (2025)
di: Vu, Hoang-Dieu, et al.
Pubblicazione: (2025)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
di: Le, Minh Khoa, et al.
Pubblicazione: (2026)
di: Le, Minh Khoa, et al.
Pubblicazione: (2026)
DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition
di: Li, Xinzhu, et al.
Pubblicazione: (2025)
di: Li, Xinzhu, et al.
Pubblicazione: (2025)
Integrating Image Features with Convolutional Sequence-to-sequence Network for Multilingual Visual Question Answering
di: Thai, Triet Minh, et al.
Pubblicazione: (2023)
di: Thai, Triet Minh, et al.
Pubblicazione: (2023)
Reinforcement Learning-Based REST API Testing with Multi-Coverage
di: Nguyen, Tien-Quang, et al.
Pubblicazione: (2024)
di: Nguyen, Tien-Quang, et al.
Pubblicazione: (2024)
Explainable Gait Abnormality Detection Using Dual-Dataset CNN-LSTM Models
di: Agarwal, Parth, et al.
Pubblicazione: (2025)
di: Agarwal, Parth, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Learning to Estimate Critical Gait Parameters from Single-View RGB Videos with Transformer-Based Attention Network
di: Le, Quoc Hung T., et al.
Pubblicazione: (2023) -
ConstStyle: Robust Domain Generalization with Unified Style Transformation
di: Tran, Nam Duong, et al.
Pubblicazione: (2025) -
Fourier-Attentive Representation Learning: A Fourier-Guided Framework for Few-Shot Generalization in Vision-Language Models
di: Pham, Hieu Dinh Trung, et al.
Pubblicazione: (2025) -
NeurFlow: Interpreting Neural Networks through Neuron Groups and Functional Interactions
di: Cao, Tue M., et al.
Pubblicazione: (2025) -
DuFal: Dual-Frequency-Aware Learning for High-Fidelity Extremely Sparse-view CBCT Reconstruction
di: Van, Cuong Tran, et al.
Pubblicazione: (2026)