Quantitative Gait Analysis from Single RGB Videos Using a Dual-Input Transformer-Based Network
Fuente:
arXiv
Saved in:
| Main Authors: | Dinh, Hiep, Le, Son, Than, My, Ho, Minh, Vuillerme, Nicolas, Pham, Hieu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Estimate Critical Gait Parameters from Single-View RGB Videos with Transformer-Based Attention Network
by: Le, Quoc Hung T., et al.
Published: (2023)
by: Le, Quoc Hung T., et al.
Published: (2023)
ConstStyle: Robust Domain Generalization with Unified Style Transformation
by: Tran, Nam Duong, et al.
Published: (2025)
by: Tran, Nam Duong, et al.
Published: (2025)
Fourier-Attentive Representation Learning: A Fourier-Guided Framework for Few-Shot Generalization in Vision-Language Models
by: Pham, Hieu Dinh Trung, et al.
Published: (2025)
by: Pham, Hieu Dinh Trung, et al.
Published: (2025)
NeurFlow: Interpreting Neural Networks through Neuron Groups and Functional Interactions
by: Cao, Tue M., et al.
Published: (2025)
by: Cao, Tue M., et al.
Published: (2025)
DuFal: Dual-Frequency-Aware Learning for High-Fidelity Extremely Sparse-view CBCT Reconstruction
by: Van, Cuong Tran, et al.
Published: (2026)
by: Van, Cuong Tran, et al.
Published: (2026)
D-SarcNet: A Dual-stream Deep Learning Framework for Automatic Analysis of Sarcomere Structures in Fluorescently Labeled hiPSC-CMs
by: Le, Huyen, et al.
Published: (2024)
by: Le, Huyen, et al.
Published: (2024)
SEMT: Static-Expansion-Mesh Transformer Network Architecture for Remote Sensing Image Captioning
by: Truong, Khang, et al.
Published: (2025)
by: Truong, Khang, et al.
Published: (2025)
Advancing Monocular Video-Based Gait Analysis Using Motion Imitation with Physics-Based Simulation
by: Smyrnakis, Nikolaos, et al.
Published: (2024)
by: Smyrnakis, Nikolaos, et al.
Published: (2024)
Using Machine Learning and Nationwide Population‐Based Data to Unravel Predictors of Treated Depression in Farmers
by: Pascal Petit, et al.
Published: (2025)
by: Pascal Petit, et al.
Published: (2025)
KGAlign: Joint Semantic-Structural Knowledge Encoding for Multimodal Fake News Detection
by: La, Tuan-Vinh, et al.
Published: (2025)
by: La, Tuan-Vinh, et al.
Published: (2025)
Structural Attention: Rethinking Transformer for Unpaired Medical Image Synthesis
by: Phan, Vu Minh Hieu, et al.
Published: (2024)
by: Phan, Vu Minh Hieu, et al.
Published: (2024)
Diffusion-based RGB-D Semantic Segmentation with Deformable Attention Transformer
by: Bui, Minh, et al.
Published: (2024)
by: Bui, Minh, et al.
Published: (2024)
Improving the Robustness of 3D Human Pose Estimation: A Benchmark and Learning from Noisy Input
by: Hoang, Trung-Hieu, et al.
Published: (2023)
by: Hoang, Trung-Hieu, et al.
Published: (2023)
VinDr-CXR-VQA: A Visual Question Answering Dataset for Explainable Chest X-Ray Analysis with Multi-Task Learning
by: Nguyen, Dang H., et al.
Published: (2025)
by: Nguyen, Dang H., et al.
Published: (2025)
TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning
by: Dinh, Quang Minh, et al.
Published: (2024)
by: Dinh, Quang Minh, et al.
Published: (2024)
Brain Network Analysis Based on Fine-tuned Self-supervised Model for Brain Disease Diagnosis
by: Tang, Yifei, et al.
Published: (2025)
by: Tang, Yifei, et al.
Published: (2025)
Importance-Based Token Merging for Efficient Image and Video Generation
by: Wu, Haoyu, et al.
Published: (2024)
by: Wu, Haoyu, et al.
Published: (2024)
High Resolution UDF Meshing via Iterative Networks
by: Stella, Federico, et al.
Published: (2025)
by: Stella, Federico, et al.
Published: (2025)
RobustGait: Robustness Analysis for Appearance Based Gait Recognition
by: Sayera, Reeshoon, et al.
Published: (2025)
by: Sayera, Reeshoon, et al.
Published: (2025)
Benchmarking Jetson Edge Devices with an End-to-end Video-based Anomaly Detection System
by: Pham, Hoang Viet, et al.
Published: (2023)
by: Pham, Hoang Viet, et al.
Published: (2023)
Enhancing Gait Video Analysis in Neurodegenerative Diseases by Knowledge Augmentation in Vision Language Model
by: Wang, Diwei, et al.
Published: (2024)
by: Wang, Diwei, et al.
Published: (2024)
Leveraging knowledge distillation for partial multi-task learning from multiple remote sensing datasets
by: Lê, Hoàng-Ân, et al.
Published: (2024)
by: Lê, Hoàng-Ân, et al.
Published: (2024)
Sampling Foundational Transformer: A Theoretical Perspective
by: Nguyen, Viet Anh, et al.
Published: (2024)
by: Nguyen, Viet Anh, et al.
Published: (2024)
Language-driven Grasp Detection
by: Vuong, An Dinh, et al.
Published: (2024)
by: Vuong, An Dinh, et al.
Published: (2024)
Exploring the Application of Visual Question Answering (VQA) for Classroom Activity Monitoring
by: Vu, Sinh Trong, et al.
Published: (2025)
by: Vu, Sinh Trong, et al.
Published: (2025)
Computer Vision for Clinical Gait Analysis: A Gait Abnormality Video Dataset
by: Ranjan, Rahm, et al.
Published: (2024)
by: Ranjan, Rahm, et al.
Published: (2024)
A2VIS: Amodal-Aware Approach to Video Instance Segmentation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
Technology Mapping with Large Language Models
by: Nguyen, Minh Hieu, et al.
Published: (2025)
by: Nguyen, Minh Hieu, et al.
Published: (2025)
Comparing Deep Neural Network for Multi-Label ECG Diagnosis From Scanned ECG
by: Nguyen, Cuong V., et al.
Published: (2025)
by: Nguyen, Cuong V., et al.
Published: (2025)
Explainable Parkinsons Disease Gait Recognition Using Multimodal RGB-D Fusion and Large Language Models
by: Alnaasan, Manar, et al.
Published: (2025)
by: Alnaasan, Manar, et al.
Published: (2025)
GaitPoint+: A Gait Recognition Network Incorporating Point Cloud Analysis and Recycling
by: Ren, Huantao, et al.
Published: (2024)
by: Ren, Huantao, et al.
Published: (2024)
Emotional Vietnamese Speech-Based Depression Diagnosis Using Dynamic Attention Mechanism
by: D., Quang-Anh N., et al.
Published: (2024)
by: D., Quang-Anh N., et al.
Published: (2024)
Combo-Gait: Unified Transformer Framework for Multi-Modal Gait Recognition and Attribute Analysis
by: Wang, Zhao-Yang, et al.
Published: (2025)
by: Wang, Zhao-Yang, et al.
Published: (2025)
Enhancing Rotated Object Detection via Anisotropic Gaussian Bounding Box and Bhattacharyya Distance
by: Thai, Chien, et al.
Published: (2025)
by: Thai, Chien, et al.
Published: (2025)
Smooth-Distill: A Self-distillation Framework for Multitask Learning with Wearable Sensor Data
by: Vu, Hoang-Dieu, et al.
Published: (2025)
by: Vu, Hoang-Dieu, et al.
Published: (2025)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
by: Le, Minh Khoa, et al.
Published: (2026)
by: Le, Minh Khoa, et al.
Published: (2026)
DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition
by: Li, Xinzhu, et al.
Published: (2025)
by: Li, Xinzhu, et al.
Published: (2025)
Integrating Image Features with Convolutional Sequence-to-sequence Network for Multilingual Visual Question Answering
by: Thai, Triet Minh, et al.
Published: (2023)
by: Thai, Triet Minh, et al.
Published: (2023)
Reinforcement Learning-Based REST API Testing with Multi-Coverage
by: Nguyen, Tien-Quang, et al.
Published: (2024)
by: Nguyen, Tien-Quang, et al.
Published: (2024)
Explainable Gait Abnormality Detection Using Dual-Dataset CNN-LSTM Models
by: Agarwal, Parth, et al.
Published: (2025)
by: Agarwal, Parth, et al.
Published: (2025)
Similar Items
-
Learning to Estimate Critical Gait Parameters from Single-View RGB Videos with Transformer-Based Attention Network
by: Le, Quoc Hung T., et al.
Published: (2023) -
ConstStyle: Robust Domain Generalization with Unified Style Transformation
by: Tran, Nam Duong, et al.
Published: (2025) -
Fourier-Attentive Representation Learning: A Fourier-Guided Framework for Few-Shot Generalization in Vision-Language Models
by: Pham, Hieu Dinh Trung, et al.
Published: (2025) -
NeurFlow: Interpreting Neural Networks through Neuron Groups and Functional Interactions
by: Cao, Tue M., et al.
Published: (2025) -
DuFal: Dual-Frequency-Aware Learning for High-Fidelity Extremely Sparse-view CBCT Reconstruction
by: Van, Cuong Tran, et al.
Published: (2026)