Enhancing Gait Video Analysis in Neurodegenerative Diseases by Knowledge Augmentation in Vision Language Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Diwei, Yuan, Kun, Muller, Candice, Blanc, Frédéric, Padoy, Nicolas, Seo, Hyewon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
KRAST: Knowledge-Augmented Robotic Action Recognition with Structured Text for Vision-Language Models
von: Nguyen, Son Hai, et al.
Veröffentlicht: (2025)
von: Nguyen, Son Hai, et al.
Veröffentlicht: (2025)
AGIR: Assessing 3D Gait Impairment with Reasoning based on LLMs
von: Wang, Diwei, et al.
Veröffentlicht: (2025)
von: Wang, Diwei, et al.
Veröffentlicht: (2025)
Procedure-Aware Surgical Video-language Pretraining with Hierarchical Knowledge Augmentation
von: Yuan, Kun, et al.
Veröffentlicht: (2024)
von: Yuan, Kun, et al.
Veröffentlicht: (2024)
fine-CLIP: Enhancing Zero-Shot Fine-Grained Surgical Action Recognition with Vision-Language Models
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)
HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition
von: Yuan, Kun, et al.
Veröffentlicht: (2024)
von: Yuan, Kun, et al.
Veröffentlicht: (2024)
CARE-PD: A Multi-Site Anonymized Clinical Dataset for Parkinson's Disease Gait Assessment
von: Adeli, Vida, et al.
Veröffentlicht: (2025)
von: Adeli, Vida, et al.
Veröffentlicht: (2025)
PhyDeformer: High-Quality Non-Rigid Garment Registration with Physics-Awareness
von: Yu, Boyang, et al.
Veröffentlicht: (2025)
von: Yu, Boyang, et al.
Veröffentlicht: (2025)
Text-driven Adaptation of Foundation Models for Few-shot Surgical Workflow Analysis
von: Chen, Tingxuan, et al.
Veröffentlicht: (2025)
von: Chen, Tingxuan, et al.
Veröffentlicht: (2025)
SurgiTrack: Fine-Grained Multi-Class Multi-Tool Tracking in Surgical Videos
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024)
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024)
Shape Conditioned Human Motion Generation with Diffusion Model
von: Xue, Kebing, et al.
Veröffentlicht: (2024)
von: Xue, Kebing, et al.
Veröffentlicht: (2024)
Advancing Surgical VQA with Scene Graph Knowledge
von: Yuan, Kun, et al.
Veröffentlicht: (2023)
von: Yuan, Kun, et al.
Veröffentlicht: (2023)
A Survey of Artificial Intelligence in Gait-Based Neurodegenerative Disease Diagnosis
von: Rao, Haocong, et al.
Veröffentlicht: (2024)
von: Rao, Haocong, et al.
Veröffentlicht: (2024)
CliPPER: Contextual Video-Language Pretraining on Long-form Intraoperative Surgical Procedures for Event Recognition
von: Stilz, Florian, et al.
Veröffentlicht: (2026)
von: Stilz, Florian, et al.
Veröffentlicht: (2026)
Adaptation of Multi-modal Representation Models for Multi-task Surgical Computer Vision
von: Walimbe, Soham, et al.
Veröffentlicht: (2025)
von: Walimbe, Soham, et al.
Veröffentlicht: (2025)
SurgTEMP: Temporal-Aware Surgical Video Question Answering with Text-guided Visual Memory for Laparoscopic Cholecystectomy
von: Li, Shi, et al.
Veröffentlicht: (2026)
von: Li, Shi, et al.
Veröffentlicht: (2026)
Computer Vision for Clinical Gait Analysis: A Gait Abnormality Video Dataset
von: Ranjan, Rahm, et al.
Veröffentlicht: (2024)
von: Ranjan, Rahm, et al.
Veröffentlicht: (2024)
On-the-Fly Point Annotation for Fast Medical Video Labeling
von: Adrien, Meyer, et al.
Veröffentlicht: (2024)
von: Adrien, Meyer, et al.
Veröffentlicht: (2024)
State-Change Learning for Prediction of Future Events in Endoscopic Videos
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)
von: Sharma, Saurav, et al.
Veröffentlicht: (2025)
OphCLIP: Hierarchical Retrieval-Augmented Learning for Ophthalmic Surgical Video-Language Pretraining
von: Hu, Ming, et al.
Veröffentlicht: (2024)
von: Hu, Ming, et al.
Veröffentlicht: (2024)
SelfPose3d: Self-Supervised Multi-Person Multi-View 3d Pose Estimation
von: Srivastav, Vinkle, et al.
Veröffentlicht: (2024)
von: Srivastav, Vinkle, et al.
Veröffentlicht: (2024)
DExTeR: Weakly Semi-Supervised Object Detection with Class and Instance Experts for Medical Imaging
von: Meyer, Adrien, et al.
Veröffentlicht: (2026)
von: Meyer, Adrien, et al.
Veröffentlicht: (2026)
BigGait: Learning Gait Representation You Want by Large Vision Models
von: Ye, Dingqiang, et al.
Veröffentlicht: (2024)
von: Ye, Dingqiang, et al.
Veröffentlicht: (2024)
Learning Multi-modal Representations by Watching Hundreds of Surgical Video Lectures
von: Yuan, Kun, et al.
Veröffentlicht: (2023)
von: Yuan, Kun, et al.
Veröffentlicht: (2023)
From Panel to Pixel: Zoom-In Vision-Language Pretraining from Biomedical Scientific Literature
von: Yuan, Kun, et al.
Veröffentlicht: (2025)
von: Yuan, Kun, et al.
Veröffentlicht: (2025)
Self-Supervised Uncalibrated Multi-View Video Anonymization in the Operating Room
von: Chen, Keqi, et al.
Veröffentlicht: (2026)
von: Chen, Keqi, et al.
Veröffentlicht: (2026)
Knowledge-Driven Vision-Language Model for Plexus Detection in Hirschsprung's Disease
von: Megahed, Youssef, et al.
Veröffentlicht: (2025)
von: Megahed, Youssef, et al.
Veröffentlicht: (2025)
CoSimGen: Controllable Diffusion Model for Simultaneous Image and Mask Generation
von: Bose, Rupak, et al.
Veröffentlicht: (2025)
von: Bose, Rupak, et al.
Veröffentlicht: (2025)
SurgVidLM: Towards Multi-grained Surgical Video Understanding with Large Language Model
von: Wang, Guankun, et al.
Veröffentlicht: (2025)
von: Wang, Guankun, et al.
Veröffentlicht: (2025)
Knowledge-Augmented Vision Language Models for Underwater Bioacoustic Spectrogram Analysis
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2025)
von: Nihal, Ragib Amin, et al.
Veröffentlicht: (2025)
4D Facial Expression Diffusion Model
von: Zou, Kaifeng, et al.
Veröffentlicht: (2023)
von: Zou, Kaifeng, et al.
Veröffentlicht: (2023)
Jumpstarting Surgical Computer Vision
von: Alapatt, Deepak, et al.
Veröffentlicht: (2023)
von: Alapatt, Deepak, et al.
Veröffentlicht: (2023)
BiggerGait: Unlocking Gait Recognition with Layer-wise Representations from Large Vision Models
von: Ye, Dingqiang, et al.
Veröffentlicht: (2025)
von: Ye, Dingqiang, et al.
Veröffentlicht: (2025)
Self-Supervised Cross-Encoder for Neurodegenerative Disease Diagnosis
von: Cheng, Fangqi, et al.
Veröffentlicht: (2025)
von: Cheng, Fangqi, et al.
Veröffentlicht: (2025)
BioGait-VLM: A Tri-Modal Vision-Language-Biomechanics Framework for Interpretable Clinical Gait Assessment
von: Chen, Erdong, et al.
Veröffentlicht: (2026)
von: Chen, Erdong, et al.
Veröffentlicht: (2026)
WiseAD: Knowledge Augmented End-to-End Autonomous Driving with Vision-Language Model
von: Zhang, Songyan, et al.
Veröffentlicht: (2024)
von: Zhang, Songyan, et al.
Veröffentlicht: (2024)
Multi-view Video-Pose Pretraining for Operating Room Surgical Activity Recognition
von: Hamoud, Idris, et al.
Veröffentlicht: (2025)
von: Hamoud, Idris, et al.
Veröffentlicht: (2025)
MimicGait: A Model Agnostic approach for Occluded Gait Recognition using Correlational Knowledge Distillation
von: Gupta, Ayush, et al.
Veröffentlicht: (2025)
von: Gupta, Ayush, et al.
Veröffentlicht: (2025)
Quantitative Gait Analysis from Single RGB Videos Using a Dual-Input Transformer-Based Network
von: Dinh, Hiep, et al.
Veröffentlicht: (2025)
von: Dinh, Hiep, et al.
Veröffentlicht: (2025)
Multi-modal Representations for Fine-grained Multi-label Critical View of Safety Recognition
von: Baby, Britty, et al.
Veröffentlicht: (2025)
von: Baby, Britty, et al.
Veröffentlicht: (2025)
Endoshare: A Publicly Available, Surgeons-Friendly Solution to De-Identify and Manage Surgical Videos
von: Arboit, Lorenzo, et al.
Veröffentlicht: (2025)
von: Arboit, Lorenzo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
KRAST: Knowledge-Augmented Robotic Action Recognition with Structured Text for Vision-Language Models
von: Nguyen, Son Hai, et al.
Veröffentlicht: (2025) -
AGIR: Assessing 3D Gait Impairment with Reasoning based on LLMs
von: Wang, Diwei, et al.
Veröffentlicht: (2025) -
Procedure-Aware Surgical Video-language Pretraining with Hierarchical Knowledge Augmentation
von: Yuan, Kun, et al.
Veröffentlicht: (2024) -
fine-CLIP: Enhancing Zero-Shot Fine-Grained Surgical Action Recognition with Vision-Language Models
von: Sharma, Saurav, et al.
Veröffentlicht: (2025) -
HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition
von: Yuan, Kun, et al.
Veröffentlicht: (2024)