CorrNet+: Sign Language Recognition and Translation via Spatial-Temporal Correlation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Lianyu, Feng, Wei, Gao, Liqing, Liu, Zekang, Wan, Liang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dynamic Spatial-Temporal Aggregation for Skeleton-Aware Sign Language Recognition
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
Improving Continuous Sign Language Recognition with Adapted Image Models
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
COMMA: Co-Articulated Multi-Modal Learning
von: Hu, Lianyu, et al.
Veröffentlicht: (2023)
von: Hu, Lianyu, et al.
Veröffentlicht: (2023)
iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal Models
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
Deep Correlated Prompting for Visual Recognition with Missing Modalities
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)
Pose-Guided Fine-Grained Sign Language Video Generation
von: Shi, Tongkai, et al.
Veröffentlicht: (2024)
von: Shi, Tongkai, et al.
Veröffentlicht: (2024)
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression
von: Hu, Lianyu, et al.
Veröffentlicht: (2025)
von: Hu, Lianyu, et al.
Veröffentlicht: (2025)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
von: Zhao, Weichao, et al.
Veröffentlicht: (2024)
von: Zhao, Weichao, et al.
Veröffentlicht: (2024)
Towards Online Continuous Sign Language Recognition and Translation
von: Zuo, Ronglai, et al.
Veröffentlicht: (2024)
von: Zuo, Ronglai, et al.
Veröffentlicht: (2024)
USTM: Unified Spatial and Temporal Modeling for Continuous Sign Language Recognition
von: Hasanaath, Ahmed Abul, et al.
Veröffentlicht: (2025)
von: Hasanaath, Ahmed Abul, et al.
Veröffentlicht: (2025)
SSL-SSAW: Self-Supervised Learning with Sigmoid Self-Attention Weighting for Question-Based Sign Language Translation
von: Liu, Zekang, et al.
Veröffentlicht: (2025)
von: Liu, Zekang, et al.
Veröffentlicht: (2025)
EvSign: Sign Language Recognition and Translation with Streaming Events
von: Zhang, Pengyu, et al.
Veröffentlicht: (2024)
von: Zhang, Pengyu, et al.
Veröffentlicht: (2024)
StepNet: Spatial-temporal Part-aware Network for Isolated Sign Language Recognition
von: Shen, Xiaolong, et al.
Veröffentlicht: (2022)
von: Shen, Xiaolong, et al.
Veröffentlicht: (2022)
Stack Transformer Based Spatial-Temporal Attention Model for Dynamic Sign Language and Fingerspelling Recognition
von: Hirooka, Koki, et al.
Veröffentlicht: (2025)
von: Hirooka, Koki, et al.
Veröffentlicht: (2025)
Multi-Stream Keypoint Attention Network for Sign Language Recognition and Translation
von: Guan, Mo, et al.
Veröffentlicht: (2024)
von: Guan, Mo, et al.
Veröffentlicht: (2024)
CorrCLIP: Reconstructing Patch Correlations in CLIP for Open-Vocabulary Semantic Segmentation
von: Zhang, Dengke, et al.
Veröffentlicht: (2024)
von: Zhang, Dengke, et al.
Veröffentlicht: (2024)
Video-based Sign Language Recognition without Temporal Segmentation
von: Huang, Jie, et al.
Veröffentlicht: (2018)
von: Huang, Jie, et al.
Veröffentlicht: (2018)
TCNet: Continuous Sign Language Recognition from Trajectories and Correlated Regions
von: Lu, Hui, et al.
Veröffentlicht: (2024)
von: Lu, Hui, et al.
Veröffentlicht: (2024)
Bengali Sign Language Recognition through Hand Pose Estimation using Multi-Branch Spatial-Temporal Attention Model
von: Miah, Abu Saleh Musa, et al.
Veröffentlicht: (2024)
von: Miah, Abu Saleh Musa, et al.
Veröffentlicht: (2024)
Direct Translation between Sign Languages
von: Wu, Zetian, et al.
Veröffentlicht: (2026)
von: Wu, Zetian, et al.
Veröffentlicht: (2026)
Generative Sign-description Prompts with Multi-positive Contrastive Learning for Sign Language Recognition
von: Liang, Siyu, et al.
Veröffentlicht: (2025)
von: Liang, Siyu, et al.
Veröffentlicht: (2025)
AutoSign: Direct Pose-to-Text Translation for Continuous Sign Language Recognition
von: Johnny, Samuel Ebimobowei, et al.
Veröffentlicht: (2025)
von: Johnny, Samuel Ebimobowei, et al.
Veröffentlicht: (2025)
RVLF: A Reinforcing Vision-Language Framework for Gloss-Free Sign Language Translation
von: Rao, Zhi, et al.
Veröffentlicht: (2025)
von: Rao, Zhi, et al.
Veröffentlicht: (2025)
LLMs are Good Sign Language Translators
von: Gong, Jia, et al.
Veröffentlicht: (2024)
von: Gong, Jia, et al.
Veröffentlicht: (2024)
Hierarchical Sub-action Tree for Continuous Sign Language Recognition
von: Yang, Dejie, et al.
Veröffentlicht: (2025)
von: Yang, Dejie, et al.
Veröffentlicht: (2025)
C${^2}$RL: Content and Context Representation Learning for Gloss-free Sign Language Translation and Retrieval
von: Chen, Zhigang, et al.
Veröffentlicht: (2024)
von: Chen, Zhigang, et al.
Veröffentlicht: (2024)
Cross-Modal Consistency Learning for Sign Language Recognition
von: Wu, Kepeng, et al.
Veröffentlicht: (2025)
von: Wu, Kepeng, et al.
Veröffentlicht: (2025)
An Efficient Sign Language Translation Using Spatial Configuration and Motion Dynamics with LLMs
von: Hwang, Eui Jun, et al.
Veröffentlicht: (2024)
von: Hwang, Eui Jun, et al.
Veröffentlicht: (2024)
Bridging Sign and Spoken Languages: Pseudo Gloss Generation for Sign Language Translation
von: Guo, Jianyuan, et al.
Veröffentlicht: (2025)
von: Guo, Jianyuan, et al.
Veröffentlicht: (2025)
GaitASMS: Gait Recognition by Adaptive Structured Spatial Representation and Multi-Scale Temporal Aggregation
von: Sun, Yan, et al.
Veröffentlicht: (2023)
von: Sun, Yan, et al.
Veröffentlicht: (2023)
LLaVA-SLT: Visual Language Tuning for Sign Language Translation
von: Liang, Han, et al.
Veröffentlicht: (2024)
von: Liang, Han, et al.
Veröffentlicht: (2024)
StgcDiff: Spatial-Temporal Graph Condition Diffusion for Sign Language Transition Generation
von: He, Jiashu, et al.
Veröffentlicht: (2025)
von: He, Jiashu, et al.
Veröffentlicht: (2025)
SignDATA: Data Pipeline for Sign Language Translation
von: Chen, Kuanwei, et al.
Veröffentlicht: (2026)
von: Chen, Kuanwei, et al.
Veröffentlicht: (2026)
Diverse Sign Language Translation
von: Shen, Xin, et al.
Veröffentlicht: (2024)
von: Shen, Xin, et al.
Veröffentlicht: (2024)
Vision-Language Model IP Protection via Prompt-based Learning
von: Wang, Lianyu, et al.
Veröffentlicht: (2025)
von: Wang, Lianyu, et al.
Veröffentlicht: (2025)
STARK: Spatio-Temporal Attention for Representation of Keypoints for Continuous Sign Language Recognition
von: Patra, Suvajit, et al.
Veröffentlicht: (2026)
von: Patra, Suvajit, et al.
Veröffentlicht: (2026)
Lost in Translation, Found in Embeddings: Sign Language Translation and Alignment
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
von: Jang, Youngjoon, et al.
Veröffentlicht: (2025)
TF-CorrNet: Leveraging Spatial Correlation for Continuous Speech Separation
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2025)
von: Shin, Ui-Hyeop, et al.
Veröffentlicht: (2025)
SNN-Driven Multimodal Human Action Recognition via Sparse Spatial-Temporal Data Fusion
von: Zheng, Naichuan, et al.
Veröffentlicht: (2025)
von: Zheng, Naichuan, et al.
Veröffentlicht: (2025)
CorrDiff: Adaptive Delay-aware Detector with Temporal Cue Inputs for Real-time Object Detection
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Dynamic Spatial-Temporal Aggregation for Skeleton-Aware Sign Language Recognition
von: Hu, Lianyu, et al.
Veröffentlicht: (2024) -
Improving Continuous Sign Language Recognition with Adapted Image Models
von: Hu, Lianyu, et al.
Veröffentlicht: (2024) -
COMMA: Co-Articulated Multi-Modal Learning
von: Hu, Lianyu, et al.
Veröffentlicht: (2023) -
iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal Models
von: Hu, Lianyu, et al.
Veröffentlicht: (2024) -
Deep Correlated Prompting for Visual Recognition with Missing Modalities
von: Hu, Lianyu, et al.
Veröffentlicht: (2024)