SHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Gueuwou, Shester, Du, Xiaodan, Shakhnarovich, Greg, Livescu, Karen, Liu, Alexander H. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
by: Gueuwou, Shester, et al.
Published: (2024)
by: Gueuwou, Shester, et al.
Published: (2024)
Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs
by: Karabüklü, Serpil, et al.
Published: (2026)
by: Karabüklü, Serpil, et al.
Published: (2026)
Cross-Modal Taxonomic Generalization in (Vision-) Language Models
by: Xu, Tianyang, et al.
Published: (2026)
by: Xu, Tianyang, et al.
Published: (2026)
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
by: Du, Xiaodan, et al.
Published: (2023)
by: Du, Xiaodan, et al.
Published: (2023)
On the Predictive Power of Representation Dispersion in Language Models
by: Li, Yanhong, et al.
Published: (2025)
by: Li, Yanhong, et al.
Published: (2025)
Teaching an Agent to Sketch One Part at a Time
by: Du, Xiaodan, et al.
Published: (2026)
by: Du, Xiaodan, et al.
Published: (2026)
Self-Supervised Speech Representations are More Phonetic than Semantic
by: Choi, Kwanghee, et al.
Published: (2024)
by: Choi, Kwanghee, et al.
Published: (2024)
SSL-SLR: Self-Supervised Representation Learning for Sign Language Recognition
by: Madjoukeng, Ariel Basso, et al.
Published: (2025)
by: Madjoukeng, Ariel Basso, et al.
Published: (2025)
SignRep: Enhancing Self-Supervised Sign Representations
by: Wong, Ryan, et al.
Published: (2025)
by: Wong, Ryan, et al.
Published: (2025)
Transcribe, Translate, or Transliterate: An Investigation of Intermediate Representations in Spoken Language Models
by: Ògúnrèmí, Tolúlopé, et al.
Published: (2025)
by: Ògúnrèmí, Tolúlopé, et al.
Published: (2025)
Self-Supervised Representation Learning with Spatial-Temporal Consistency for Sign Language Recognition
by: Zhao, Weichao, et al.
Published: (2024)
by: Zhao, Weichao, et al.
Published: (2024)
What Do Self-Supervised Speech Models Know About Words?
by: Pasad, Ankita, et al.
Published: (2023)
by: Pasad, Ankita, et al.
Published: (2023)
Flow-SLM: Joint Learning of Linguistic and Acoustic Information for Spoken Language Modeling
by: Chou, Ju-Chieh, et al.
Published: (2025)
by: Chou, Ju-Chieh, et al.
Published: (2025)
DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding
by: Shon, Suwon, et al.
Published: (2024)
by: Shon, Suwon, et al.
Published: (2024)
Chunk-Distilled Language Modeling
by: Li, Yanhong, et al.
Published: (2024)
by: Li, Yanhong, et al.
Published: (2024)
SignMAE: Segmentation-Driven Self-Supervised Learning for Sign Language Recognition
by: Xie, Kunyuan, et al.
Published: (2026)
by: Xie, Kunyuan, et al.
Published: (2026)
Meta-Representational Predictive Coding: Biomimetic Self-Supervised Learning
by: Ororbia, Alexander, et al.
Published: (2025)
by: Ororbia, Alexander, et al.
Published: (2025)
Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion
by: Guizilini, Vitor, et al.
Published: (2025)
by: Guizilini, Vitor, et al.
Published: (2025)
Alpha Invariance: On Inverse Scaling Between Distance and Volume Density in Neural Radiance Fields
by: Ahn, Joshua, et al.
Published: (2024)
by: Ahn, Joshua, et al.
Published: (2024)
Robust Human Trajectory Prediction via Self-Supervised Skeleton Representation Learning
by: Arashima, Taishu, et al.
Published: (2026)
by: Arashima, Taishu, et al.
Published: (2026)
LoopDraw: a Loop-Based Autoregressive Model for Shape Synthesis and Editing
by: Dinh, Nam Anh, et al.
Published: (2022)
by: Dinh, Nam Anh, et al.
Published: (2022)
Stream State-tying for Sign Language Recognition
by: Ma, Jiyong, et al.
Published: (2024)
by: Ma, Jiyong, et al.
Published: (2024)
DinoSR: Self-Distillation and Online Clustering for Self-supervised Speech Representation Learning
by: Liu, Alexander H., et al.
Published: (2023)
by: Liu, Alexander H., et al.
Published: (2023)
EvRepSL: Event-Stream Representation via Self-Supervised Learning for Event-Based Vision
by: Qu, Qiang, et al.
Published: (2024)
by: Qu, Qiang, et al.
Published: (2024)
AV2Wav: Diffusion-Based Re-synthesis from Continuous Self-supervised Features for Audio-Visual Speech Enhancement
by: Chou, Ju-Chieh, et al.
Published: (2023)
by: Chou, Ju-Chieh, et al.
Published: (2023)
Supervised Contrastive Frame Aggregation for Video Representation Learning
by: Chowdhury, Shaif, et al.
Published: (2025)
by: Chowdhury, Shaif, et al.
Published: (2025)
Expand BERT Representation with Visual Information via Grounded Language Learning with Multimodal Partial Alignment
by: Nguyen, Cong-Duy, et al.
Published: (2023)
by: Nguyen, Cong-Duy, et al.
Published: (2023)
EvSign: Sign Language Recognition and Translation with Streaming Events
by: Zhang, Pengyu, et al.
Published: (2024)
by: Zhang, Pengyu, et al.
Published: (2024)
Towards Robust Speech Representation Learning for Thousands of Languages
by: Chen, William, et al.
Published: (2024)
by: Chen, William, et al.
Published: (2024)
Structured Tree Alignment for Evaluation of (Speech) Constituency Parsing
by: Shi, Freda, et al.
Published: (2024)
by: Shi, Freda, et al.
Published: (2024)
Multi-Stream Keypoint Attention Network for Sign Language Recognition and Translation
by: Guan, Mo, et al.
Published: (2024)
by: Guan, Mo, et al.
Published: (2024)
SSL-SSAW: Self-Supervised Learning with Sigmoid Self-Attention Weighting for Question-Based Sign Language Translation
by: Liu, Zekang, et al.
Published: (2025)
by: Liu, Zekang, et al.
Published: (2025)
OmniShape: Zero-Shot Multi-Hypothesis Shape and Pose Estimation in the Real World
by: Liu, Katherine, et al.
Published: (2025)
by: Liu, Katherine, et al.
Published: (2025)
Supervised and Contrastive Self-Supervised In-Domain Representation Learning for Dense Prediction Problems in Remote Sensing
by: Ghanbarzade, Ali, et al.
Published: (2023)
by: Ghanbarzade, Ali, et al.
Published: (2023)
Self-Supervised Representation Learning via Hyperspherical Density Shaping
by: Rodríguez-Betancourt, Esteban, et al.
Published: (2026)
by: Rodríguez-Betancourt, Esteban, et al.
Published: (2026)
UniBERT: Adversarial Training for Language-Universal Representations
by: Avram, Andrei-Marius, et al.
Published: (2025)
by: Avram, Andrei-Marius, et al.
Published: (2025)
Clustering via Self-Supervised Diffusion
by: Uziel, Roy, et al.
Published: (2025)
by: Uziel, Roy, et al.
Published: (2025)
On the Discriminability of Self-Supervised Representation Learning
by: Song, Zeen, et al.
Published: (2024)
by: Song, Zeen, et al.
Published: (2024)
C${^2}$RL: Content and Context Representation Learning for Gloss-free Sign Language Translation and Retrieval
by: Chen, Zhigang, et al.
Published: (2024)
by: Chen, Zhigang, et al.
Published: (2024)
Point2Vec for Self-Supervised Representation Learning on Point Clouds
by: Knaebel, Karim, et al.
Published: (2023)
by: Knaebel, Karim, et al.
Published: (2023)
Similar Items
-
SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
by: Gueuwou, Shester, et al.
Published: (2024) -
Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs
by: Karabüklü, Serpil, et al.
Published: (2026) -
Cross-Modal Taxonomic Generalization in (Vision-) Language Models
by: Xu, Tianyang, et al.
Published: (2026) -
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
by: Du, Xiaodan, et al.
Published: (2023) -
On the Predictive Power of Representation Dispersion in Language Models
by: Li, Yanhong, et al.
Published: (2025)