Lost in Translation, Found in Embeddings: Sign Language Translation and Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Jang, Youngjoon, Momeni, Liliane, Jiang, Zifan, Chung, Joon Son, Varol, Gül, Zisserman, Andrew |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lost in Translation, Found in Context: Sign Language Translation with Contextual Cues
by: Jang, Youngjoon, et al.
Published: (2025)
by: Jang, Youngjoon, et al.
Published: (2025)
Segment, Embed, and Align: A Universal Recipe for Aligning Subtitles to Signing
by: Jiang, Zifan, et al.
Published: (2025)
by: Jiang, Zifan, et al.
Published: (2025)
Deep Understanding of Sign Language for Sign to Subtitle Alignment
by: Jang, Youngjoon, et al.
Published: (2025)
by: Jang, Youngjoon, et al.
Published: (2025)
A Tale of Two Languages: Large-Vocabulary Continuous Sign Language Recognition from Spoken Language Supervision
by: Raude, Charles, et al.
Published: (2024)
by: Raude, Charles, et al.
Published: (2024)
AVCD: Mitigating Hallucinations in Audio-Visual Large Language Models through Contrastive Decoding
by: Jung, Chaeyoung, et al.
Published: (2025)
by: Jung, Chaeyoung, et al.
Published: (2025)
Test-Time Augmentation for Pose-invariant Face Recognition
by: Jung, Jaemin, et al.
Published: (2025)
by: Jung, Jaemin, et al.
Published: (2025)
On the Nature of Attention Sink that Shapes Decoding Strategy in Omni-LLMs
by: Yoo, Suho, et al.
Published: (2026)
by: Yoo, Suho, et al.
Published: (2026)
Fork-Merge Decoding: Enhancing Multimodal Understanding in Audio-Visual Large Language Models
by: Jung, Chaeyoung, et al.
Published: (2025)
by: Jung, Chaeyoung, et al.
Published: (2025)
AutoAD III: The Prequel -- Back to the Pixels
by: Han, Tengda, et al.
Published: (2024)
by: Han, Tengda, et al.
Published: (2024)
Let Me Finish My Sentence: Video Temporal Grounding with Holistic Text Understanding
by: Woo, Jongbhin, et al.
Published: (2024)
by: Woo, Jongbhin, et al.
Published: (2024)
Hierarchical Feature Alignment for Gloss-Free Sign Language Translation
by: Asasi, Sobhan, et al.
Published: (2025)
by: Asasi, Sobhan, et al.
Published: (2025)
AutoAD-Zero: A Training-Free Framework for Zero-Shot Audio Description
by: Xie, Junyu, et al.
Published: (2024)
by: Xie, Junyu, et al.
Published: (2024)
Recognising BSL Fingerspelling in Continuous Signing Sequences
by: Chan, Alyssa, et al.
Published: (2026)
by: Chan, Alyssa, et al.
Published: (2026)
Personalizing Retrieval using Joint Embeddings or "the Return of Fluffy"
by: Korbar, Bruno, et al.
Published: (2025)
by: Korbar, Bruno, et al.
Published: (2025)
Diverse Sign Language Translation
by: Shen, Xin, et al.
Published: (2024)
by: Shen, Xin, et al.
Published: (2024)
Shot-by-Shot: Film-Grammar-Aware Training-Free Audio Description Generation
by: Xie, Junyu, et al.
Published: (2025)
by: Xie, Junyu, et al.
Published: (2025)
EvSign: Sign Language Recognition and Translation with Streaming Events
by: Zhang, Pengyu, et al.
Published: (2024)
by: Zhang, Pengyu, et al.
Published: (2024)
SignDATA: Data Pipeline for Sign Language Translation
by: Chen, Kuanwei, et al.
Published: (2026)
by: Chen, Kuanwei, et al.
Published: (2026)
Text-Driven 3D Hand Motion Generation from Sign Language Data
by: Bensabath, Léore, et al.
Published: (2025)
by: Bensabath, Léore, et al.
Published: (2025)
New keypoint-based approach for recognising British Sign Language (BSL) from sequences
by: Deb, Oishi, et al.
Published: (2024)
by: Deb, Oishi, et al.
Published: (2024)
Direct Translation between Sign Languages
by: Wu, Zetian, et al.
Published: (2026)
by: Wu, Zetian, et al.
Published: (2026)
Fingerspelling within Sign Language Translation
by: Tanzer, Garrett
Published: (2024)
by: Tanzer, Garrett
Published: (2024)
LLMs are Good Sign Language Translators
by: Gong, Jia, et al.
Published: (2024)
by: Gong, Jia, et al.
Published: (2024)
Lost in Translation? Vocabulary Alignment for Source-Free Adaptation in Open-Vocabulary Semantic Segmentation
by: Mazzucco, Silvio, et al.
Published: (2025)
by: Mazzucco, Silvio, et al.
Published: (2025)
Saudi Sign Language Translation Using T5
by: Alhejab, Ali, et al.
Published: (2025)
by: Alhejab, Ali, et al.
Published: (2025)
More than a Moment: Towards Coherent Sequences of Audio Descriptions
by: Khandelwal, Eshika, et al.
Published: (2025)
by: Khandelwal, Eshika, et al.
Published: (2025)
Bridging Sign and Spoken Languages: Pseudo Gloss Generation for Sign Language Translation
by: Guo, Jianyuan, et al.
Published: (2025)
by: Guo, Jianyuan, et al.
Published: (2025)
Scaling Sign Language Translation
by: Zhang, Biao, et al.
Published: (2024)
by: Zhang, Biao, et al.
Published: (2024)
Segmentation-before-Staining Improves Structural Fidelity in Virtual IHC-to-Multiplex IF Translation
by: Lee, Junhyeok, et al.
Published: (2026)
by: Lee, Junhyeok, et al.
Published: (2026)
Pose-Based Sign Language Appearance Transfer
by: Moryossef, Amit, et al.
Published: (2024)
by: Moryossef, Amit, et al.
Published: (2024)
Towards Online Continuous Sign Language Recognition and Translation
by: Zuo, Ronglai, et al.
Published: (2024)
by: Zuo, Ronglai, et al.
Published: (2024)
Spatio-temporal Sign Language Representation and Translation
by: Hamidullah, Yasser, et al.
Published: (2025)
by: Hamidullah, Yasser, et al.
Published: (2025)
American Sign Language Video to Text Translation
by: Roy, Parsheeta, et al.
Published: (2024)
by: Roy, Parsheeta, et al.
Published: (2024)
A Cross-Dataset Study for Text-based 3D Human Motion Retrieval
by: Bensabath, Léore, et al.
Published: (2024)
by: Bensabath, Léore, et al.
Published: (2024)
Learning text-to-video retrieval from image captioning
by: Ventura, Lucas, et al.
Published: (2024)
by: Ventura, Lucas, et al.
Published: (2024)
Using Sign Language Production as Data Augmentation to enhance Sign Language Translation
by: Walsh, Harry, et al.
Published: (2025)
by: Walsh, Harry, et al.
Published: (2025)
AutoSign: Direct Pose-to-Text Translation for Continuous Sign Language Recognition
by: Johnny, Samuel Ebimobowei, et al.
Published: (2025)
by: Johnny, Samuel Ebimobowei, et al.
Published: (2025)
Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation
by: Wong, Ryan, et al.
Published: (2024)
by: Wong, Ryan, et al.
Published: (2024)
Revisiting Misalignment in Multispectral Pedestrian Detection: A Language-Driven Approach for Cross-modal Alignment Fusion
by: Kim, Taeheon, et al.
Published: (2024)
by: Kim, Taeheon, et al.
Published: (2024)
Think in Latent Thoughts: A New Paradigm for Gloss-Free Sign Language Translation
by: Jiang, Yiyang, et al.
Published: (2026)
by: Jiang, Yiyang, et al.
Published: (2026)
Similar Items
-
Lost in Translation, Found in Context: Sign Language Translation with Contextual Cues
by: Jang, Youngjoon, et al.
Published: (2025) -
Segment, Embed, and Align: A Universal Recipe for Aligning Subtitles to Signing
by: Jiang, Zifan, et al.
Published: (2025) -
Deep Understanding of Sign Language for Sign to Subtitle Alignment
by: Jang, Youngjoon, et al.
Published: (2025) -
A Tale of Two Languages: Large-Vocabulary Continuous Sign Language Recognition from Spoken Language Supervision
by: Raude, Charles, et al.
Published: (2024) -
AVCD: Mitigating Hallucinations in Audio-Visual Large Language Models through Contrastive Decoding
by: Jung, Chaeyoung, et al.
Published: (2025)