Wav2Gloss: Generating Interlinear Glossed Text from Speech
Fuente:
arXiv
Saved in:
| Main Authors: | He, Taiqi, Choi, Kwanghee, Tjuatja, Lindia, Robinson, Nathaniel R., Shi, Jiatong, Watanabe, Shinji, Neubig, Graham, Mortensen, David R., Levin, Lori |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GlossLM: A Massively Multilingual Corpus and Pretrained Model for Interlinear Glossed Text
by: Ginn, Michael, et al.
Published: (2024)
by: Ginn, Michael, et al.
Published: (2024)
CWoMP: Morpheme Representation Learning for Interlinear Glossing
by: Alper, Morris, et al.
Published: (2026)
by: Alper, Morris, et al.
Published: (2026)
Massively Multilingual Joint Segmentation and Glossing
by: Ginn, Michael, et al.
Published: (2026)
by: Ginn, Michael, et al.
Published: (2026)
BehaviorBox: Automated Discovery of Fine-Grained Performance Differences Between Language Models
by: Tjuatja, Lindia, et al.
Published: (2025)
by: Tjuatja, Lindia, et al.
Published: (2025)
What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length
by: Tjuatja, Lindia, et al.
Published: (2024)
by: Tjuatja, Lindia, et al.
Published: (2024)
Leveraging Allophony in Self-Supervised Speech Models for Atypical Pronunciation Assessment
by: Choi, Kwanghee, et al.
Published: (2025)
by: Choi, Kwanghee, et al.
Published: (2025)
Do LLMs exhibit human-like response biases? A case study in survey design
by: Tjuatja, Lindia, et al.
Published: (2023)
by: Tjuatja, Lindia, et al.
Published: (2023)
CMULAB: An Open-Source Framework for Training and Deployment of Natural Language Processing Models
by: Sheikh, Zaid, et al.
Published: (2024)
by: Sheikh, Zaid, et al.
Published: (2024)
Gloss2Text: Sign Language Gloss translation using LLMs and Semantically Aware Label Smoothing
by: Fayyazsanavi, Pooya, et al.
Published: (2024)
by: Fayyazsanavi, Pooya, et al.
Published: (2024)
Constructions Are So Difficult That Even Large Language Models Get Them Right for the Wrong Reasons
by: Zhou, Shijia, et al.
Published: (2024)
by: Zhou, Shijia, et al.
Published: (2024)
Beyond Gloss: A Hand-Centric Framework for Gloss-Free Sign Language Translation
by: Asasi, Sobhan, et al.
Published: (2025)
by: Asasi, Sobhan, et al.
Published: (2025)
On-device Streaming Discrete Speech Units
by: Choi, Kwanghee, et al.
Published: (2025)
by: Choi, Kwanghee, et al.
Published: (2025)
On the Effects of Heterogeneous Data Sources on Speech-to-Text Foundation Models
by: Tian, Jinchuan, et al.
Published: (2024)
by: Tian, Jinchuan, et al.
Published: (2024)
What do Language Models Learn and When? The Implicit Curriculum Hypothesis
by: Liu, Emmy, et al.
Published: (2026)
by: Liu, Emmy, et al.
Published: (2026)
Sign Language Gloss Embedding Models
by: McGill, Euan, et al.
Published: (2024)
by: McGill, Euan, et al.
Published: (2024)
Predicting Perceived Gloss: Do Weak Labels Suffice?
by: Guerrero-Viu, Julia, et al.
Published: (2024)
by: Guerrero-Viu, Julia, et al.
Published: (2024)
Breaking the Silence: A Dataset and Benchmark for Bangla Text-to-Gloss Translation
by: Abdullah, Sharif Mohammad, et al.
Published: (2025)
by: Abdullah, Sharif Mohammad, et al.
Published: (2025)
Artist‐Inator: Text‐based, Gloss‐aware Non‐photorealistic Stylization
by: J. Daniel Subias, et al.
Published: (2025)
by: J. Daniel Subias, et al.
Published: (2025)
POWSM: A Phonetic Open Whisper-Style Speech Foundation Model
by: Li, Chin-Jou, et al.
Published: (2025)
by: Li, Chin-Jou, et al.
Published: (2025)
Embedded Translations for Low-resource Automated Glossing
by: Yang, Changbing, et al.
Published: (2024)
by: Yang, Changbing, et al.
Published: (2024)
Style-Aware Gloss Control for Generative Non-Photorealistic Rendering
by: Jimenez-Navarro, Santiago, et al.
Published: (2026)
by: Jimenez-Navarro, Santiago, et al.
Published: (2026)
Text2Sign Diffusion: A Generative Approach for Gloss-Free Sign Language Production
by: Feng, Liqian, et al.
Published: (2025)
by: Feng, Liqian, et al.
Published: (2025)
[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic
by: Choi, Kwanghee, et al.
Published: (2026)
by: Choi, Kwanghee, et al.
Published: (2026)
Uni-VERSA: Versatile Speech Assessment with a Unified Network
by: Shi, Jiatong, et al.
Published: (2025)
by: Shi, Jiatong, et al.
Published: (2025)
Self-Supervised Speech Models Encode Phonetic Context via Position-dependent Orthogonal Subspaces
by: Choi, Kwanghee, et al.
Published: (2026)
by: Choi, Kwanghee, et al.
Published: (2026)
Robust Generalization Strategies for Morpheme Glossing in an Endangered Language Documentation Context
by: Ginn, Michael, et al.
Published: (2023)
by: Ginn, Michael, et al.
Published: (2023)
GLOS: Sign Language Generation with Temporally Aligned Gloss-Level Conditioning
by: Lee, Taeryung, et al.
Published: (2025)
by: Lee, Taeryung, et al.
Published: (2025)
Philosophy [IO Islamic 509] Glosses on الرسالة الشمسية
by: ʿAlī ibn Muḥammad, al-Sayyid al-Sharīf Jurjānī, 1340-1413
Published: (2021)
by: ʿAlī ibn Muḥammad, al-Sayyid al-Sharīf Jurjānī, 1340-1413
Published: (2021)
High‐Gloss SVBRDF Capture Using Bounce Light
by: Tomáš Iser, et al.
Published: (2026)
by: Tomáš Iser, et al.
Published: (2026)
Gloss-Free Sign Language Translation: An Unbiased Evaluation of Progress in the Field
by: Sincan, Ozge Mercanoglu, et al.
Published: (2026)
by: Sincan, Ozge Mercanoglu, et al.
Published: (2026)
An Empirical Recipe for Universal Phone Recognition
by: Bharadwaj, Shikhar, et al.
Published: (2026)
by: Bharadwaj, Shikhar, et al.
Published: (2026)
Bridge to Non-Barrier Communication: Gloss-Prompted Fine-grained Cued Speech Gesture Generation with Diffusion Model
by: Lei, Wentao, et al.
Published: (2024)
by: Lei, Wentao, et al.
Published: (2024)
Self-Supervised Speech Representations are More Phonetic than Semantic
by: Choi, Kwanghee, et al.
Published: (2024)
by: Choi, Kwanghee, et al.
Published: (2024)
Discrete Speech Unit Extraction via Independent Component Analysis
by: Nakamura, Tomohiko, et al.
Published: (2025)
by: Nakamura, Tomohiko, et al.
Published: (2025)
Bridging Sign and Spoken Languages: Pseudo Gloss Generation for Sign Language Translation
by: Guo, Jianyuan, et al.
Published: (2025)
by: Guo, Jianyuan, et al.
Published: (2025)
Selective Contrastive Learning For Gloss Free Sign Language Translation
by: Lai, Changhao, et al.
Published: (2026)
by: Lai, Changhao, et al.
Published: (2026)
Leveraging the Power of MLLMs for Gloss-Free Sign Language Translation
by: Kim, Jungeun, et al.
Published: (2024)
by: Kim, Jungeun, et al.
Published: (2024)
Hierarchical Feature Alignment for Gloss-Free Sign Language Translation
by: Asasi, Sobhan, et al.
Published: (2025)
by: Asasi, Sobhan, et al.
Published: (2025)
SAGE: Segment-Aware Gloss-Free Encoding for Token-Efficient Sign Language Translation
by: Low, JianHe, et al.
Published: (2025)
by: Low, JianHe, et al.
Published: (2025)
Spec-Gloss Surfels and Normal-Diffuse Priors for Relightable Glossy Objects
by: Kouros, Georgios, et al.
Published: (2025)
by: Kouros, Georgios, et al.
Published: (2025)
Similar Items
-
GlossLM: A Massively Multilingual Corpus and Pretrained Model for Interlinear Glossed Text
by: Ginn, Michael, et al.
Published: (2024) -
CWoMP: Morpheme Representation Learning for Interlinear Glossing
by: Alper, Morris, et al.
Published: (2026) -
Massively Multilingual Joint Segmentation and Glossing
by: Ginn, Michael, et al.
Published: (2026) -
BehaviorBox: Automated Discovery of Fine-Grained Performance Differences Between Language Models
by: Tjuatja, Lindia, et al.
Published: (2025) -
What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length
by: Tjuatja, Lindia, et al.
Published: (2024)