MS2SL: Multimodal Spoken Data-Driven Continuous Sign Language Production
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Jian, Wang, Wenguan, Yang, Yi, Zheng, Feng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Data-Driven Representation for Sign Language Production
by: Walsh, Harry, et al.
Published: (2024)
by: Walsh, Harry, et al.
Published: (2024)
Language Models as Continuous Self-Evolving Data Engineers
by: Wang, Peidong, et al.
Published: (2024)
by: Wang, Peidong, et al.
Published: (2024)
GLaM-Sign: Greek Language Multimodal Lip Reading with Integrated Sign Language Accessibility
by: Kouremenos, Dimitris, et al.
Published: (2025)
by: Kouremenos, Dimitris, et al.
Published: (2025)
Scaling Properties of Continuous Diffusion Spoken Language Models
by: Ramapuram, Jason, et al.
Published: (2026)
by: Ramapuram, Jason, et al.
Published: (2026)
Session-Level Spoken Language Assessment with a Multimodal Foundation Model via Multi-Target Learning
by: Lin, Hong-Yun, et al.
Published: (2025)
by: Lin, Hong-Yun, et al.
Published: (2025)
CALM: Curiosity-Driven Auditing for Large Language Models
by: Zheng, Xiang, et al.
Published: (2025)
by: Zheng, Xiang, et al.
Published: (2025)
Evaluating and Improving Continual Learning in Spoken Language Understanding
by: Yang, Muqiao, et al.
Published: (2024)
by: Yang, Muqiao, et al.
Published: (2024)
Synthesize-on-Graph: Knowledgeable Synthetic Data Generation for Continue Pre-training of Large Language Models
by: Ma, Shengjie, et al.
Published: (2025)
by: Ma, Shengjie, et al.
Published: (2025)
Evaluating Bias in Spoken Dialogue LLMs for Real-World Decisions and Recommendations
by: Wu, Yihao, et al.
Published: (2025)
by: Wu, Yihao, et al.
Published: (2025)
CNSL-bench: Benchmarking the Sign Language Understanding Capabilities of MLLMs on Chinese National Sign Language
by: Zhao, Rui, et al.
Published: (2026)
by: Zhao, Rui, et al.
Published: (2026)
On the Fallacy of Global Token Perplexity in Spoken Language Model Evaluation
by: Hsu, Chan-Jan, et al.
Published: (2026)
by: Hsu, Chan-Jan, et al.
Published: (2026)
Mutual Learning for Acoustic Matching and Dereverberation via Visual Scene-driven Diffusion
by: Ma, Jian, et al.
Published: (2024)
by: Ma, Jian, et al.
Published: (2024)
BabelBench: An Omni Benchmark for Code-Driven Analysis of Multimodal and Multistructured Data
by: Wang, Xuwu, et al.
Published: (2024)
by: Wang, Xuwu, et al.
Published: (2024)
Chronological Thinking in Full-Duplex Spoken Dialogue Language Models
by: Wu, Donghang, et al.
Published: (2025)
by: Wu, Donghang, et al.
Published: (2025)
Continuous Saudi Sign Language Recognition: A Vision Transformer Approach
by: Elhassen, Soukeina, et al.
Published: (2025)
by: Elhassen, Soukeina, et al.
Published: (2025)
Key-Point-Driven Mathematical Reasoning Distillation of Large Language Model
by: Zhu, Xunyu, et al.
Published: (2024)
by: Zhu, Xunyu, et al.
Published: (2024)
Multi-Intent Spoken Language Understanding: Methods, Trends, and Challenges
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
MultimodalHugs: Enabling Sign Language Processing in Hugging Face
by: Sant, Gerard, et al.
Published: (2025)
by: Sant, Gerard, et al.
Published: (2025)
Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners
by: Yuan, Haidong, et al.
Published: (2026)
by: Yuan, Haidong, et al.
Published: (2026)
Optimal Transport Regularization for Speech Text Alignment in Spoken Language Models
by: Xu, Wenze, et al.
Published: (2025)
by: Xu, Wenze, et al.
Published: (2025)
Stream State-tying for Sign Language Recognition
by: Ma, Jiyong, et al.
Published: (2024)
by: Ma, Jiyong, et al.
Published: (2024)
Improving Mathematical Reasoning Capabilities of Small Language Models via Feedback-Driven Distillation
by: Zhu, Xunyu, et al.
Published: (2024)
by: Zhu, Xunyu, et al.
Published: (2024)
SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents
by: Si, Shuzheng, et al.
Published: (2023)
by: Si, Shuzheng, et al.
Published: (2023)
Enhanced Sign Language Translation between American Sign Language (ASL) and Indian Sign Language (ISL) Using LLMs
by: Kumar, Malay, et al.
Published: (2024)
by: Kumar, Malay, et al.
Published: (2024)
VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing
by: Xu, Jiacheng, et al.
Published: (2026)
by: Xu, Jiacheng, et al.
Published: (2026)
Towards Data-and Knowledge-Driven Artificial Intelligence: A Survey on Neuro-Symbolic Computing
by: Wang, Wenguan, et al.
Published: (2022)
by: Wang, Wenguan, et al.
Published: (2022)
ECLM: Entity Level Language Model for Spoken Language Understanding with Chain of Intent
by: Yin, Shangjian, et al.
Published: (2024)
by: Yin, Shangjian, et al.
Published: (2024)
AutoSign: Direct Pose-to-Text Translation for Continuous Sign Language Recognition
by: Johnny, Samuel Ebimobowei, et al.
Published: (2025)
by: Johnny, Samuel Ebimobowei, et al.
Published: (2025)
Continuous Bangla Sign Language Translation: Mitigating the Expense of Gloss Annotation with the Assistance of Graph
by: Arib, Safaeid Hossain, et al.
Published: (2025)
by: Arib, Safaeid Hossain, et al.
Published: (2025)
Devising a Set of Compact and Explainable Spoken Language Feature for Screening Alzheimer's Disease
by: Li, Junan, et al.
Published: (2024)
by: Li, Junan, et al.
Published: (2024)
SignAttention: On the Interpretability of Transformer Models for Sign Language Translation
by: Bianco, Pedro Alejandro Dal, et al.
Published: (2024)
by: Bianco, Pedro Alejandro Dal, et al.
Published: (2024)
HC$^2$L: Hybrid and Cooperative Contrastive Learning for Cross-lingual Spoken Language Understanding
by: Xing, Bowen, et al.
Published: (2024)
by: Xing, Bowen, et al.
Published: (2024)
New Semantic Task for the French Spoken Language Understanding MEDIA Benchmark
by: Alavoine, Nadège, et al.
Published: (2024)
by: Alavoine, Nadège, et al.
Published: (2024)
An Adapter-Based Unified Model for Multiple Spoken Language Processing Tasks
by: Suresh, Varsha, et al.
Published: (2024)
by: Suresh, Varsha, et al.
Published: (2024)
Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language Understanding
by: Schmidt, Fabian David, et al.
Published: (2025)
by: Schmidt, Fabian David, et al.
Published: (2025)
FormalASR: End-to-End Spoken Chinese to Formal Text
by: Ning, Wanyi, et al.
Published: (2026)
by: Ning, Wanyi, et al.
Published: (2026)
2M-NER: Contrastive Learning for Multilingual and Multimodal NER with Language and Modal Fusion
by: Wang, Dongsheng, et al.
Published: (2024)
by: Wang, Dongsheng, et al.
Published: (2024)
Shape2Scene: 3D Scene Representation Learning Through Pre-training on Shape Data
by: Feng, Tuo, et al.
Published: (2024)
by: Feng, Tuo, et al.
Published: (2024)
Stance-Driven Multimodal Controlled Statement Generation: New Dataset and Task
by: Wang, Bingqian, et al.
Published: (2025)
by: Wang, Bingqian, et al.
Published: (2025)
Contrastive and Consistency Learning for Neural Noisy-Channel Model in Spoken Language Understanding
by: Kim, Suyoung, et al.
Published: (2024)
by: Kim, Suyoung, et al.
Published: (2024)
Similar Items
-
A Data-Driven Representation for Sign Language Production
by: Walsh, Harry, et al.
Published: (2024) -
Language Models as Continuous Self-Evolving Data Engineers
by: Wang, Peidong, et al.
Published: (2024) -
GLaM-Sign: Greek Language Multimodal Lip Reading with Integrated Sign Language Accessibility
by: Kouremenos, Dimitris, et al.
Published: (2025) -
Scaling Properties of Continuous Diffusion Spoken Language Models
by: Ramapuram, Jason, et al.
Published: (2026) -
Session-Level Spoken Language Assessment with a Multimodal Foundation Model via Multi-Target Learning
by: Lin, Hong-Yun, et al.
Published: (2025)