Conditional Collapse in Sign Language Production: A Diagnostic and a Scaling Argument
Fuente:
arXiv
Salvato in:
| Autori principali: | Hong, Rui, Košecká, Jana |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Toward Phonology-Guided Sign Language Motion Generation: A Diffusion Baseline and Conditioning Analysis
di: Hong, Rui, et al.
Pubblicazione: (2026)
di: Hong, Rui, et al.
Pubblicazione: (2026)
Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation
di: Hong, Rui, et al.
Pubblicazione: (2026)
di: Hong, Rui, et al.
Pubblicazione: (2026)
Gloss2Text: Sign Language Gloss translation using LLMs and Semantically Aware Label Smoothing
di: Fayyazsanavi, Pooya, et al.
Pubblicazione: (2024)
di: Fayyazsanavi, Pooya, et al.
Pubblicazione: (2024)
Towards Grounded Visual Spatial Reasoning in Multi-Modal Vision Language Models
di: Rajabi, Navid, et al.
Pubblicazione: (2023)
di: Rajabi, Navid, et al.
Pubblicazione: (2023)
Q-GroundCAM: Quantifying Grounding in Vision Language Models via GradCAM
di: Rajabi, Navid, et al.
Pubblicazione: (2024)
di: Rajabi, Navid, et al.
Pubblicazione: (2024)
PointSplat: Efficient Geometry-Driven Pruning and Transformer Refinement for 3D Gaussian Splatting
di: Tran, Anh Thuan, et al.
Pubblicazione: (2026)
di: Tran, Anh Thuan, et al.
Pubblicazione: (2026)
VarSplat: Uncertainty-aware 3D Gaussian Splatting for Robust RGB-D SLAM
di: Tran, Anh Thuan, et al.
Pubblicazione: (2026)
di: Tran, Anh Thuan, et al.
Pubblicazione: (2026)
GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs
di: Rajabi, Navid, et al.
Pubblicazione: (2024)
di: Rajabi, Navid, et al.
Pubblicazione: (2024)
TRAVEL: Training-Free Retrieval and Alignment for Vision-and-Language Navigation
di: Rajabi, Navid, et al.
Pubblicazione: (2025)
di: Rajabi, Navid, et al.
Pubblicazione: (2025)
Structured Spatial Reasoning with Open Vocabulary Object Detectors
di: Nejatishahidin, Negar, et al.
Pubblicazione: (2024)
di: Nejatishahidin, Negar, et al.
Pubblicazione: (2024)
Compositional Image-Text Matching and Retrieval by Grounding Entities
di: Vongala, Madhukar Reddy, et al.
Pubblicazione: (2025)
di: Vongala, Madhukar Reddy, et al.
Pubblicazione: (2025)
Multi-temporal Adaptive Red-Green-Blue and Long-Wave Infrared Fusion for You Only Look Once-Based Landmine Detection from Unmanned Aerial Systems
di: Gallagher, James E., et al.
Pubblicazione: (2025)
di: Gallagher, James E., et al.
Pubblicazione: (2025)
Sign-IDD: Iconicity Disentangled Diffusion for Sign Language Production
di: Tang, Shengeng, et al.
Pubblicazione: (2024)
di: Tang, Shengeng, et al.
Pubblicazione: (2024)
Uni-Sign: Toward Unified Sign Language Understanding at Scale
di: Li, Zecheng, et al.
Pubblicazione: (2025)
di: Li, Zecheng, et al.
Pubblicazione: (2025)
SignDiff: Diffusion Model for American Sign Language Production
di: Fang, Sen, et al.
Pubblicazione: (2023)
di: Fang, Sen, et al.
Pubblicazione: (2023)
SignLLM: Sign Language Production Large Language Models
di: Fang, Sen, et al.
Pubblicazione: (2024)
di: Fang, Sen, et al.
Pubblicazione: (2024)
Sign Stitching: A Novel Approach to Sign Language Production
di: Walsh, Harry, et al.
Pubblicazione: (2024)
di: Walsh, Harry, et al.
Pubblicazione: (2024)
Beyond Image-Text Matching: Verb Understanding in Multimodal Transformers Using Guided Masking
di: Beňová, Ivana, et al.
Pubblicazione: (2024)
di: Beňová, Ivana, et al.
Pubblicazione: (2024)
StgcDiff: Spatial-Temporal Graph Condition Diffusion for Sign Language Transition Generation
di: He, Jiashu, et al.
Pubblicazione: (2025)
di: He, Jiashu, et al.
Pubblicazione: (2025)
Linguistics-Vision Monotonic Consistent Network for Sign Language Production
di: Wang, Xu, et al.
Pubblicazione: (2024)
di: Wang, Xu, et al.
Pubblicazione: (2024)
Using Sign Language Production as Data Augmentation to enhance Sign Language Translation
di: Walsh, Harry, et al.
Pubblicazione: (2025)
di: Walsh, Harry, et al.
Pubblicazione: (2025)
Text-Driven Diffusion Model for Sign Language Production
di: He, Jiayi, et al.
Pubblicazione: (2025)
di: He, Jiayi, et al.
Pubblicazione: (2025)
SignSparK: Efficient Multilingual Sign Language Production via Sparse Keyframe Learning
di: Low, Jianhe, et al.
Pubblicazione: (2026)
di: Low, Jianhe, et al.
Pubblicazione: (2026)
SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation
di: Wang, Xu, et al.
Pubblicazione: (2025)
di: Wang, Xu, et al.
Pubblicazione: (2025)
Scaling Sign Language Translation
di: Zhang, Biao, et al.
Pubblicazione: (2024)
di: Zhang, Biao, et al.
Pubblicazione: (2024)
Bootstrapping Sign Language Annotations with Sign Language Models
di: Lea, Colin, et al.
Pubblicazione: (2026)
di: Lea, Colin, et al.
Pubblicazione: (2026)
Teach Me Sign: Stepwise Prompting LLM for Sign Language Production
di: An, Zhaoyi, et al.
Pubblicazione: (2025)
di: An, Zhaoyi, et al.
Pubblicazione: (2025)
Autoregressive Sign Language Production: A Gloss-Free Approach with Discrete Representations
di: Hwang, Eui Jun, et al.
Pubblicazione: (2023)
di: Hwang, Eui Jun, et al.
Pubblicazione: (2023)
CCSPNet-Joint: Efficient Joint Training Method for Traffic Sign Detection Under Extreme Conditions
di: Hong, Haoqin, et al.
Pubblicazione: (2023)
di: Hong, Haoqin, et al.
Pubblicazione: (2023)
Hybrid Autoregressive-Diffusion Model for Real-Time Sign Language Production
di: Ye, Maoxiao, et al.
Pubblicazione: (2025)
di: Ye, Maoxiao, et al.
Pubblicazione: (2025)
Isharah: A Large-Scale Multi-Scene Dataset for Continuous Sign Language Recognition
di: Alyami, Sarah, et al.
Pubblicazione: (2025)
di: Alyami, Sarah, et al.
Pubblicazione: (2025)
GLOS: Sign Language Generation with Temporally Aligned Gloss-Level Conditioning
di: Lee, Taeryung, et al.
Pubblicazione: (2025)
di: Lee, Taeryung, et al.
Pubblicazione: (2025)
Diversity-Aware Sign Language Production through a Pose Encoding Variational Autoencoder
di: Lakhal, Mohamed Ilyes, et al.
Pubblicazione: (2024)
di: Lakhal, Mohamed Ilyes, et al.
Pubblicazione: (2024)
Deep Understanding of Sign Language for Sign to Subtitle Alignment
di: Jang, Youngjoon, et al.
Pubblicazione: (2025)
di: Jang, Youngjoon, et al.
Pubblicazione: (2025)
Scaling up Multimodal Pre-training for Sign Language Understanding
di: Zhou, Wengang, et al.
Pubblicazione: (2024)
di: Zhou, Wengang, et al.
Pubblicazione: (2024)
M3T: Discrete Multi-Modal Motion Tokens for Sign Language Production
di: Symeonidis-Herzig, Alexandre, et al.
Pubblicazione: (2026)
di: Symeonidis-Herzig, Alexandre, et al.
Pubblicazione: (2026)
Bridging Sign and Spoken Languages: Pseudo Gloss Generation for Sign Language Translation
di: Guo, Jianyuan, et al.
Pubblicazione: (2025)
di: Guo, Jianyuan, et al.
Pubblicazione: (2025)
SignIT: A Comprehensive Dataset and Multimodal Analysis for Italian Sign Language Recognition
di: Micieli, Alessia, et al.
Pubblicazione: (2025)
di: Micieli, Alessia, et al.
Pubblicazione: (2025)
EmoSign: A Multimodal Dataset for Understanding Emotions in American Sign Language
di: Chua, Phoebe, et al.
Pubblicazione: (2025)
di: Chua, Phoebe, et al.
Pubblicazione: (2025)
Towards Continuous Sign Language Conversation from Isolated Signs
di: Kim, Youngmin, et al.
Pubblicazione: (2026)
di: Kim, Youngmin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Toward Phonology-Guided Sign Language Motion Generation: A Diffusion Baseline and Conditioning Analysis
di: Hong, Rui, et al.
Pubblicazione: (2026) -
Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation
di: Hong, Rui, et al.
Pubblicazione: (2026) -
Gloss2Text: Sign Language Gloss translation using LLMs and Semantically Aware Label Smoothing
di: Fayyazsanavi, Pooya, et al.
Pubblicazione: (2024) -
Towards Grounded Visual Spatial Reasoning in Multi-Modal Vision Language Models
di: Rajabi, Navid, et al.
Pubblicazione: (2023) -
Q-GroundCAM: Quantifying Grounding in Vision Language Models via GradCAM
di: Rajabi, Navid, et al.
Pubblicazione: (2024)