Stable Signer: Hierarchical Sign Language Generative Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fang, Sen, Feng, Yalin, Zhong, Hongbin, Zhang, Yanxin, Metaxas, Dimitris N. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SignVerse-2M: A Two-Million-Clip Pose-Native Universe of 55+ Sign Languages
von: Fang, Sen, et al.
Veröffentlicht: (2026)
von: Fang, Sen, et al.
Veröffentlicht: (2026)
SignX: Continuous Sign Recognition in Compact Pose-Rich Latent Space
von: Fang, Sen, et al.
Veröffentlicht: (2025)
von: Fang, Sen, et al.
Veröffentlicht: (2025)
StreamFlow: Theory, Algorithm, and Implementation for High-Efficiency Rectified Flow Generation
von: Fang, Sen, et al.
Veröffentlicht: (2025)
von: Fang, Sen, et al.
Veröffentlicht: (2025)
RAC: Rectified Flow Auto Coder
von: Fang, Sen, et al.
Veröffentlicht: (2026)
von: Fang, Sen, et al.
Veröffentlicht: (2026)
SignLLM: Sign Language Production Large Language Models
von: Fang, Sen, et al.
Veröffentlicht: (2024)
von: Fang, Sen, et al.
Veröffentlicht: (2024)
Emotion Recognition in Signers
von: Funakoshi, Kotaro, et al.
Veröffentlicht: (2025)
von: Funakoshi, Kotaro, et al.
Veröffentlicht: (2025)
New Capability to Look Up an ASL Sign from a Video Example
von: Neidle, Carol, et al.
Veröffentlicht: (2024)
von: Neidle, Carol, et al.
Veröffentlicht: (2024)
SignDiff: Diffusion Model for American Sign Language Production
von: Fang, Sen, et al.
Veröffentlicht: (2023)
von: Fang, Sen, et al.
Veröffentlicht: (2023)
Improving Continuous Sign Language Recognition with Consistency Constraints and Signer Removal
von: Zuo, Ronglai, et al.
Veröffentlicht: (2022)
von: Zuo, Ronglai, et al.
Veröffentlicht: (2022)
Diverse Signer Avatars with Manual and Non-Manual Feature Modelling for Sign Language Production
von: Lakhal, Mohamed Ilyes, et al.
Veröffentlicht: (2025)
von: Lakhal, Mohamed Ilyes, et al.
Veröffentlicht: (2025)
Large Sign Language Models: Toward 3D American Sign Language Translation
von: Zhang, Sen, et al.
Veröffentlicht: (2025)
von: Zhang, Sen, et al.
Veröffentlicht: (2025)
MLLM-as-a-Judge for Image Safety without Human Labeling
von: Wang, Zhenting, et al.
Veröffentlicht: (2024)
von: Wang, Zhenting, et al.
Veröffentlicht: (2024)
The Aftermath of DrawEduMath: Vision Language Models Underperform with Struggling Students and Misdiagnose Errors
von: Lucy, Li, et al.
Veröffentlicht: (2026)
von: Lucy, Li, et al.
Veröffentlicht: (2026)
Are Vision Language Models Cross-Cultural Theory of Mind Reasoners?
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2025)
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2025)
Modeling Intensification for Sign Language Generation: A Computational Approach
von: İnan, Mert, et al.
Veröffentlicht: (2022)
von: İnan, Mert, et al.
Veröffentlicht: (2022)
Signs as Tokens: A Retrieval-Enhanced Multilingual Sign Language Generator
von: Zuo, Ronglai, et al.
Veröffentlicht: (2024)
von: Zuo, Ronglai, et al.
Veröffentlicht: (2024)
A Unified Framework and Dataset for Assessing Societal Bias in Vision-Language Models
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024)
von: Sathe, Ashutosh, et al.
Veröffentlicht: (2024)
MM-Soc: Benchmarking Multimodal Large Language Models in Social Media Platforms
von: Jin, Yiqiao, et al.
Veröffentlicht: (2024)
von: Jin, Yiqiao, et al.
Veröffentlicht: (2024)
Large Language Models and Provenance Metadata for Determining the Relevance of Images and Videos in News Stories
von: Peterka, Tomas, et al.
Veröffentlicht: (2025)
von: Peterka, Tomas, et al.
Veröffentlicht: (2025)
The Use of Multimodal Large Language Models to Detect Objects from Thermal Images: Transportation Applications
von: Ashqar, Huthaifa I., et al.
Veröffentlicht: (2024)
von: Ashqar, Huthaifa I., et al.
Veröffentlicht: (2024)
Examining Gender and Racial Bias in Large Vision-Language Models Using a Novel Dataset of Parallel Images
von: Fraser, Kathleen C., et al.
Veröffentlicht: (2024)
von: Fraser, Kathleen C., et al.
Veröffentlicht: (2024)
Generative Sign-description Prompts with Multi-positive Contrastive Learning for Sign Language Recognition
von: Liang, Siyu, et al.
Veröffentlicht: (2025)
von: Liang, Siyu, et al.
Veröffentlicht: (2025)
Signformer is all you need: Towards Edge AI for Sign Language
von: Yang, Eta
Veröffentlicht: (2024)
von: Yang, Eta
Veröffentlicht: (2024)
MPDiT: Multi-Patch Global-to-Local Transformer Architecture For Efficient Flow Matching and Diffusion Model
von: Dao, Quan, et al.
Veröffentlicht: (2026)
von: Dao, Quan, et al.
Veröffentlicht: (2026)
Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
von: Si, Chenglei, et al.
Veröffentlicht: (2024)
von: Si, Chenglei, et al.
Veröffentlicht: (2024)
MHB: Multimodal Handshape-aware Boundary Detection for Continuous Sign Language Recognition
von: Zhao, Mingyu, et al.
Veröffentlicht: (2025)
von: Zhao, Mingyu, et al.
Veröffentlicht: (2025)
Describing Differences in Image Sets with Natural Language
von: Dunlap, Lisa, et al.
Veröffentlicht: (2023)
von: Dunlap, Lisa, et al.
Veröffentlicht: (2023)
SignDATA: Data Pipeline for Sign Language Translation
von: Chen, Kuanwei, et al.
Veröffentlicht: (2026)
von: Chen, Kuanwei, et al.
Veröffentlicht: (2026)
Identifying Implicit Social Biases in Vision-Language Models
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
von: Hamidieh, Kimia, et al.
Veröffentlicht: (2024)
Beyond Translation: Cross-Cultural Meme Transcreation with Vision-Language Models
von: Zhao, Yuming, et al.
Veröffentlicht: (2026)
von: Zhao, Yuming, et al.
Veröffentlicht: (2026)
Representing Signs as Signs: One-Shot ISLR to Facilitate Functional Sign Language Technologies
von: Vandendriessche, Toon, et al.
Veröffentlicht: (2025)
von: Vandendriessche, Toon, et al.
Veröffentlicht: (2025)
Using Sign Language Production as Data Augmentation to enhance Sign Language Translation
von: Walsh, Harry, et al.
Veröffentlicht: (2025)
von: Walsh, Harry, et al.
Veröffentlicht: (2025)
FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes
von: Chen, Liuliu, et al.
Veröffentlicht: (2026)
von: Chen, Liuliu, et al.
Veröffentlicht: (2026)
Studying and Mitigating Biases in Sign Language Understanding Models
von: Atwell, Katherine, et al.
Veröffentlicht: (2024)
von: Atwell, Katherine, et al.
Veröffentlicht: (2024)
Sign Stitching: A Novel Approach to Sign Language Production
von: Walsh, Harry, et al.
Veröffentlicht: (2024)
von: Walsh, Harry, et al.
Veröffentlicht: (2024)
Hierarchical Windowed Graph Attention Network and a Large Scale Dataset for Isolated Indian Sign Language Recognition
von: Patra, Suvajit, et al.
Veröffentlicht: (2024)
von: Patra, Suvajit, et al.
Veröffentlicht: (2024)
ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation
von: Jha, Akshita, et al.
Veröffentlicht: (2024)
von: Jha, Akshita, et al.
Veröffentlicht: (2024)
Secure Traffic Sign Recognition: An Attention-Enabled Universal Image Inpainting Mechanism against Light Patch Attacks
von: Cao, Hangcheng, et al.
Veröffentlicht: (2024)
von: Cao, Hangcheng, et al.
Veröffentlicht: (2024)
GLOS: Sign Language Generation with Temporally Aligned Gloss-Level Conditioning
von: Lee, Taeryung, et al.
Veröffentlicht: (2025)
von: Lee, Taeryung, et al.
Veröffentlicht: (2025)
Geometry-Aware Losses for Structure-Preserving Text-to-Sign Language Generation
von: Wu, Zetian, et al.
Veröffentlicht: (2025)
von: Wu, Zetian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SignVerse-2M: A Two-Million-Clip Pose-Native Universe of 55+ Sign Languages
von: Fang, Sen, et al.
Veröffentlicht: (2026) -
SignX: Continuous Sign Recognition in Compact Pose-Rich Latent Space
von: Fang, Sen, et al.
Veröffentlicht: (2025) -
StreamFlow: Theory, Algorithm, and Implementation for High-Efficiency Rectified Flow Generation
von: Fang, Sen, et al.
Veröffentlicht: (2025) -
RAC: Rectified Flow Auto Coder
von: Fang, Sen, et al.
Veröffentlicht: (2026) -
SignLLM: Sign Language Production Large Language Models
von: Fang, Sen, et al.
Veröffentlicht: (2024)