Towards Continuous Sign Language Conversation from Isolated Signs
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Youngmin, Choo, Kyobin, Park, Jiwoo, Kim, Minseo, Kim, Chanyoung, Kim, Junhyeok, Hwang, Seong Jae |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Interpreting Attention Heads for Image-to-Text Information Flow in Large Vision-Language Models
di: Kim, Jinyeong, et al.
Pubblicazione: (2025)
di: Kim, Jinyeong, et al.
Pubblicazione: (2025)
Disentangling Disentangled Representations: Towards Improved Latent Units via Diffusion Models
di: Jun, Youngjun, et al.
Pubblicazione: (2024)
di: Jun, Youngjun, et al.
Pubblicazione: (2024)
Delaunay Canopy: Building Wireframe Reconstruction from Airborne LiDAR Point Clouds via Delaunay Graph
di: Kim, Donghyun, et al.
Pubblicazione: (2026)
di: Kim, Donghyun, et al.
Pubblicazione: (2026)
Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding
di: Kang, Seil, et al.
Pubblicazione: (2025)
di: Kang, Seil, et al.
Pubblicazione: (2025)
CoBra: Complementary Branch Fusing Class and Semantic Knowledge for Robust Weakly Supervised Semantic Segmentation
di: Han, Woojung, et al.
Pubblicazione: (2024)
di: Han, Woojung, et al.
Pubblicazione: (2024)
See What You Are Told: Visual Attention Sink in Large Multimodal Models
di: Kang, Seil, et al.
Pubblicazione: (2025)
di: Kang, Seil, et al.
Pubblicazione: (2025)
FEAST: Fully Connected Expressive Attention for Spatial Transcriptomics
di: Jeong, Taejin, et al.
Pubblicazione: (2026)
di: Jeong, Taejin, et al.
Pubblicazione: (2026)
Fourier Decomposition for Explicit Representation of 3D Point Cloud Attributes
di: Kim, Donghyun, et al.
Pubblicazione: (2025)
di: Kim, Donghyun, et al.
Pubblicazione: (2025)
Mono-Modalizing Extremely Heterogeneous Multi-Modal Medical Image Registration
di: Choo, Kyobin, et al.
Pubblicazione: (2025)
di: Choo, Kyobin, et al.
Pubblicazione: (2025)
EAGLE: Eigen Aggregation Learning for Object-Centric Unsupervised Semantic Segmentation
di: Kim, Chanyoung, et al.
Pubblicazione: (2024)
di: Kim, Chanyoung, et al.
Pubblicazione: (2024)
Spatial Transport Optimization by Repositioning Attention Map for Training-Free Text-to-Image Synthesis
di: Han, Woojung, et al.
Pubblicazione: (2025)
di: Han, Woojung, et al.
Pubblicazione: (2025)
Rethinking Graph Convolution for 2D-to-3D Hand Pose Lifting
di: Kim, Chanyoung, et al.
Pubblicazione: (2026)
di: Kim, Chanyoung, et al.
Pubblicazione: (2026)
Real-Time Visual Attribution Streaming in Thinking Model
di: Kang, Seil, et al.
Pubblicazione: (2026)
di: Kang, Seil, et al.
Pubblicazione: (2026)
Interpreting vision transformers via residual replacement model
di: Kim, Jinyeong, et al.
Pubblicazione: (2025)
di: Kim, Jinyeong, et al.
Pubblicazione: (2025)
Advancing Text-Driven Chest X-Ray Generation with Policy-Based Reinforcement Learning
di: Han, Woojung, et al.
Pubblicazione: (2024)
di: Han, Woojung, et al.
Pubblicazione: (2024)
Pathology-Aware Adaptive Watermarking for Text-Driven Medical Image Synthesis
di: Kim, Chanyoung, et al.
Pubblicazione: (2025)
di: Kim, Chanyoung, et al.
Pubblicazione: (2025)
ViKey: Enhancing Temporal Understanding in Videos via Visual Prompting
di: Lee, Yeonkyung, et al.
Pubblicazione: (2026)
di: Lee, Yeonkyung, et al.
Pubblicazione: (2026)
Anchoring and Rescaling Attention for Semantically Coherent Inbetweening
di: Choi, Tae Eun, et al.
Pubblicazione: (2026)
di: Choi, Tae Eun, et al.
Pubblicazione: (2026)
Distilling Spectral Graph for Object-Context Aware Open-Vocabulary Semantic Segmentation
di: Kim, Chanyoung, et al.
Pubblicazione: (2024)
di: Kim, Chanyoung, et al.
Pubblicazione: (2024)
DiffSLT: Enhancing Diversity in Sign Language Translation via Diffusion Model
di: Moon, JiHwan, et al.
Pubblicazione: (2024)
di: Moon, JiHwan, et al.
Pubblicazione: (2024)
Leveraging the Power of MLLMs for Gloss-Free Sign Language Translation
di: Kim, Jungeun, et al.
Pubblicazione: (2024)
di: Kim, Jungeun, et al.
Pubblicazione: (2024)
PLATYPUS: Progressive Local Surface Estimator for Arbitrary-Scale Point Cloud Upsampling
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
MonoWAD: Weather-Adaptive Diffusion Model for Robust Monocular 3D Object Detection
di: Oh, Youngmin, et al.
Pubblicazione: (2024)
di: Oh, Youngmin, et al.
Pubblicazione: (2024)
Slice-Consistent 3D Volumetric Brain CT-to-MRI Translation with 2D Brownian Bridge Diffusion Model
di: Choo, Kyobin, et al.
Pubblicazione: (2024)
di: Choo, Kyobin, et al.
Pubblicazione: (2024)
FALCON: Frequency Adjoint Link with CONtinuous Density Mask for Fast Single Image Dehazing
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
LRSLAM: Low-rank Representation of Signed Distance Fields in Dense Visual SLAM System
di: Park, Hongbeen, et al.
Pubblicazione: (2025)
di: Park, Hongbeen, et al.
Pubblicazione: (2025)
Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
di: Kim, Youngmin, et al.
Pubblicazione: (2025)
di: Kim, Youngmin, et al.
Pubblicazione: (2025)
KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts
di: Hwang, Taebaek, et al.
Pubblicazione: (2025)
di: Hwang, Taebaek, et al.
Pubblicazione: (2025)
Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding
di: Kim, Jiwan, et al.
Pubblicazione: (2026)
di: Kim, Jiwan, et al.
Pubblicazione: (2026)
WAVE: Warp-Based View Guidance for Consistent Novel View Synthesis Using a Single Image
di: Park, Jiwoo, et al.
Pubblicazione: (2025)
di: Park, Jiwoo, et al.
Pubblicazione: (2025)
CompoDistill: Attention Distillation for Compositional Reasoning in Multimodal LLMs
di: Kim, Jiwan, et al.
Pubblicazione: (2025)
di: Kim, Jiwan, et al.
Pubblicazione: (2025)
Towards Online Continuous Sign Language Recognition and Translation
di: Zuo, Ronglai, et al.
Pubblicazione: (2024)
di: Zuo, Ronglai, et al.
Pubblicazione: (2024)
OVS Meets Continual Learning: Towards Sustainable Open-Vocabulary Segmentation
di: Hwang, Dongjun, et al.
Pubblicazione: (2024)
di: Hwang, Dongjun, et al.
Pubblicazione: (2024)
Training Strategies for Isolated Sign Language Recognition
di: Kvanchiani, Karina, et al.
Pubblicazione: (2024)
di: Kvanchiani, Karina, et al.
Pubblicazione: (2024)
Test-Time Training for Visual Foresight Vision-Language-Action Models
di: Park, Sangwu, et al.
Pubblicazione: (2026)
di: Park, Sangwu, et al.
Pubblicazione: (2026)
Weakly Supervised Video Scene Graph Generation via Natural Language Supervision
di: Kim, Kibum, et al.
Pubblicazione: (2025)
di: Kim, Kibum, et al.
Pubblicazione: (2025)
Uni-Sign: Toward Unified Sign Language Understanding at Scale
di: Li, Zecheng, et al.
Pubblicazione: (2025)
di: Li, Zecheng, et al.
Pubblicazione: (2025)
OpenFS: Multi-Hand-Capable Fingerspelling Recognition with Implicit Signing-Hand Detection and Frame-Wise Letter-Conditioned Synthesis
di: Cha, Junuk, et al.
Pubblicazione: (2026)
di: Cha, Junuk, et al.
Pubblicazione: (2026)
RA-SGG: Retrieval-Augmented Scene Graph Generation Framework via Multi-Prototype Learning
di: Yoon, Kanghoon, et al.
Pubblicazione: (2024)
di: Yoon, Kanghoon, et al.
Pubblicazione: (2024)
EgoSpeak: Learning When to Speak for Egocentric Conversational Agents in the Wild
di: Kim, Junhyeok, et al.
Pubblicazione: (2025)
di: Kim, Junhyeok, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Interpreting Attention Heads for Image-to-Text Information Flow in Large Vision-Language Models
di: Kim, Jinyeong, et al.
Pubblicazione: (2025) -
Disentangling Disentangled Representations: Towards Improved Latent Units via Diffusion Models
di: Jun, Youngjun, et al.
Pubblicazione: (2024) -
Delaunay Canopy: Building Wireframe Reconstruction from Airborne LiDAR Point Clouds via Delaunay Graph
di: Kim, Donghyun, et al.
Pubblicazione: (2026) -
Your Large Vision-Language Model Only Needs A Few Attention Heads For Visual Grounding
di: Kang, Seil, et al.
Pubblicazione: (2025) -
CoBra: Complementary Branch Fusing Class and Semantic Knowledge for Robust Weakly Supervised Semantic Segmentation
di: Han, Woojung, et al.
Pubblicazione: (2024)