Understanding Co-speech Gestures in-the-wild
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hegde, Sindhu B, Prajwal, K R, Kwon, Taein, Zisserman, Andrew |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Recognizing Co-Speech Gestures in-the-Wild
von: Hegde, Sindhu B, et al.
Veröffentlicht: (2026)
von: Hegde, Sindhu B, et al.
Veröffentlicht: (2026)
Recognising BSL Fingerspelling in Continuous Signing Sequences
von: Chan, Alyssa, et al.
Veröffentlicht: (2026)
von: Chan, Alyssa, et al.
Veröffentlicht: (2026)
New keypoint-based approach for recognising British Sign Language (BSL) from sequences
von: Deb, Oishi, et al.
Veröffentlicht: (2024)
von: Deb, Oishi, et al.
Veröffentlicht: (2024)
CoCoGesture: Toward Coherent Co-speech 3D Gesture Generation in the Wild
von: Qi, Xingqun, et al.
Veröffentlicht: (2024)
von: Qi, Xingqun, et al.
Veröffentlicht: (2024)
Co$^{3}$Gesture: Towards Coherent Concurrent Co-speech 3D Gesture Generation with Interactive Diffusion
von: Qi, Xingqun, et al.
Veröffentlicht: (2025)
von: Qi, Xingqun, et al.
Veröffentlicht: (2025)
EgoWorld: Translating Exocentric View to Egocentric View using Rich Exocentric Observations
von: Park, Junho, et al.
Veröffentlicht: (2025)
von: Park, Junho, et al.
Veröffentlicht: (2025)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
von: Song, Yafei, et al.
Veröffentlicht: (2025)
von: Song, Yafei, et al.
Veröffentlicht: (2025)
Multi Activity Sequence Alignment via Implicit Clustering
von: Kwon, Taein, et al.
Veröffentlicht: (2025)
von: Kwon, Taein, et al.
Veröffentlicht: (2025)
Motion-example-controlled Co-speech Gesture Generation Leveraging Large Language Models
von: Chen, Bohong, et al.
Veröffentlicht: (2025)
von: Chen, Bohong, et al.
Veröffentlicht: (2025)
Self-Supervised Learning of Deviation in Latent Representation for Co-speech Gesture Video Generation
von: Yang, Huan, et al.
Veröffentlicht: (2024)
von: Yang, Huan, et al.
Veröffentlicht: (2024)
A Tale of Two Languages: Large-Vocabulary Continuous Sign Language Recognition from Spoken Language Supervision
von: Raude, Charles, et al.
Veröffentlicht: (2024)
von: Raude, Charles, et al.
Veröffentlicht: (2024)
HoloGest: Decoupled Diffusion and Motion Priors for Generating Holisticly Expressive Co-speech Gestures
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
von: Cheng, Yongkang, et al.
Veröffentlicht: (2025)
Weakly-Supervised Emotion Transition Learning for Diverse 3D Co-speech Gesture Generation
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
von: Qi, Xingqun, et al.
Veröffentlicht: (2023)
Personalizing Retrieval using Joint Embeddings or "the Return of Fluffy"
von: Korbar, Bruno, et al.
Veröffentlicht: (2025)
von: Korbar, Bruno, et al.
Veröffentlicht: (2025)
Chirality in Action: Time-Aware Video Representation Learning by Latent Straightening
von: Bagad, Piyush, et al.
Veröffentlicht: (2025)
von: Bagad, Piyush, et al.
Veröffentlicht: (2025)
From Panels to Prose: Generating Literary Narratives from Comics
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2025)
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2025)
The Manga Whisperer: Automatically Generating Transcriptions for Comics
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2024)
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2024)
A General Protocol to Probe Large Vision Models for 3D Physical Understanding
von: Zhan, Guanqi, et al.
Veröffentlicht: (2023)
von: Zhan, Guanqi, et al.
Veröffentlicht: (2023)
Character-Centric Understanding of Animated Movies
von: Gui, Zhongrui, et al.
Veröffentlicht: (2025)
von: Gui, Zhongrui, et al.
Veröffentlicht: (2025)
Text-Conditioned Resampler For Long Form Video Understanding
von: Korbar, Bruno, et al.
Veröffentlicht: (2023)
von: Korbar, Bruno, et al.
Veröffentlicht: (2023)
CountGD++: Generalized Prompting for Open-World Counting
von: Amini-Naieni, Niki, et al.
Veröffentlicht: (2025)
von: Amini-Naieni, Niki, et al.
Veröffentlicht: (2025)
SocialGesture: Delving into Multi-person Gesture Understanding
von: Cao, Xu, et al.
Veröffentlicht: (2025)
von: Cao, Xu, et al.
Veröffentlicht: (2025)
DiffTED: One-shot Audio-driven TED Talk Video Generation with Diffusion-based Co-speech Gestures
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
von: Hogue, Steven, et al.
Veröffentlicht: (2024)
LiveGesture Streamable Co-Speech Gesture Generation Model
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2026)
von: Saleem, Muhammad Usama, et al.
Veröffentlicht: (2026)
Tails Tell Tales: Chapter-Wide Manga Transcriptions with Character Names
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2024)
von: Sachdeva, Ragav, et al.
Veröffentlicht: (2024)
Appearance-Based Refinement for Object-Centric Motion Segmentation
von: Xie, Junyu, et al.
Veröffentlicht: (2023)
von: Xie, Junyu, et al.
Veröffentlicht: (2023)
Adapting MLLMs for Nuanced Video Retrieval
von: Bagad, Piyush, et al.
Veröffentlicht: (2025)
von: Bagad, Piyush, et al.
Veröffentlicht: (2025)
EgoPressure: A Dataset for Hand Pressure and Pose Estimation in Egocentric Vision
von: Zhao, Yiming, et al.
Veröffentlicht: (2024)
von: Zhao, Yiming, et al.
Veröffentlicht: (2024)
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026)
von: Zhang, Xiangyue, et al.
Veröffentlicht: (2026)
CountGD: Multi-Modal Open-World Counting
von: Amini-Naieni, Niki, et al.
Veröffentlicht: (2024)
von: Amini-Naieni, Niki, et al.
Veröffentlicht: (2024)
Open-World Object Counting in Videos
von: Amini-Naieni, Niki, et al.
Veröffentlicht: (2025)
von: Amini-Naieni, Niki, et al.
Veröffentlicht: (2025)
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
von: Fang, Fengyi, et al.
Veröffentlicht: (2025)
von: Fang, Fengyi, et al.
Veröffentlicht: (2025)
Amodal Ground Truth and Completion in the Wild
von: Zhan, Guanqi, et al.
Veröffentlicht: (2023)
von: Zhan, Guanqi, et al.
Veröffentlicht: (2023)
It's Just Another Day: Unique Video Captioning by Discriminative Prompting
von: Perrett, Toby, et al.
Veröffentlicht: (2024)
von: Perrett, Toby, et al.
Veröffentlicht: (2024)
Moving Object Segmentation: All You Need Is SAM (and Flow)
von: Xie, Junyu, et al.
Veröffentlicht: (2024)
von: Xie, Junyu, et al.
Veröffentlicht: (2024)
GMOS: Grounding Moving Object Segmentation in 3D Space and Time
von: Xie, Junyu, et al.
Veröffentlicht: (2026)
von: Xie, Junyu, et al.
Veröffentlicht: (2026)
EgoXtreme: A Dataset for Robust Object Pose Estimation in Egocentric Views under Extreme Conditions
von: Yoon, Taegyoon, et al.
Veröffentlicht: (2026)
von: Yoon, Taegyoon, et al.
Veröffentlicht: (2026)
Made to Order: Discovering monotonic temporal changes via self-supervised video ordering
von: Yang, Charig, et al.
Veröffentlicht: (2024)
von: Yang, Charig, et al.
Veröffentlicht: (2024)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
von: Liu, Pinxin, et al.
Veröffentlicht: (2025)
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling
von: Liu, Haiyang, et al.
Veröffentlicht: (2023)
von: Liu, Haiyang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Recognizing Co-Speech Gestures in-the-Wild
von: Hegde, Sindhu B, et al.
Veröffentlicht: (2026) -
Recognising BSL Fingerspelling in Continuous Signing Sequences
von: Chan, Alyssa, et al.
Veröffentlicht: (2026) -
New keypoint-based approach for recognising British Sign Language (BSL) from sequences
von: Deb, Oishi, et al.
Veröffentlicht: (2024) -
CoCoGesture: Toward Coherent Co-speech 3D Gesture Generation in the Wild
von: Qi, Xingqun, et al.
Veröffentlicht: (2024) -
Co$^{3}$Gesture: Towards Coherent Concurrent Co-speech 3D Gesture Generation with Interactive Diffusion
von: Qi, Xingqun, et al.
Veröffentlicht: (2025)