InteracTalker: Prompt-Based Human-Object Interaction with Co-Speech Gesture Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Rajan, Sreehari, Bhosikar, Kunal, Sharma, Charu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MOGRAS: Human Motion with Grasping in 3D Scenes
por: Bhosikar, Kunal, et al.
Publicado: (2025)
por: Bhosikar, Kunal, et al.
Publicado: (2025)
PatchPoison: Poisoning Multi-View Datasets to Degrade 3D Reconstruction
por: Wadekar, Prajas, et al.
Publicado: (2026)
por: Wadekar, Prajas, et al.
Publicado: (2026)
LiveGesture Streamable Co-Speech Gesture Generation Model
por: Saleem, Muhammad Usama, et al.
Publicado: (2026)
por: Saleem, Muhammad Usama, et al.
Publicado: (2026)
Recognizing Co-Speech Gestures in-the-Wild
por: Hegde, Sindhu B, et al.
Publicado: (2026)
por: Hegde, Sindhu B, et al.
Publicado: (2026)
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
por: Fang, Fengyi, et al.
Publicado: (2025)
por: Fang, Fengyi, et al.
Publicado: (2025)
Conveying Meaning through Gestures: An Investigation into Semantic Co-Speech Gesture Generation
por: Voss, Hendric, et al.
Publicado: (2025)
por: Voss, Hendric, et al.
Publicado: (2025)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
por: Liu, Pinxin, et al.
Publicado: (2025)
por: Liu, Pinxin, et al.
Publicado: (2025)
Co$^{3}$Gesture: Towards Coherent Concurrent Co-speech 3D Gesture Generation with Interactive Diffusion
por: Qi, Xingqun, et al.
Publicado: (2025)
por: Qi, Xingqun, et al.
Publicado: (2025)
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
por: Zhang, Xiangyue, et al.
Publicado: (2026)
por: Zhang, Xiangyue, et al.
Publicado: (2026)
GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling
por: Liu, Pinxin, et al.
Publicado: (2025)
por: Liu, Pinxin, et al.
Publicado: (2025)
DuoGesture: Neuro-Inspired and Biomechanically Informed Dual-Stream Co-Speech Gesture Generation
por: Paar, Ferdinand, et al.
Publicado: (2026)
por: Paar, Ferdinand, et al.
Publicado: (2026)
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling
por: Liu, Haiyang, et al.
Publicado: (2023)
por: Liu, Haiyang, et al.
Publicado: (2023)
Prompt-to-Gesture: Measuring the Capabilities of Image-to-Video Deictic Gesture Generation
por: Ali, Hassan, et al.
Publicado: (2026)
por: Ali, Hassan, et al.
Publicado: (2026)
Chain-of-Talkers (CoTalk): Fast Human Annotation of Dense Image Captions
por: Shen, Yijun, et al.
Publicado: (2025)
por: Shen, Yijun, et al.
Publicado: (2025)
Democratizing High-Fidelity Co-Speech Gesture Video Generation
por: Yang, Xu, et al.
Publicado: (2025)
por: Yang, Xu, et al.
Publicado: (2025)
Joint Co-Speech Gesture and Expressive Talking Face Generation using Diffusion with Adapters
por: Hogue, Steven, et al.
Publicado: (2024)
por: Hogue, Steven, et al.
Publicado: (2024)
CoCoGesture: Toward Coherent Co-speech 3D Gesture Generation in the Wild
por: Qi, Xingqun, et al.
Publicado: (2024)
por: Qi, Xingqun, et al.
Publicado: (2024)
AnyTalker: Scaling Multi-Person Talking Video Generation with Interactivity Refinement
por: Zhong, Zhizhou, et al.
Publicado: (2025)
por: Zhong, Zhizhou, et al.
Publicado: (2025)
HOP: Heterogeneous Topology-based Multimodal Entanglement for Co-Speech Gesture Generation
por: Cheng, Hongye, et al.
Publicado: (2025)
por: Cheng, Hongye, et al.
Publicado: (2025)
MMGT: Motion Mask Guided Two-Stage Network for Co-Speech Gesture Video Generation
por: Wang, Siyuan, et al.
Publicado: (2025)
por: Wang, Siyuan, et al.
Publicado: (2025)
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation
por: Mao, Xiaofeng, et al.
Publicado: (2024)
por: Mao, Xiaofeng, et al.
Publicado: (2024)
CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation
por: Luo, Xiangyang, et al.
Publicado: (2026)
por: Luo, Xiangyang, et al.
Publicado: (2026)
Co-Speech Gesture Detection through Multi-Phase Sequence Labeling
por: Ghaleb, Esam, et al.
Publicado: (2023)
por: Ghaleb, Esam, et al.
Publicado: (2023)
EmotionGesture: Audio-Driven Diverse Emotional Co-Speech 3D Gesture Generation
por: Qi, Xingqun, et al.
Publicado: (2023)
por: Qi, Xingqun, et al.
Publicado: (2023)
DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion Autoencoder
por: Du, Chenpeng, et al.
Publicado: (2023)
por: Du, Chenpeng, et al.
Publicado: (2023)
Bridge to Non-Barrier Communication: Gloss-Prompted Fine-grained Cued Speech Gesture Generation with Diffusion Model
por: Lei, Wentao, et al.
Publicado: (2024)
por: Lei, Wentao, et al.
Publicado: (2024)
VarGes: Improving Variation in Co-Speech 3D Gesture Generation via StyleCLIPS
por: Meng, Ming, et al.
Publicado: (2025)
por: Meng, Ming, et al.
Publicado: (2025)
SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning
por: Liu, Lanmiao, et al.
Publicado: (2025)
por: Liu, Lanmiao, et al.
Publicado: (2025)
HolisticSemGes: Semantic Grounding of Holistic Co-Speech Gesture Generation with Contrastive Flow-Matching
por: Liu, Lanmiao, et al.
Publicado: (2026)
por: Liu, Lanmiao, et al.
Publicado: (2026)
Enabling Synergistic Full-Body Control in Prompt-Based Co-Speech Motion Generation
por: Chen, Bohong, et al.
Publicado: (2024)
por: Chen, Bohong, et al.
Publicado: (2024)
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture Synthesis
por: Mughal, Muhammad Hamza, et al.
Publicado: (2024)
por: Mughal, Muhammad Hamza, et al.
Publicado: (2024)
Orchestrating the Symphony of Prompt Distribution Learning for Human-Object Interaction Detection
por: Jia, Mingda, et al.
Publicado: (2024)
por: Jia, Mingda, et al.
Publicado: (2024)
Co-speech Gesture Video Generation via Motion-Based Graph Retrieval
por: Song, Yafei, et al.
Publicado: (2025)
por: Song, Yafei, et al.
Publicado: (2025)
Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model
por: He, Xu, et al.
Publicado: (2024)
por: He, Xu, et al.
Publicado: (2024)
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion
por: Vu, Evgeniia, et al.
Publicado: (2025)
por: Vu, Evgeniia, et al.
Publicado: (2025)
VIZOR: Viewpoint-Invariant Zero-Shot Scene Graph Generation for 3D Scene Reasoning
por: Madhavaram, Vivek, et al.
Publicado: (2026)
por: Madhavaram, Vivek, et al.
Publicado: (2026)
Open-World Human-Object Interaction Detection via Multi-modal Prompts
por: Yang, Jie, et al.
Publicado: (2024)
por: Yang, Jie, et al.
Publicado: (2024)
Emphasizing Semantic Consistency of Salient Posture for Speech-Driven Gesture Generation
por: Liu, Fengqi, et al.
Publicado: (2024)
por: Liu, Fengqi, et al.
Publicado: (2024)
GLDiTalker: Speech-Driven 3D Facial Animation with Graph Latent Diffusion Transformer
por: Lin, Yihong, et al.
Publicado: (2024)
por: Lin, Yihong, et al.
Publicado: (2024)
TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation
por: Liu, Haiyang, et al.
Publicado: (2024)
por: Liu, Haiyang, et al.
Publicado: (2024)
Ejemplares similares
-
MOGRAS: Human Motion with Grasping in 3D Scenes
por: Bhosikar, Kunal, et al.
Publicado: (2025) -
PatchPoison: Poisoning Multi-View Datasets to Degrade 3D Reconstruction
por: Wadekar, Prajas, et al.
Publicado: (2026) -
LiveGesture Streamable Co-Speech Gesture Generation Model
por: Saleem, Muhammad Usama, et al.
Publicado: (2026) -
Recognizing Co-Speech Gestures in-the-Wild
por: Hegde, Sindhu B, et al.
Publicado: (2026) -
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
por: Fang, Fengyi, et al.
Publicado: (2025)