SocialGesture: Delving into Multi-person Gesture Understanding
Fuente:
arXiv
Guardado en:
| Autores principales: | Cao, Xu, Virupaksha, Pranav, Jia, Wenqi, Lai, Bolin, Ryan, Fiona, Lee, Sangmin, Rehg, James M. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations
por: Lee, Sangmin, et al.
Publicado: (2024)
por: Lee, Sangmin, et al.
Publicado: (2024)
Listen to Look into the Future: Audio-Visual Egocentric Gaze Anticipation
por: Lai, Bolin, et al.
Publicado: (2023)
por: Lai, Bolin, et al.
Publicado: (2023)
GRASP: Learning to Ground Social Reasoning in Multi-Person Non-Verbal Interactions
por: Kim, Junho, et al.
Publicado: (2026)
por: Kim, Junho, et al.
Publicado: (2026)
Unified Text-Image-to-Video Generation: A Training-Free Approach to Flexible Visual Conditioning
por: Lai, Bolin, et al.
Publicado: (2025)
por: Lai, Bolin, et al.
Publicado: (2025)
In the Eye of Transformer: Global-Local Correlation for Egocentric Gaze Estimation
por: Lai, Bolin, et al.
Publicado: (2022)
por: Lai, Bolin, et al.
Publicado: (2022)
Learning Predictive Visuomotor Coordination
por: Jia, Wenqi, et al.
Publicado: (2025)
por: Jia, Wenqi, et al.
Publicado: (2025)
Towards Online Multi-Modal Social Interaction Understanding
por: Li, Xinpeng, et al.
Publicado: (2025)
por: Li, Xinpeng, et al.
Publicado: (2025)
Gaze-LLE: Gaze Target Estimation via Large-Scale Learned Encoders
por: Ryan, Fiona, et al.
Publicado: (2024)
por: Ryan, Fiona, et al.
Publicado: (2024)
Boosting Gesture Recognition with an Automatic Gesture Annotation Framework
por: Shen, Junxiao, et al.
Publicado: (2024)
por: Shen, Junxiao, et al.
Publicado: (2024)
Omni-MMSI: Toward Identity-attributed Social Interaction Understanding
por: Li, Xinpeng, et al.
Publicado: (2026)
por: Li, Xinpeng, et al.
Publicado: (2026)
Understanding Co-speech Gestures in-the-wild
por: Hegde, Sindhu B, et al.
Publicado: (2025)
por: Hegde, Sindhu B, et al.
Publicado: (2025)
LiveGesture Streamable Co-Speech Gesture Generation Model
por: Saleem, Muhammad Usama, et al.
Publicado: (2026)
por: Saleem, Muhammad Usama, et al.
Publicado: (2026)
Intentional Gesture: Deliver Your Intentions with Gestures for Speech
por: Liu, Pinxin, et al.
Publicado: (2025)
por: Liu, Pinxin, et al.
Publicado: (2025)
Prompt-to-Gesture: Measuring the Capabilities of Image-to-Video Deictic Gesture Generation
por: Ali, Hassan, et al.
Publicado: (2026)
por: Ali, Hassan, et al.
Publicado: (2026)
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation
por: Liu, Pinxin, et al.
Publicado: (2025)
por: Liu, Pinxin, et al.
Publicado: (2025)
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
por: Gu, Jihao, et al.
Publicado: (2025)
por: Gu, Jihao, et al.
Publicado: (2025)
PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers
por: Zhang, Xiangyue, et al.
Publicado: (2026)
por: Zhang, Xiangyue, et al.
Publicado: (2026)
MM-SpuBench: Towards Better Understanding of Spurious Biases in Multimodal LLMs
por: Ye, Wenqian, et al.
Publicado: (2024)
por: Ye, Wenqian, et al.
Publicado: (2024)
Leveraging Object Priors for Point Tracking
por: Boote, Bikram, et al.
Publicado: (2024)
por: Boote, Bikram, et al.
Publicado: (2024)
GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling
por: Liu, Pinxin, et al.
Publicado: (2025)
por: Liu, Pinxin, et al.
Publicado: (2025)
Unleashing In-context Learning of Autoregressive Models for Few-shot Image Manipulation
por: Lai, Bolin, et al.
Publicado: (2024)
por: Lai, Bolin, et al.
Publicado: (2024)
Multi-task Learning For Joint Action and Gesture Recognition
por: Spathis, Konstantinos, et al.
Publicado: (2025)
por: Spathis, Konstantinos, et al.
Publicado: (2025)
Towards Fine-Grained Emotion Understanding via Skeleton-Based Micro-Gesture Recognition
por: Xu, Hao, et al.
Publicado: (2025)
por: Xu, Hao, et al.
Publicado: (2025)
LEGO: Learning EGOcentric Action Frame Generation via Visual Instruction Tuning
por: Lai, Bolin, et al.
Publicado: (2023)
por: Lai, Bolin, et al.
Publicado: (2023)
Towards Open-World Gesture Recognition
por: Shen, Junxiao, et al.
Publicado: (2024)
por: Shen, Junxiao, et al.
Publicado: (2024)
CoordSpeaker: Exploiting Gesture Captioning for Coordinated Caption-Empowered Co-Speech Gesture Generation
por: Fang, Fengyi, et al.
Publicado: (2025)
por: Fang, Fengyi, et al.
Publicado: (2025)
CoCoGesture: Toward Coherent Co-speech 3D Gesture Generation in the Wild
por: Qi, Xingqun, et al.
Publicado: (2024)
por: Qi, Xingqun, et al.
Publicado: (2024)
Hand2World: Autoregressive Egocentric Interaction Generation via Free-Space Hand Gestures
por: Wang, Yuxi, et al.
Publicado: (2026)
por: Wang, Yuxi, et al.
Publicado: (2026)
DyaDiT: A Multi-Modal Diffusion Transformer for Socially Favorable Dyadic Gesture Generation
por: Peng, Yichen, et al.
Publicado: (2026)
por: Peng, Yichen, et al.
Publicado: (2026)
Recognizing Co-Speech Gestures in-the-Wild
por: Hegde, Sindhu B, et al.
Publicado: (2026)
por: Hegde, Sindhu B, et al.
Publicado: (2026)
Interpretable Underwater Diver Gesture Recognition
por: Mangalvedhekar, Sudeep, et al.
Publicado: (2023)
por: Mangalvedhekar, Sudeep, et al.
Publicado: (2023)
Zero-Shot Underwater Gesture Recognition
por: Sarma, Sandipan, et al.
Publicado: (2024)
por: Sarma, Sandipan, et al.
Publicado: (2024)
DuoGesture: Neuro-Inspired and Biomechanically Informed Dual-Stream Co-Speech Gesture Generation
por: Paar, Ferdinand, et al.
Publicado: (2026)
por: Paar, Ferdinand, et al.
Publicado: (2026)
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling
por: Liu, Haiyang, et al.
Publicado: (2023)
por: Liu, Haiyang, et al.
Publicado: (2023)
Co$^{3}$Gesture: Towards Coherent Concurrent Co-speech 3D Gesture Generation with Interactive Diffusion
por: Qi, Xingqun, et al.
Publicado: (2025)
por: Qi, Xingqun, et al.
Publicado: (2025)
Elastic Spiking Transformers for Efficient Gesture Understanding
por: Ancilotto, Alberto, et al.
Publicado: (2026)
por: Ancilotto, Alberto, et al.
Publicado: (2026)
Co-Speech Gesture Detection through Multi-Phase Sequence Labeling
por: Ghaleb, Esam, et al.
Publicado: (2023)
por: Ghaleb, Esam, et al.
Publicado: (2023)
Toward Diffusible High-Dimensional Latent Spaces: A Frequency Perspective
por: Lai, Bolin, et al.
Publicado: (2025)
por: Lai, Bolin, et al.
Publicado: (2025)
Conveying Meaning through Gestures: An Investigation into Semantic Co-Speech Gesture Generation
por: Voss, Hendric, et al.
Publicado: (2025)
por: Voss, Hendric, et al.
Publicado: (2025)
Emotion Detection through Body Gesture and Face
por: Liu, Haoyang
Publicado: (2024)
por: Liu, Haoyang
Publicado: (2024)
Ejemplares similares
-
Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations
por: Lee, Sangmin, et al.
Publicado: (2024) -
Listen to Look into the Future: Audio-Visual Egocentric Gaze Anticipation
por: Lai, Bolin, et al.
Publicado: (2023) -
GRASP: Learning to Ground Social Reasoning in Multi-Person Non-Verbal Interactions
por: Kim, Junho, et al.
Publicado: (2026) -
Unified Text-Image-to-Video Generation: A Training-Free Approach to Flexible Visual Conditioning
por: Lai, Bolin, et al.
Publicado: (2025) -
In the Eye of Transformer: Global-Local Correlation for Egocentric Gaze Estimation
por: Lai, Bolin, et al.
Publicado: (2022)