Towards Fine-Grained Emotion Understanding via Skeleton-Based Micro-Gesture Recognition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Hao, Cheng, Lechao, Wang, Yaxiong, Tang, Shengeng, Zhong, Zhun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Micro-Action Recognition with Limited Annotations: An Asynchronous Pseudo Labeling and Training Approach
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
von: Zhang, Yan, et al.
Veröffentlicht: (2025)
Knowledge Swapping via Learning and Unlearning
von: Xing, Mingyu, et al.
Veröffentlicht: (2025)
von: Xing, Mingyu, et al.
Veröffentlicht: (2025)
Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline
von: Li, Haiyang, et al.
Veröffentlicht: (2025)
von: Li, Haiyang, et al.
Veröffentlicht: (2025)
OmniVL-Guard: Towards Unified Vision-Language Forgery Detection and Grounding via Balanced RL
von: Shen, Jinjie, et al.
Veröffentlicht: (2026)
von: Shen, Jinjie, et al.
Veröffentlicht: (2026)
CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition
von: Wang, Xu, et al.
Veröffentlicht: (2026)
von: Wang, Xu, et al.
Veröffentlicht: (2026)
TDEdit: A Unified Diffusion Framework for Text-Drag Guided Image Manipulation
von: Wang, Qihang, et al.
Veröffentlicht: (2025)
von: Wang, Qihang, et al.
Veröffentlicht: (2025)
Motion is the Choreographer: Learning Latent Pose Dynamics for Seamless Sign Language Generation
von: He, Jiayi, et al.
Veröffentlicht: (2025)
von: He, Jiayi, et al.
Veröffentlicht: (2025)
Text-Driven Diffusion Model for Sign Language Production
von: He, Jiayi, et al.
Veröffentlicht: (2025)
von: He, Jiayi, et al.
Veröffentlicht: (2025)
EntityCLIP: Entity-Centric Image-Text Matching via Multimodal Attentive Contrastive Learning
von: Wang, Yaxiong, et al.
Veröffentlicht: (2024)
von: Wang, Yaxiong, et al.
Veröffentlicht: (2024)
Beyond Artificial Misalignment: Detecting and Grounding Semantic-Coordinated Multimodal Manipulations
von: Shen, Jinjie, et al.
Veröffentlicht: (2025)
von: Shen, Jinjie, et al.
Veröffentlicht: (2025)
SSAM: Self-Supervised Association Modeling for Test-Time Adaption
von: Wang, Yaxiong, et al.
Veröffentlicht: (2025)
von: Wang, Yaxiong, et al.
Veröffentlicht: (2025)
Modality Alignment Meets Federated Broadcasting
von: Ma, Yuting, et al.
Veröffentlicht: (2024)
von: Ma, Yuting, et al.
Veröffentlicht: (2024)
FedHPL: Efficient Heterogeneous Federated Learning with Prompt Tuning and Logit Distillation
von: Ma, Yuting, et al.
Veröffentlicht: (2024)
von: Ma, Yuting, et al.
Veröffentlicht: (2024)
OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics
von: Shen, Jinjie, et al.
Veröffentlicht: (2026)
von: Shen, Jinjie, et al.
Veröffentlicht: (2026)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
Synchronized and Fine-Grained Head for Skeleton-Based Ambiguous Action Recognition
von: Huang, Hao, et al.
Veröffentlicht: (2024)
von: Huang, Hao, et al.
Veröffentlicht: (2024)
Shaping a Stabilized Video by Mitigating Unintended Changes for Concept-Augmented Video Editing
von: Guo, Mingce, et al.
Veröffentlicht: (2024)
von: Guo, Mingce, et al.
Veröffentlicht: (2024)
Efficient Vision Language Model Fine-tuning for Text-based Person Anomaly Search
von: He, Jiayi, et al.
Veröffentlicht: (2025)
von: He, Jiayi, et al.
Veröffentlicht: (2025)
FG-SGL: Fine-Grained Semantic Guidance Learning via Motion Process Decomposition for Micro-Gesture Recognition
von: Wei, Jinsheng, et al.
Veröffentlicht: (2026)
von: Wei, Jinsheng, et al.
Veröffentlicht: (2026)
SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation
von: Wang, Xu, et al.
Veröffentlicht: (2025)
von: Wang, Xu, et al.
Veröffentlicht: (2025)
Enhancing Micro Gesture Recognition for Emotion Understanding via Context-aware Visual-Text Contrastive Learning
von: Li, Deng, et al.
Veröffentlicht: (2024)
von: Li, Deng, et al.
Veröffentlicht: (2024)
Open-World 3D Scene Graph Generation for Retrieval-Augmented Reasoning
von: Yu, Fei, et al.
Veröffentlicht: (2025)
von: Yu, Fei, et al.
Veröffentlicht: (2025)
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
von: Gu, Jihao, et al.
Veröffentlicht: (2025)
SplitGaussian: Reconstructing Dynamic Scenes via Visual Geometry Decomposition
von: Li, Jiahui, et al.
Veröffentlicht: (2025)
von: Li, Jiahui, et al.
Veröffentlicht: (2025)
Prior-Constrained Association Learning for Fine-Grained Generalized Category Discovery
von: Wang, Menglin, et al.
Veröffentlicht: (2025)
von: Wang, Menglin, et al.
Veröffentlicht: (2025)
Micro-Expression Recognition via Fine-Grained Dynamic Perception
von: Shao, Zhiwen, et al.
Veröffentlicht: (2025)
von: Shao, Zhiwen, et al.
Veröffentlicht: (2025)
Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering
von: Wang, Xu, et al.
Veröffentlicht: (2025)
von: Wang, Xu, et al.
Veröffentlicht: (2025)
Discrete to Continuous: Generating Smooth Transition Poses from Sign Language Observation
von: Tang, Shengeng, et al.
Veröffentlicht: (2024)
von: Tang, Shengeng, et al.
Veröffentlicht: (2024)
StgcDiff: Spatial-Temporal Graph Condition Diffusion for Sign Language Transition Generation
von: He, Jiashu, et al.
Veröffentlicht: (2025)
von: He, Jiashu, et al.
Veröffentlicht: (2025)
Cross-Block Fine-Grained Semantic Cascade for Skeleton-Based Sports Action Recognition
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
von: Liu, Zhendong, et al.
Veröffentlicht: (2024)
Navigating Semantic Drift in Task-Agnostic Class-Incremental Learning
von: Wu, Fangwen, et al.
Veröffentlicht: (2025)
von: Wu, Fangwen, et al.
Veröffentlicht: (2025)
Dataset Distillers Are Good Label Denoisers In the Wild
von: Cheng, Lechao, et al.
Veröffentlicht: (2024)
von: Cheng, Lechao, et al.
Veröffentlicht: (2024)
Towards Universal Skeleton-Based Action Recognition
von: Kuang, Jidong, et al.
Veröffentlicht: (2026)
von: Kuang, Jidong, et al.
Veröffentlicht: (2026)
Multi-Track Multimodal Learning on iMiGUE: Micro-Gesture and Emotion Recognition
von: Martirosyan, Arman, et al.
Veröffentlicht: (2025)
von: Martirosyan, Arman, et al.
Veröffentlicht: (2025)
Decoupled Training with Local Reinforcement Fine-Tuning in Federated Learning
von: Ma, Yuting, et al.
Veröffentlicht: (2026)
von: Ma, Yuting, et al.
Veröffentlicht: (2026)
Micro-gesture Online Recognition using Learnable Query Points
von: Liu, Pengyu, et al.
Veröffentlicht: (2024)
von: Liu, Pengyu, et al.
Veröffentlicht: (2024)
FineTec: Fine-Grained Action Recognition Under Temporal Corruption via Skeleton Decomposition and Sequence Completion
von: Shao, Dian, et al.
Veröffentlicht: (2025)
von: Shao, Dian, et al.
Veröffentlicht: (2025)
OMG-Bench: A New Challenging Benchmark for Skeleton-based Online Micro Hand Gesture Recognition
von: Chang, Haochen, et al.
Veröffentlicht: (2025)
von: Chang, Haochen, et al.
Veröffentlicht: (2025)
Identity-free Artificial Emotional Intelligence via Micro-Gesture Understanding
von: Gao, Rong, et al.
Veröffentlicht: (2024)
von: Gao, Rong, et al.
Veröffentlicht: (2024)
Wavelet-Decoupling Contrastive Enhancement Network for Fine-Grained Skeleton-Based Action Recognition
von: Chang, Haochen, et al.
Veröffentlicht: (2024)
von: Chang, Haochen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Micro-Action Recognition with Limited Annotations: An Asynchronous Pseudo Labeling and Training Approach
von: Zhang, Yan, et al.
Veröffentlicht: (2025) -
Knowledge Swapping via Learning and Unlearning
von: Xing, Mingyu, et al.
Veröffentlicht: (2025) -
Towards Unified Multimodal Misinformation Detection in Social Media: A Benchmark Dataset and Baseline
von: Li, Haiyang, et al.
Veröffentlicht: (2025) -
OmniVL-Guard: Towards Unified Vision-Language Forgery Detection and Grounding via Balanced RL
von: Shen, Jinjie, et al.
Veröffentlicht: (2026) -
CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition
von: Wang, Xu, et al.
Veröffentlicht: (2026)