Micro-gesture Online Recognition using Learnable Query Points
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Pengyu, Wang, Fei, Li, Kun, Chen, Guoliang, Wei, Yanyan, Tang, Shengeng, Wu, Zhiliang, Guo, Dan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Online Micro-gesture Recognition Using Data Augmentation and Spatial-Temporal Attention
by: Liu, Pengyu, et al.
Published: (2025)
by: Liu, Pengyu, et al.
Published: (2025)
Prototype Learning for Micro-gesture Classification
by: Chen, Guoliang, et al.
Published: (2024)
by: Chen, Guoliang, et al.
Published: (2024)
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
by: Gu, Jihao, et al.
Published: (2025)
by: Gu, Jihao, et al.
Published: (2025)
MMAD: Multi-label Micro-Action Detection in Videos
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
Prototypical Calibrating Ambiguous Samples for Micro-Action Recognition
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
Motion Matters: Motion-guided Modulation Network for Skeleton-based Micro-Action Recognition
by: Gu, Jihao, et al.
Published: (2025)
by: Gu, Jihao, et al.
Published: (2025)
MA-Bench: Towards Fine-grained Micro-Action Understanding
by: Li, Kun, et al.
Published: (2026)
by: Li, Kun, et al.
Published: (2026)
Sign-IDD: Iconicity Disentangled Diffusion for Sign Language Production
by: Tang, Shengeng, et al.
Published: (2024)
by: Tang, Shengeng, et al.
Published: (2024)
Towards Fine-Grained Emotion Understanding via Skeleton-Based Micro-Gesture Recognition
by: Xu, Hao, et al.
Published: (2025)
by: Xu, Hao, et al.
Published: (2025)
Hybrid-supervised Hypergraph-enhanced Transformer for Micro-gesture Based Emotion Recognition
by: Xia, Zhaoqiang, et al.
Published: (2025)
by: Xia, Zhaoqiang, et al.
Published: (2025)
Efficient Vision Language Model Fine-tuning for Text-based Person Anomaly Search
by: He, Jiayi, et al.
Published: (2025)
by: He, Jiayi, et al.
Published: (2025)
Exploiting Ensemble Learning for Cross-View Isolated Sign Language Recognition
by: Wang, Fei, et al.
Published: (2025)
by: Wang, Fei, et al.
Published: (2025)
A Two-Stage Adverse Weather Semantic Segmentation Method for WeatherProof Challenge CVPR 2024 Workshop UG2+
by: Wang, Jianzhao, et al.
Published: (2024)
by: Wang, Jianzhao, et al.
Published: (2024)
Benchmarking Micro-action Recognition: Dataset, Methods, and Applications
by: Guo, Dan, et al.
Published: (2024)
by: Guo, Dan, et al.
Published: (2024)
Towards Pixel-Level Prediction for Gaze Following: Benchmark and Approach
by: Liu, Feiyang, et al.
Published: (2024)
by: Liu, Feiyang, et al.
Published: (2024)
Patch-level Sounding Object Tracking for Audio-Visual Question Answering
by: Li, Zhangbin, et al.
Published: (2024)
by: Li, Zhangbin, et al.
Published: (2024)
3D Smoke Scene Reconstruction Guided by Vision Priors from Multimodal Large Language Models
by: Zheng, Xinye, et al.
Published: (2026)
by: Zheng, Xinye, et al.
Published: (2026)
Dense Audio-Visual Event Localization under Cross-Modal Consistency and Multi-Temporal Granularity Collaboration
by: Zhou, Ziheng, et al.
Published: (2024)
by: Zhou, Ziheng, et al.
Published: (2024)
Discrete to Continuous: Generating Smooth Transition Poses from Sign Language Observation
by: Tang, Shengeng, et al.
Published: (2024)
by: Tang, Shengeng, et al.
Published: (2024)
Wi-CBR: Salient-aware Adaptive WiFi Sensing for Cross-domain Behavior Recognition
by: Zhang, Ruobei, et al.
Published: (2025)
by: Zhang, Ruobei, et al.
Published: (2025)
Online hand gesture recognition using Continual Graph Transformers
by: Slama, Rim, et al.
Published: (2025)
by: Slama, Rim, et al.
Published: (2025)
Motion is the Choreographer: Learning Latent Pose Dynamics for Seamless Sign Language Generation
by: He, Jiayi, et al.
Published: (2025)
by: He, Jiayi, et al.
Published: (2025)
Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering
by: Wang, Xu, et al.
Published: (2025)
by: Wang, Xu, et al.
Published: (2025)
Open-World 3D Scene Graph Generation for Retrieval-Augmented Reasoning
by: Yu, Fei, et al.
Published: (2025)
by: Yu, Fei, et al.
Published: (2025)
Learnable Query Aggregation with KV Routing for Cross-view Geo-localisation
by: Ye, Hualin, et al.
Published: (2025)
by: Ye, Hualin, et al.
Published: (2025)
CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition
by: Wang, Xu, et al.
Published: (2026)
by: Wang, Xu, et al.
Published: (2026)
Linguistics-Vision Monotonic Consistent Network for Sign Language Production
by: Wang, Xu, et al.
Published: (2024)
by: Wang, Xu, et al.
Published: (2024)
Image Aesthetics Assessment via Learnable Queries
by: Xiong, Zhiwei, et al.
Published: (2023)
by: Xiong, Zhiwei, et al.
Published: (2023)
Multimodal Protein Language Models for Enzyme Kinetic Parameters: From Substrate Recognition to Conformational Adaptation
by: Wang, Fei, et al.
Published: (2026)
by: Wang, Fei, et al.
Published: (2026)
Shaping a Stabilized Video by Mitigating Unintended Changes for Concept-Augmented Video Editing
by: Guo, Mingce, et al.
Published: (2024)
by: Guo, Mingce, et al.
Published: (2024)
Multi-modal Learnable Queries for Image Aesthetics Assessment
by: Xiong, Zhiwei, et al.
Published: (2024)
by: Xiong, Zhiwei, et al.
Published: (2024)
Interpreting Hand gestures using Object Detection and Digits Classification
by: K, Sangeetha, et al.
Published: (2024)
by: K, Sangeetha, et al.
Published: (2024)
Modality Alignment Meets Federated Broadcasting
by: Ma, Yuting, et al.
Published: (2024)
by: Ma, Yuting, et al.
Published: (2024)
EdgeOAR: Real-time Online Action Recognition On Edge Devices
by: Luo, Wei, et al.
Published: (2024)
by: Luo, Wei, et al.
Published: (2024)
AMMSM: Adaptive Motion Magnification and Sparse Mamba for Micro-Expression Recognition
by: Liu, Xuxiong, et al.
Published: (2025)
by: Liu, Xuxiong, et al.
Published: (2025)
DEFT-LLM: Disentangled Expert Feature Tuning for Micro-Expression Recognition
by: Zhang, Ren, et al.
Published: (2025)
by: Zhang, Ren, et al.
Published: (2025)
EPIR: An Efficient Patch Tokenization, Integration and Representation Framework for Micro-expression Recognition
by: Wang, Junbo, et al.
Published: (2026)
by: Wang, Junbo, et al.
Published: (2026)
MEDN: Motion-Emotion Feature Decoupling Network for Micro-Expression Recognition
by: Hu, Chenxing, et al.
Published: (2026)
by: Hu, Chenxing, et al.
Published: (2026)
Hyperbolic Contrastive Learning for Hierarchical 3D Point Cloud Embedding
by: Liu, Yingjie, et al.
Published: (2025)
by: Liu, Yingjie, et al.
Published: (2025)
Point Cloud Resampling with Learnable Heat Diffusion
by: Xu, Wenqiang, et al.
Published: (2024)
by: Xu, Wenqiang, et al.
Published: (2024)
Similar Items
-
Online Micro-gesture Recognition Using Data Augmentation and Spatial-Temporal Attention
by: Liu, Pengyu, et al.
Published: (2025) -
Prototype Learning for Micro-gesture Classification
by: Chen, Guoliang, et al.
Published: (2024) -
MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion
by: Gu, Jihao, et al.
Published: (2025) -
MMAD: Multi-label Micro-Action Detection in Videos
by: Li, Kun, et al.
Published: (2024) -
Prototypical Calibrating Ambiguous Samples for Micro-Action Recognition
by: Li, Kun, et al.
Published: (2024)