Hierarchical Vision-Language Interaction for Facial Action Unit Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yong, Ren, Yi, Zhang, Yizhe, Zhang, Wenhua, Zhang, Tianyi, Jiang, Muyun, Xie, Guo-Sen, Guan, Cuntai |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Decoupled Doubly Contrastive Learning for Cross Domain Facial Action Unit Detection
by: Li, Yong, et al.
Published: (2025)
by: Li, Yong, et al.
Published: (2025)
Decoupled Hierarchical Distillation for Multimodal Emotion Recognition
by: Li, Yong, et al.
Published: (2026)
by: Li, Yong, et al.
Published: (2026)
Contrastive Learning of Person-independent Representations for Facial Action Unit Detection
by: Li, Yong, et al.
Published: (2024)
by: Li, Yong, et al.
Published: (2024)
TAG: Thinking with Action Unit Grounding for Facial Expression Recognition
by: Lin, Haobo, et al.
Published: (2026)
by: Lin, Haobo, et al.
Published: (2026)
Multi-scale Dynamic and Hierarchical Relationship Modeling for Facial Action Units Recognition
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
Beyond Overfitting: Doubly Adaptive Dropout for Generalizable AU Detection
by: Li, Yong, et al.
Published: (2025)
by: Li, Yong, et al.
Published: (2025)
AU-TTT: Vision Test-Time Training model for Facial Action Unit Detection
by: Xing, Bohao, et al.
Published: (2025)
by: Xing, Bohao, et al.
Published: (2025)
AUFormer: Vision Transformers are Parameter-Efficient Facial Action Unit Detectors
by: Yuan, Kaishen, et al.
Published: (2024)
by: Yuan, Kaishen, et al.
Published: (2024)
Towards Unified Facial Action Unit Recognition Framework by Large Language Models
by: Hu, Guohong, et al.
Published: (2024)
by: Hu, Guohong, et al.
Published: (2024)
Facial Action Unit Detection by Adaptively Constraining Self-Attention and Causally Deconfounding Sample
by: Shao, Zhiwen, et al.
Published: (2024)
by: Shao, Zhiwen, et al.
Published: (2024)
Towards End-to-End Explainable Facial Action Unit Recognition via Vision-Language Joint Learning
by: Ge, Xuri, et al.
Published: (2024)
by: Ge, Xuri, et al.
Published: (2024)
Unveiling Deep Semantic Uncertainty Perception for Language-Anchored Multi-modal Vision-Brain Alignment
by: Feng, Zehui, et al.
Published: (2025)
by: Feng, Zehui, et al.
Published: (2025)
Leveraging Synthetic Data for Generalizable and Fair Facial Action Unit Detection
by: Lu, Liupei, et al.
Published: (2024)
by: Lu, Liupei, et al.
Published: (2024)
Bidirectional Learning of Facial Action Units and Expressions via Structured Semantic Mapping across Heterogeneous Datasets
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
Hierarchical Action Recognition: A Contrastive Video-Language Approach with Hierarchical Interactions
by: Zhang, Rui, et al.
Published: (2024)
by: Zhang, Rui, et al.
Published: (2024)
Capturing the Unseen: Vision-Free Facial Motion Capture Using Inertial Measurement Units
by: Wang, Youjia, et al.
Published: (2024)
by: Wang, Youjia, et al.
Published: (2024)
Learning Contrastive Feature Representations for Facial Action Unit Detection
by: Shang, Ziqiao, et al.
Published: (2024)
by: Shang, Ziqiao, et al.
Published: (2024)
Occlusion Aware Student Emotion Recognition based on Facial Action Unit Detection
by: Wally, Shrouk, et al.
Published: (2023)
by: Wally, Shrouk, et al.
Published: (2023)
FauForensics: Boosting Audio-Visual Deepfake Detection with Facial Action Units
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
Trend-Aware Supervision: On Learning Invariance for Semi-Supervised Facial Action Unit Intensity Estimation
by: Chen, Yingjie, et al.
Published: (2025)
by: Chen, Yingjie, et al.
Published: (2025)
Guide, Think, Act: Interactive Embodied Reasoning in Vision-Language-Action Models
by: Ling, Yiran, et al.
Published: (2026)
by: Ling, Yiran, et al.
Published: (2026)
Action Unit Enhance Dynamic Facial Expression Recognition
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
MagicFace: High-Fidelity Facial Expression Editing with Action-Unit Control
by: Wei, Mengting, et al.
Published: (2025)
by: Wei, Mengting, et al.
Published: (2025)
Grounding Actions in Camera Space: Observation-Centric Vision-Language-Action Policy
by: Zhang, Tianyi, et al.
Published: (2025)
by: Zhang, Tianyi, et al.
Published: (2025)
Talking Head Generation Driven by Speech-Related Facial Action Units and Audio- Based on Multimodal Representation Fusion
by: Chen, Sen, et al.
Published: (2022)
by: Chen, Sen, et al.
Published: (2022)
Causal Intervention for Subject-Deconfounded Facial Action Unit Recognition
by: Chen, Yingjie, et al.
Published: (2022)
by: Chen, Yingjie, et al.
Published: (2022)
MGRR-Net: Multi-level Graph Relational Reasoning Network for Facial Action Units Detection
by: Ge, Xuri, et al.
Published: (2022)
by: Ge, Xuri, et al.
Published: (2022)
GeoConv: Geodesic Guided Convolution for Facial Action Unit Recognition
by: Chen, Yuedong, et al.
Published: (2020)
by: Chen, Yuedong, et al.
Published: (2020)
Exploring Facial Biomarkers for Depression through Temporal Analysis of Action Units
by: Parikh, Aditya, et al.
Published: (2024)
by: Parikh, Aditya, et al.
Published: (2024)
AUGlasses: Continuous Action Unit based Facial Reconstruction with Low-power IMUs on Smart Glasses
by: Li, Yanrong, et al.
Published: (2024)
by: Li, Yanrong, et al.
Published: (2024)
Hugging Rain Man: A Novel Facial Action Units Dataset for Analyzing Atypical Facial Expressions in Children with Autism Spectrum Disorder
by: Ji, Yanfeng, et al.
Published: (2024)
by: Ji, Yanfeng, et al.
Published: (2024)
UAU-Net: Uncertainty-aware Representation Learning and Evidential Classification for Facial Action Unit Detection
by: Li, Yuze, et al.
Published: (2026)
by: Li, Yuze, et al.
Published: (2026)
Diffusion Facial Forgery Detection
by: Cheng, Harry, et al.
Published: (2024)
by: Cheng, Harry, et al.
Published: (2024)
Exploring the Limits of End-to-End Feature-Affinity Propagation for Single-Point Supervised Infrared Small Target Detection
by: Zhou, Qiancheng, et al.
Published: (2026)
by: Zhou, Qiancheng, et al.
Published: (2026)
Guided Interpretable Facial Expression Recognition via Spatial Action Unit Cues
by: Belharbi, Soufiane, et al.
Published: (2024)
by: Belharbi, Soufiane, et al.
Published: (2024)
Training-Free Zero-Shot Temporal Action Detection with Vision-Language Models
by: Han, Chaolei, et al.
Published: (2025)
by: Han, Chaolei, et al.
Published: (2025)
Bidirectional Channel-selective Semantic Interaction for Semi-Supervised Medical Segmentation
by: Huang, Kaiwen, et al.
Published: (2026)
by: Huang, Kaiwen, et al.
Published: (2026)
Boosting Facial Action Unit Detection Through Jointly Learning Facial Landmark Detection and Domain Separation and Reconstruction
by: Shang, Ziqiao, et al.
Published: (2023)
by: Shang, Ziqiao, et al.
Published: (2023)
Hierarchical Spatio-temporal Segmentation Network for Ejection Fraction Estimation in Echocardiography Videos
by: Wang, Dongfang, et al.
Published: (2025)
by: Wang, Dongfang, et al.
Published: (2025)
AUEditNet: Dual-Branch Facial Action Unit Intensity Manipulation with Implicit Disentanglement
by: Jin, Shiwei, et al.
Published: (2024)
by: Jin, Shiwei, et al.
Published: (2024)
Similar Items
-
Decoupled Doubly Contrastive Learning for Cross Domain Facial Action Unit Detection
by: Li, Yong, et al.
Published: (2025) -
Decoupled Hierarchical Distillation for Multimodal Emotion Recognition
by: Li, Yong, et al.
Published: (2026) -
Contrastive Learning of Person-independent Representations for Facial Action Unit Detection
by: Li, Yong, et al.
Published: (2024) -
TAG: Thinking with Action Unit Grounding for Facial Expression Recognition
by: Lin, Haobo, et al.
Published: (2026) -
Multi-scale Dynamic and Hierarchical Relationship Modeling for Facial Action Units Recognition
by: Wang, Zihan, et al.
Published: (2024)