AUFormer: Vision Transformers are Parameter-Efficient Facial Action Unit Detectors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yuan, Kaishen, Yu, Zitong, Liu, Xin, Xie, Weicheng, Yue, Huanjing, Yang, Jingyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning
von: Xing, Bohao, et al.
Veröffentlicht: (2024)
von: Xing, Bohao, et al.
Veröffentlicht: (2024)
AU-TTT: Vision Test-Time Training model for Facial Action Unit Detection
von: Xing, Bohao, et al.
Veröffentlicht: (2025)
von: Xing, Bohao, et al.
Veröffentlicht: (2025)
FEALLM: Advancing Facial Emotion Analysis in Multimodal Large Language Models with Emotional Synergy and Reasoning
von: Hu, Zhuozhao, et al.
Veröffentlicht: (2025)
von: Hu, Zhuozhao, et al.
Veröffentlicht: (2025)
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation
von: Liu, Xin, et al.
Veröffentlicht: (2024)
von: Liu, Xin, et al.
Veröffentlicht: (2024)
A Simple yet Effective Network based on Vision Transformer for Camouflaged Object and Salient Object Detection
von: Hao, Chao, et al.
Veröffentlicht: (2024)
von: Hao, Chao, et al.
Veröffentlicht: (2024)
Adversarial Robustness in RGB-Skeleton Action Recognition: Leveraging Attention Modality Reweighter
von: Liu, Chao, et al.
Veröffentlicht: (2024)
von: Liu, Chao, et al.
Veröffentlicht: (2024)
Distribution-Specific Learning for Joint Salient and Camouflaged Object Detection
von: Hao, Chao, et al.
Veröffentlicht: (2025)
von: Hao, Chao, et al.
Veröffentlicht: (2025)
AU-LLM: Micro-Expression Action Unit Detection via Enhanced LLM-Based Feature Fusion
von: Liu, Zhishu, et al.
Veröffentlicht: (2025)
von: Liu, Zhishu, et al.
Veröffentlicht: (2025)
RViDeformer: Efficient Raw Video Denoising Transformer with a Larger Benchmark Dataset
von: Yue, Huanjing, et al.
Veröffentlicht: (2023)
von: Yue, Huanjing, et al.
Veröffentlicht: (2023)
$Δ$VLA: Prior-Guided Vision-Language-Action Models via World Knowledge Variation
von: Zhu, Yijie, et al.
Veröffentlicht: (2026)
von: Zhu, Yijie, et al.
Veröffentlicht: (2026)
Zero-Shot Video Restoration and Enhancement with Assistance of Video Diffusion Models
von: Cao, Cong, et al.
Veröffentlicht: (2026)
von: Cao, Cong, et al.
Veröffentlicht: (2026)
CA-Edit: Causality-Aware Condition Adapter for High-Fidelity Local Facial Attribute Editing
von: Xian, Xiaole, et al.
Veröffentlicht: (2024)
von: Xian, Xiaole, et al.
Veröffentlicht: (2024)
Multi-scale Dynamic and Hierarchical Relationship Modeling for Facial Action Units Recognition
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
von: Wang, Zihan, et al.
Veröffentlicht: (2024)
Zero-Shot Video Restoration and Enhancement Using Pre-Trained Image Diffusion Model
von: Cao, Cong, et al.
Veröffentlicht: (2024)
von: Cao, Cong, et al.
Veröffentlicht: (2024)
Hierarchical Vision-Language Interaction for Facial Action Unit Detection
von: Li, Yong, et al.
Veröffentlicht: (2026)
von: Li, Yong, et al.
Veröffentlicht: (2026)
PA-FAS: Towards Interpretable and Generalizable Multimodal Face Anti-Spoofing via Path-Augmented Reinforcement Learning
von: Ma, Yingjie, et al.
Veröffentlicht: (2025)
von: Ma, Yingjie, et al.
Veröffentlicht: (2025)
TAG: Thinking with Action Unit Grounding for Facial Expression Recognition
von: Lin, Haobo, et al.
Veröffentlicht: (2026)
von: Lin, Haobo, et al.
Veröffentlicht: (2026)
AULLM++: Structural Reasoning with Large Language Models for Micro-Expression Recognition
von: Liu, Zhishu, et al.
Veröffentlicht: (2026)
von: Liu, Zhishu, et al.
Veröffentlicht: (2026)
KeDuSR: Real-World Dual-Lens Super-Resolution via Kernel-Free Matching
von: Yue, Huanjing, et al.
Veröffentlicht: (2023)
von: Yue, Huanjing, et al.
Veröffentlicht: (2023)
F2HDR: Two-Stage HDR Video Reconstruction via Flow Adapter and Physical Motion Modeling
von: Yue, Huanjing, et al.
Veröffentlicht: (2026)
von: Yue, Huanjing, et al.
Veröffentlicht: (2026)
Learning Contrastive Feature Representations for Facial Action Unit Detection
von: Shang, Ziqiao, et al.
Veröffentlicht: (2024)
von: Shang, Ziqiao, et al.
Veröffentlicht: (2024)
RISAM: Referring Image Segmentation via Mutual-Aware Attention Features
von: Zhang, Mengxi, et al.
Veröffentlicht: (2023)
von: Zhang, Mengxi, et al.
Veröffentlicht: (2023)
Boosting Facial Action Unit Detection Through Jointly Learning Facial Landmark Detection and Domain Separation and Reconstruction
von: Shang, Ziqiao, et al.
Veröffentlicht: (2023)
von: Shang, Ziqiao, et al.
Veröffentlicht: (2023)
CameraMaster: Unified Camera Semantic-Parameter Control for Photography Retouching
von: Yang, Qirui, et al.
Veröffentlicht: (2025)
von: Yang, Qirui, et al.
Veröffentlicht: (2025)
DeeDSR: Towards Real-World Image Super-Resolution via Degradation-Aware Stable Diffusion
von: Bi, Chunyang, et al.
Veröffentlicht: (2024)
von: Bi, Chunyang, et al.
Veröffentlicht: (2024)
Efficient HDR Reconstruction from Real-World Raw Images
von: Yang, Qirui, et al.
Veröffentlicht: (2023)
von: Yang, Qirui, et al.
Veröffentlicht: (2023)
Enhancing Adversarial Transferability by Balancing Exploration and Exploitation with Gradient-Guided Sampling
von: Niu, Zenghao, et al.
Veröffentlicht: (2025)
von: Niu, Zenghao, et al.
Veröffentlicht: (2025)
One-Frame Calibration with Siamese Network in Facial Action Unit Recognition
von: Feng, Shuangquan, et al.
Veröffentlicht: (2024)
von: Feng, Shuangquan, et al.
Veröffentlicht: (2024)
Accelerating Vision Transformers on Brain Processing Unit
von: Tang, Jinchi, et al.
Veröffentlicht: (2026)
von: Tang, Jinchi, et al.
Veröffentlicht: (2026)
Denoising and Alignment: Rethinking Domain Generalization for Multimodal Face Anti-Spoofing
von: Ma, Yingjie, et al.
Veröffentlicht: (2025)
von: Ma, Yingjie, et al.
Veröffentlicht: (2025)
MALT: Multi-scale Action Learning Transformer for Online Action Detection
von: Yang, Zhipeng, et al.
Veröffentlicht: (2024)
von: Yang, Zhipeng, et al.
Veröffentlicht: (2024)
Action Unit Enhance Dynamic Facial Expression Recognition
von: Liu, Feng, et al.
Veröffentlicht: (2025)
von: Liu, Feng, et al.
Veröffentlicht: (2025)
Efficient Adaptation of Pre-trained Vision Transformer via Householder Transformation
von: Dong, Wei, et al.
Veröffentlicht: (2024)
von: Dong, Wei, et al.
Veröffentlicht: (2024)
HalluRNN: Mitigating Hallucinations via Recurrent Cross-Layer Reasoning in Large Vision-Language Models
von: Yu, Le, et al.
Veröffentlicht: (2025)
von: Yu, Le, et al.
Veröffentlicht: (2025)
SpiralDiff: Spiral Diffusion with LoRA for RGB-to-RAW Conversion Across Cameras
von: Yue, Huanjing, et al.
Veröffentlicht: (2026)
von: Yue, Huanjing, et al.
Veröffentlicht: (2026)
CF-VLA: Efficient Coarse-to-Fine Action Generation for Vision-Language-Action Policies
von: Du, Fan, et al.
Veröffentlicht: (2026)
von: Du, Fan, et al.
Veröffentlicht: (2026)
Decoupled Doubly Contrastive Learning for Cross Domain Facial Action Unit Detection
von: Li, Yong, et al.
Veröffentlicht: (2025)
von: Li, Yong, et al.
Veröffentlicht: (2025)
SFDA-rPPG: Source-Free Domain Adaptive Remote Physiological Measurement with Spatio-Temporal Consistency
von: Xie, Yiping, et al.
Veröffentlicht: (2024)
von: Xie, Yiping, et al.
Veröffentlicht: (2024)
Vision-based Discovery of Nonlinear Dynamics for 3D Moving Target
von: Zhang, Zitong, et al.
Veröffentlicht: (2024)
von: Zhang, Zitong, et al.
Veröffentlicht: (2024)
Probing the Efficacy of Federated Parameter-Efficient Fine-Tuning of Vision Transformers for Medical Image Classification
von: Alkhunaizi, Naif, et al.
Veröffentlicht: (2024)
von: Alkhunaizi, Naif, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning
von: Xing, Bohao, et al.
Veröffentlicht: (2024) -
AU-TTT: Vision Test-Time Training model for Facial Action Unit Detection
von: Xing, Bohao, et al.
Veröffentlicht: (2025) -
FEALLM: Advancing Facial Emotion Analysis in Multimodal Large Language Models with Emotional Synergy and Reasoning
von: Hu, Zhuozhao, et al.
Veröffentlicht: (2025) -
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation
von: Liu, Xin, et al.
Veröffentlicht: (2024) -
A Simple yet Effective Network based on Vision Transformer for Camouflaged Object and Salient Object Detection
von: Hao, Chao, et al.
Veröffentlicht: (2024)