AU-TTT: Vision Test-Time Training model for Facial Action Unit Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Xing, Bohao, Yuan, Kaishen, Yu, Zitong, Liu, Xin, Kälviäinen, Heikki |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning
by: Xing, Bohao, et al.
Published: (2024)
by: Xing, Bohao, et al.
Published: (2024)
AU-LLM: Micro-Expression Action Unit Detection via Enhanced LLM-Based Feature Fusion
by: Liu, Zhishu, et al.
Published: (2025)
by: Liu, Zhishu, et al.
Published: (2025)
AUFormer: Vision Transformers are Parameter-Efficient Facial Action Unit Detectors
by: Yuan, Kaishen, et al.
Published: (2024)
by: Yuan, Kaishen, et al.
Published: (2024)
Insights from Visual Cognition: Understanding Human Action Dynamics with Overall Glance and Refined Gaze Transformer
by: Xing, Bohao, et al.
Published: (2026)
by: Xing, Bohao, et al.
Published: (2026)
FSBench: A Figure Skating Benchmark for Advancing Artistic Sports Understanding
by: Gao, Rong, et al.
Published: (2025)
by: Gao, Rong, et al.
Published: (2025)
Identity-free Artificial Emotional Intelligence via Micro-Gesture Understanding
by: Gao, Rong, et al.
Published: (2024)
by: Gao, Rong, et al.
Published: (2024)
DEEMO: De-identity Multimodal Emotion Recognition and Reasoning
by: Li, Deng, et al.
Published: (2025)
by: Li, Deng, et al.
Published: (2025)
EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models
by: Xing, Bohao, et al.
Published: (2025)
by: Xing, Bohao, et al.
Published: (2025)
EALD-MLLM: Emotion Analysis in Long-sequential and De-identity videos with Multi-modal Large Language Model
by: Li, Deng, et al.
Published: (2024)
by: Li, Deng, et al.
Published: (2024)
MSF-Mamba: Motion-aware State Fusion Mamba for Efficient Micro-Gesture Recognition
by: Li, Deng, et al.
Published: (2025)
by: Li, Deng, et al.
Published: (2025)
$Δ$VLA: Prior-Guided Vision-Language-Action Models via World Knowledge Variation
by: Zhu, Yijie, et al.
Published: (2026)
by: Zhu, Yijie, et al.
Published: (2026)
FEALLM: Advancing Facial Emotion Analysis in Multimodal Large Language Models with Emotional Synergy and Reasoning
by: Hu, Zhuozhao, et al.
Published: (2025)
by: Hu, Zhuozhao, et al.
Published: (2025)
Vision-TTT: Efficient and Expressive Visual Representation Learning with Test-Time Training
by: Kong, Quan, et al.
Published: (2026)
by: Kong, Quan, et al.
Published: (2026)
DiffFAS: Face Anti-Spoofing via Generative Diffusion Models
by: Ge, Xinxu, et al.
Published: (2024)
by: Ge, Xinxu, et al.
Published: (2024)
Med-TTT: Vision Test-Time Training model for Medical Image Segmentation
by: Xu, Jiashu
Published: (2024)
by: Xu, Jiashu
Published: (2024)
LoRA-TTT: Low-Rank Test-Time Training for Vision-Language Models
by: Kojima, Yuto, et al.
Published: (2025)
by: Kojima, Yuto, et al.
Published: (2025)
Hierarchical Vision-Language Interaction for Facial Action Unit Detection
by: Li, Yong, et al.
Published: (2026)
by: Li, Yong, et al.
Published: (2026)
ForgeryTTT: Zero-Shot Image Manipulation Localization with Test-Time Training
by: Liu, Weihuang, et al.
Published: (2024)
by: Liu, Weihuang, et al.
Published: (2024)
TTT3R: 3D Reconstruction as Test-Time Training
by: Chen, Xingyu, et al.
Published: (2025)
by: Chen, Xingyu, et al.
Published: (2025)
SAM-TTT: Segment Anything Model via Reverse Parameter Configuration and Test-Time Training for Camouflaged Object Detection
by: Yu, Zhenni, et al.
Published: (2025)
by: Yu, Zhenni, et al.
Published: (2025)
AULLM++: Structural Reasoning with Large Language Models for Micro-Expression Recognition
by: Liu, Zhishu, et al.
Published: (2026)
by: Liu, Zhishu, et al.
Published: (2026)
CustomTTT: Motion and Appearance Customized Video Generation via Test-Time Training
by: Bi, Xiuli, et al.
Published: (2024)
by: Bi, Xiuli, et al.
Published: (2024)
Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training
by: Liu, Fangfu, et al.
Published: (2026)
by: Liu, Fangfu, et al.
Published: (2026)
ClipTTT: CLIP-Guided Test-Time Training Helps LVLMs See Better
by: Nath, Mriganka, et al.
Published: (2026)
by: Nath, Mriganka, et al.
Published: (2026)
NC-TTT: A Noise Contrastive Approach for Test-Time Training
by: Osowiechi, David, et al.
Published: (2024)
by: Osowiechi, David, et al.
Published: (2024)
ReC-TTT: Contrastive Feature Reconstruction for Test-Time Training
by: Colussi, Marco, et al.
Published: (2024)
by: Colussi, Marco, et al.
Published: (2024)
Decoupled Doubly Contrastive Learning for Cross Domain Facial Action Unit Detection
by: Li, Yong, et al.
Published: (2025)
by: Li, Yong, et al.
Published: (2025)
DocTTT: Test-Time Training for Handwritten Document Recognition Using Meta-Auxiliary Learning
by: Gu, Wenhao, et al.
Published: (2025)
by: Gu, Wenhao, et al.
Published: (2025)
FauForensics: Boosting Audio-Visual Deepfake Detection with Facial Action Units
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
AU-vMAE: Knowledge-Guide Action Units Detection via Video Masked Autoencoder
by: Jin, Qiaoqiao, et al.
Published: (2024)
by: Jin, Qiaoqiao, et al.
Published: (2024)
Action Unit Enhance Dynamic Facial Expression Recognition
by: Liu, Feng, et al.
Published: (2025)
by: Liu, Feng, et al.
Published: (2025)
Contrastive Learning of Person-independent Representations for Facial Action Unit Detection
by: Li, Yong, et al.
Published: (2024)
by: Li, Yong, et al.
Published: (2024)
Leveraging Synthetic Data for Generalizable and Fair Facial Action Unit Detection
by: Lu, Liupei, et al.
Published: (2024)
by: Lu, Liupei, et al.
Published: (2024)
Towards Unified Facial Action Unit Recognition Framework by Large Language Models
by: Hu, Guohong, et al.
Published: (2024)
by: Hu, Guohong, et al.
Published: (2024)
Facial Action Unit Detection by Adaptively Constraining Self-Attention and Causally Deconfounding Sample
by: Shao, Zhiwen, et al.
Published: (2024)
by: Shao, Zhiwen, et al.
Published: (2024)
TTT-KD: Test-Time Training for 3D Semantic Segmentation through Knowledge Distillation from Foundation Models
by: Weijler, Lisa, et al.
Published: (2024)
by: Weijler, Lisa, et al.
Published: (2024)
Occlusion Aware Student Emotion Recognition based on Facial Action Unit Detection
by: Wally, Shrouk, et al.
Published: (2023)
by: Wally, Shrouk, et al.
Published: (2023)
MGRR-Net: Multi-level Graph Relational Reasoning Network for Facial Action Units Detection
by: Ge, Xuri, et al.
Published: (2022)
by: Ge, Xuri, et al.
Published: (2022)
UAU-Net: Uncertainty-aware Representation Learning and Evidential Classification for Facial Action Unit Detection
by: Li, Yuze, et al.
Published: (2026)
by: Li, Yuze, et al.
Published: (2026)
TTT-Unet: Enhancing U-Net with Test-Time Training Layers for Biomedical Image Segmentation
by: Zhou, Rong, et al.
Published: (2024)
by: Zhou, Rong, et al.
Published: (2024)
Similar Items
-
EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning
by: Xing, Bohao, et al.
Published: (2024) -
AU-LLM: Micro-Expression Action Unit Detection via Enhanced LLM-Based Feature Fusion
by: Liu, Zhishu, et al.
Published: (2025) -
AUFormer: Vision Transformers are Parameter-Efficient Facial Action Unit Detectors
by: Yuan, Kaishen, et al.
Published: (2024) -
Insights from Visual Cognition: Understanding Human Action Dynamics with Overall Glance and Refined Gaze Transformer
by: Xing, Bohao, et al.
Published: (2026) -
FSBench: A Figure Skating Benchmark for Advancing Artistic Sports Understanding
by: Gao, Rong, et al.
Published: (2025)