EmoCLIP: A Vision-Language Method for Zero-Shot Video Facial Expression Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Foteinopoulou, Niki Maria, Patras, Ioannis |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VLLMs Provide Better Context for Emotion Understanding Through Common Sense Reasoning
by: Xenos, Alexandros, et al.
Published: (2024)
by: Xenos, Alexandros, et al.
Published: (2024)
Video Joint-Embedding Predictive Architectures for Facial Expression Recognition
by: Eing, Lennart, et al.
Published: (2026)
by: Eing, Lennart, et al.
Published: (2026)
FineCLIPER: Multi-modal Fine-grained CLIP for Dynamic Facial Expression Recognition with AdaptERs
by: Chen, Haodong, et al.
Published: (2024)
by: Chen, Haodong, et al.
Published: (2024)
Enhancing Zero-Shot Facial Expression Recognition by LLM Knowledge Transfer
by: Zhao, Zengqun, et al.
Published: (2024)
by: Zhao, Zengqun, et al.
Published: (2024)
Prompting Visual-Language Models for Dynamic Facial Expression Recognition
by: Zhao, Zengqun, et al.
Published: (2023)
by: Zhao, Zengqun, et al.
Published: (2023)
Exploring Thermography Technology: A Comprehensive Facial Dataset for Face Detection, Recognition, and Emotion
by: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Published: (2024)
by: Abuhussein, Mohamed Fawzi Abdelshafie, et al.
Published: (2024)
Unsupervised learning of Data-driven Facial Expression Coding System (DFECS) using keypoint tracking
by: Tripathi, Shivansh Chandra, et al.
Published: (2024)
by: Tripathi, Shivansh Chandra, et al.
Published: (2024)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
by: Voss, Hendric, et al.
Published: (2025)
by: Voss, Hendric, et al.
Published: (2025)
DEFT-LLM: Disentangled Expert Feature Tuning for Micro-Expression Recognition
by: Zhang, Ren, et al.
Published: (2025)
by: Zhang, Ren, et al.
Published: (2025)
Efficient Expression Neutrality Estimation with Application to Face Recognition Utility Prediction
by: Grimmer, Marcel, et al.
Published: (2024)
by: Grimmer, Marcel, et al.
Published: (2024)
SVFAP: Self-supervised Video Facial Affect Perceiver
by: Sun, Licai, et al.
Published: (2023)
by: Sun, Licai, et al.
Published: (2023)
Foundation Models for Zero-Shot Segmentation of Scientific Images without AI-Ready Data
by: Mukherjee, Shubhabrata, et al.
Published: (2025)
by: Mukherjee, Shubhabrata, et al.
Published: (2025)
VideoA11y: Method and Dataset for Accessible Video Description
by: Li, Chaoyu, et al.
Published: (2025)
by: Li, Chaoyu, et al.
Published: (2025)
MVTN: A Multiscale Video Transformer Network for Hand Gesture Recognition
by: Garg, Mallika, et al.
Published: (2024)
by: Garg, Mallika, et al.
Published: (2024)
A Dataset for Crucial Object Recognition in Blind and Low-Vision Individuals' Navigation
by: Islam, Md Touhidul, et al.
Published: (2024)
by: Islam, Md Touhidul, et al.
Published: (2024)
Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models
by: Chen, Tianrun, et al.
Published: (2024)
by: Chen, Tianrun, et al.
Published: (2024)
Has the Virtualization of the Face Changed Facial Perception? A Study of the Impact of Photo Editing and Augmented Reality on Facial Perception
by: Conwill, Louisa, et al.
Published: (2023)
by: Conwill, Louisa, et al.
Published: (2023)
Exploring the "Great Unseen" in Medieval Manuscripts: Instance-Level Labeling of Legacy Image Collections with Zero-Shot Models
by: Meinecke, Christofer, et al.
Published: (2025)
by: Meinecke, Christofer, et al.
Published: (2025)
Face-LLaVA: Facial Expression and Attribute Understanding through Instruction Tuning
by: Chaubey, Ashutosh, et al.
Published: (2025)
by: Chaubey, Ashutosh, et al.
Published: (2025)
Plug-and-Play Clarifier: A Zero-Shot Multimodal Framework for Egocentric Intent Disambiguation
by: Yang, Sicheng, et al.
Published: (2025)
by: Yang, Sicheng, et al.
Published: (2025)
Reading Smiles: Proxy Bias in Foundation Models for Facial Emotion Recognition
by: Tsangko, Iosif, et al.
Published: (2025)
by: Tsangko, Iosif, et al.
Published: (2025)
Multiscaled Multi-Head Attention-based Video Transformer Network for Hand Gesture Recognition
by: Garg, Mallika, et al.
Published: (2025)
by: Garg, Mallika, et al.
Published: (2025)
MathWriting: A Dataset For Handwritten Mathematical Expression Recognition
by: Gervais, Philippe, et al.
Published: (2024)
by: Gervais, Philippe, et al.
Published: (2024)
Vision Language Models as Values Detectors
by: Abbo, Giulio Antonio, et al.
Published: (2025)
by: Abbo, Giulio Antonio, et al.
Published: (2025)
egoEMOTION: Egocentric Vision and Physiological Signals for Emotion and Personality Recognition in Real-World Tasks
by: Jammot, Matthias, et al.
Published: (2025)
by: Jammot, Matthias, et al.
Published: (2025)
Customizable Avatars with Dynamic Facial Action Coded Expressions (CADyFACE) for Improved User Engagement
by: Witherow, Megan A., et al.
Published: (2024)
by: Witherow, Megan A., et al.
Published: (2024)
Facial Movement Dynamics Reveal Workload During Complex Multitasking
by: Sale, Carter, et al.
Published: (2026)
by: Sale, Carter, et al.
Published: (2026)
When Less Is More: A Sparse Facial Motion Structure For Listening Motion Learning
by: Nguyen, Tri Tung Nguyen, et al.
Published: (2025)
by: Nguyen, Tri Tung Nguyen, et al.
Published: (2025)
A Comparison of Bounding Box and Landmark Detection Methods for Video-Based Heart Rate Estimation
by: Liang, Laurence
Published: (2023)
by: Liang, Laurence
Published: (2023)
Efficient Listener: Dyadic Facial Motion Synthesis via Action Diffusion
by: Wang, Zesheng, et al.
Published: (2025)
by: Wang, Zesheng, et al.
Published: (2025)
AuraMask: An Extensible Pipeline for Developing Aesthetic Anti-Facial Recognition Image Filters
by: Lagogiannis, Jacob, et al.
Published: (2026)
by: Lagogiannis, Jacob, et al.
Published: (2026)
AV-EmoDialog: Chat with Audio-Visual Users Leveraging Emotional Cues
by: Park, Se Jin, et al.
Published: (2024)
by: Park, Se Jin, et al.
Published: (2024)
Bring Your Own Character: A Holistic Solution for Automatic Facial Animation Generation of Customized Characters
by: Bai, Zechen, et al.
Published: (2024)
by: Bai, Zechen, et al.
Published: (2024)
Exploring Emotion Expression Recognition in Older Adults Interacting with a Virtual Coach
by: Palmero, Cristina, et al.
Published: (2023)
by: Palmero, Cristina, et al.
Published: (2023)
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
by: Duan, Junwen, et al.
Published: (2025)
by: Duan, Junwen, et al.
Published: (2025)
AUGlasses: Continuous Action Unit based Facial Reconstruction with Low-power IMUs on Smart Glasses
by: Li, Yanrong, et al.
Published: (2024)
by: Li, Yanrong, et al.
Published: (2024)
MedFoundationHub: A Lightweight and Secure Toolkit for Deploying Medical Vision Language Foundation Models
by: Li, Xiao, et al.
Published: (2025)
by: Li, Xiao, et al.
Published: (2025)
EduGage: Methods and Dataset for Sensor-Based Momentary Assessment of Engagement in Self-Guided Video Learning
by: Leng, Zikang, et al.
Published: (2026)
by: Leng, Zikang, et al.
Published: (2026)
An Egocentric Vision-Language Model based Portable Real-time Smart Assistant
by: Huang, Yifei, et al.
Published: (2025)
by: Huang, Yifei, et al.
Published: (2025)
Visual Affect Analysis: Predicting Emotions of Image Viewers with Vision-Language Models
by: Nowicki, Filip, et al.
Published: (2026)
by: Nowicki, Filip, et al.
Published: (2026)
Similar Items
-
VLLMs Provide Better Context for Emotion Understanding Through Common Sense Reasoning
by: Xenos, Alexandros, et al.
Published: (2024) -
Video Joint-Embedding Predictive Architectures for Facial Expression Recognition
by: Eing, Lennart, et al.
Published: (2026) -
FineCLIPER: Multi-modal Fine-grained CLIP for Dynamic Facial Expression Recognition with AdaptERs
by: Chen, Haodong, et al.
Published: (2024) -
Enhancing Zero-Shot Facial Expression Recognition by LLM Knowledge Transfer
by: Zhao, Zengqun, et al.
Published: (2024) -
Prompting Visual-Language Models for Dynamic Facial Expression Recognition
by: Zhao, Zengqun, et al.
Published: (2023)