Saved in:
| Main Authors: | Li, Deng, Xing, Bohao, Liu, Xin, Xia, Baiqiang, Wen, Bihan, Kälviäinen, Heikki |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2504.19549 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
EALD-MLLM: Emotion Analysis in Long-sequential and De-identity videos with Multi-modal Large Language Model
by: Li, Deng, et al.
Published: (2024)
by: Li, Deng, et al.
Published: (2024)
MSF-Mamba: Motion-aware State Fusion Mamba for Efficient Micro-Gesture Recognition
by: Li, Deng, et al.
Published: (2025)
by: Li, Deng, et al.
Published: (2025)
FSBench: A Figure Skating Benchmark for Advancing Artistic Sports Understanding
by: Gao, Rong, et al.
Published: (2025)
by: Gao, Rong, et al.
Published: (2025)
EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models
by: Xing, Bohao, et al.
Published: (2025)
by: Xing, Bohao, et al.
Published: (2025)
Insights from Visual Cognition: Understanding Human Action Dynamics with Overall Glance and Refined Gaze Transformer
by: Xing, Bohao, et al.
Published: (2026)
by: Xing, Bohao, et al.
Published: (2026)
AU-TTT: Vision Test-Time Training model for Facial Action Unit Detection
by: Xing, Bohao, et al.
Published: (2025)
by: Xing, Bohao, et al.
Published: (2025)
Identity-free Artificial Emotional Intelligence via Micro-Gesture Understanding
by: Gao, Rong, et al.
Published: (2024)
by: Gao, Rong, et al.
Published: (2024)
Enhancing Micro Gesture Recognition for Emotion Understanding via Context-aware Visual-Text Contrastive Learning
by: Li, Deng, et al.
Published: (2024)
by: Li, Deng, et al.
Published: (2024)
EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning
by: Xing, Bohao, et al.
Published: (2024)
by: Xing, Bohao, et al.
Published: (2024)
Self-Supervised Pretraining for Fine-Grained Plankton Recognition
by: Kareinen, Joona, et al.
Published: (2025)
by: Kareinen, Joona, et al.
Published: (2025)
Open-Set Plankton Recognition
by: Kareinen, Joona, et al.
Published: (2025)
by: Kareinen, Joona, et al.
Published: (2025)
DiffFAS: Face Anti-Spoofing via Generative Diffusion Models
by: Ge, Xinxu, et al.
Published: (2024)
by: Ge, Xinxu, et al.
Published: (2024)
TiCAL:Typicality-Based Consistency-Aware Learning for Multimodal Emotion Recognition
by: Yin, Wen, et al.
Published: (2025)
by: Yin, Wen, et al.
Published: (2025)
Beyond Emotion Recognition: A Multi-Turn Multimodal Emotion Understanding and Reasoning Benchmark
by: Hu, Jinpeng, et al.
Published: (2025)
by: Hu, Jinpeng, et al.
Published: (2025)
AffectAgent: Collaborative Multi-Agent Reasoning for Retrieval-Augmented Multimodal Emotion Recognition
by: Wang, Zeheng, et al.
Published: (2026)
by: Wang, Zeheng, et al.
Published: (2026)
DAPlankton: Benchmark Dataset for Multi-instrument Plankton Recognition via Fine-grained Domain Adaptation
by: Batrakhanov, Daniel, et al.
Published: (2024)
by: Batrakhanov, Daniel, et al.
Published: (2024)
RSGround-R1: Rethinking Remote Sensing Visual Grounding through Spatial Reasoning
by: Huang, Shiqi, et al.
Published: (2026)
by: Huang, Shiqi, et al.
Published: (2026)
Unsupervised Pelage Pattern Unwrapping for Animal Re-identification
by: Algasov, Aleksandr, et al.
Published: (2025)
by: Algasov, Aleksandr, et al.
Published: (2025)
XEmoGPT: An Explainable Multimodal Emotion Recognition Framework with Cue-Level Perception and Reasoning
by: Zhang, Hanwen, et al.
Published: (2026)
by: Zhang, Hanwen, et al.
Published: (2026)
Multimodal Video Emotion Recognition with Reliable Reasoning Priors
by: Wang, Zhepeng, et al.
Published: (2025)
by: Wang, Zhepeng, et al.
Published: (2025)
Cross-modal learning for plankton recognition
by: Kareinen, Joona, et al.
Published: (2026)
by: Kareinen, Joona, et al.
Published: (2026)
GroundFlow: A Plug-in Module for Temporal Reasoning on 3D Point Cloud Sequential Grounding
by: Lin, Zijun, et al.
Published: (2025)
by: Lin, Zijun, et al.
Published: (2025)
A Trustworthy Method for Multimodal Emotion Recognition
by: Xue, Junxiao, et al.
Published: (2025)
by: Xue, Junxiao, et al.
Published: (2025)
Decoupled Hierarchical Distillation for Multimodal Emotion Recognition
by: Li, Yong, et al.
Published: (2026)
by: Li, Yong, et al.
Published: (2026)
FEALLM: Advancing Facial Emotion Analysis in Multimodal Large Language Models with Emotional Synergy and Reasoning
by: Hu, Zhuozhao, et al.
Published: (2025)
by: Hu, Zhuozhao, et al.
Published: (2025)
Facial-R1: Aligning Reasoning and Recognition for Facial Emotion Analysis
by: Wu, Jiulong, et al.
Published: (2025)
by: Wu, Jiulong, et al.
Published: (2025)
Feature-Based Dual Visual Feature Extraction Model for Compound Multimodal Emotion Recognition
by: Liu, Ran, et al.
Published: (2025)
by: Liu, Ran, et al.
Published: (2025)
Attribute-Grounded Selective Reasoning for Artwork Emotion Understanding with Multimodal Large Language Models
by: Zhang, Cheng, et al.
Published: (2026)
by: Zhang, Cheng, et al.
Published: (2026)
Leveraging CLIP Encoder for Multimodal Emotion Recognition
by: Song, Yehun, et al.
Published: (2025)
by: Song, Yehun, et al.
Published: (2025)
Complementarity-Supervised Spectral-Band Routing for Multimodal Emotion Recognition
by: Huang, Zhexian, et al.
Published: (2026)
by: Huang, Zhexian, et al.
Published: (2026)
Understanding the Impact of Training Set Size on Animal Re-identification
by: Algasov, Aleksandr, et al.
Published: (2024)
by: Algasov, Aleksandr, et al.
Published: (2024)
ZoRI: Towards Discriminative Zero-Shot Remote Sensing Instance Segmentation
by: Huang, Shiqi, et al.
Published: (2024)
by: Huang, Shiqi, et al.
Published: (2024)
RF4D:Neural Radar Fields for Novel View Synthesis in Outdoor Dynamic Scenes
by: Zhang, Jiarui, et al.
Published: (2025)
by: Zhang, Jiarui, et al.
Published: (2025)
Navigating the Emotion Tree: Hierarchical Hyperbolic RAG for Multimodal Emotion Recognition
by: Wang, Zeheng, et al.
Published: (2026)
by: Wang, Zeheng, et al.
Published: (2026)
Enhancing Modal Fusion by Alignment and Label Matching for Multimodal Emotion Recognition
by: Li, Qifei, et al.
Published: (2024)
by: Li, Qifei, et al.
Published: (2024)
Calibrating Multimodal Consensus for Emotion Recognition
by: Zhong, Guowei, et al.
Published: (2025)
by: Zhong, Guowei, et al.
Published: (2025)
LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation
by: Chern, Ethan, et al.
Published: (2025)
by: Chern, Ethan, et al.
Published: (2025)
LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model
by: Yang, Senqiao, et al.
Published: (2023)
by: Yang, Senqiao, et al.
Published: (2023)
Video Emotion Open-vocabulary Recognition Based on Multimodal Large Language Model
by: Ge, Mengying, et al.
Published: (2024)
by: Ge, Mengying, et al.
Published: (2024)
VISTANet: VIsual Spoken Textual Additive Net for Interpretable Multimodal Emotion Recognition
by: Kumar, Puneet, et al.
Published: (2022)
by: Kumar, Puneet, et al.
Published: (2022)
Similar Items
-
EALD-MLLM: Emotion Analysis in Long-sequential and De-identity videos with Multi-modal Large Language Model
by: Li, Deng, et al.
Published: (2024) -
MSF-Mamba: Motion-aware State Fusion Mamba for Efficient Micro-Gesture Recognition
by: Li, Deng, et al.
Published: (2025) -
FSBench: A Figure Skating Benchmark for Advancing Artistic Sports Understanding
by: Gao, Rong, et al.
Published: (2025) -
EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models
by: Xing, Bohao, et al.
Published: (2025) -
Insights from Visual Cognition: Understanding Human Action Dynamics with Overall Glance and Refined Gaze Transformer
by: Xing, Bohao, et al.
Published: (2026)