VISTANet: VIsual Spoken Textual Additive Net for Interpretable Multimodal Emotion Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Puneet, Malik, Sarthak, Raman, Balasubramanian, Li, Xiaobai |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Interpretable Image Emotion Recognition: A Domain Adaptation Approach Using Facial Expressions
by: Kumar, Puneet, et al.
Published: (2020)
by: Kumar, Puneet, et al.
Published: (2020)
GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos
by: Kumar, Deepak, et al.
Published: (2026)
by: Kumar, Deepak, et al.
Published: (2026)
TCCT-Net: Two-Stream Network Architecture for Fast and Efficient Engagement Estimation via Behavioral Feature Signals
by: Vedernikov, Alexander, et al.
Published: (2024)
by: Vedernikov, Alexander, et al.
Published: (2024)
Synthesizing Sentiment-Controlled Feedback For Multimodal Text and Image Data
by: Kumar, Puneet, et al.
Published: (2024)
by: Kumar, Puneet, et al.
Published: (2024)
Textualized and Feature-based Models for Compound Multimodal Emotion Recognition in the Wild
by: Richet, Nicolas, et al.
Published: (2024)
by: Richet, Nicolas, et al.
Published: (2024)
Vision Large Language Models Are Good Noise Handlers in Engagement Analysis
by: Vedernikov, Alexander, et al.
Published: (2025)
by: Vedernikov, Alexander, et al.
Published: (2025)
VisioPhysioENet: Visual Physiological Engagement Detection Network
by: Singh, Alakhsimar, et al.
Published: (2024)
by: Singh, Alakhsimar, et al.
Published: (2024)
Visual and Textual Prompts in VLLMs for Enhancing Emotion Recognition
by: Wang, Zhifeng, et al.
Published: (2025)
by: Wang, Zhifeng, et al.
Published: (2025)
VISOR: VIsual Spatial Object Reasoning for Language-driven Object Navigation
by: Taioli, Francesco, et al.
Published: (2026)
by: Taioli, Francesco, et al.
Published: (2026)
Decoupled Hierarchical Distillation for Multimodal Emotion Recognition
by: Li, Yong, et al.
Published: (2026)
by: Li, Yong, et al.
Published: (2026)
Active Multimodal Distillation for Few-shot Action Recognition
by: Feng, Weijia, et al.
Published: (2025)
by: Feng, Weijia, et al.
Published: (2025)
VISTA: A Visual and Textual Attention Dataset for Interpreting Multimodal Models
by: Harshit, et al.
Published: (2024)
by: Harshit, et al.
Published: (2024)
Biasing & Debiasing based Approach Towards Fair Knowledge Transfer for Equitable Skin Analysis
by: Pundhir, Anshul, et al.
Published: (2024)
by: Pundhir, Anshul, et al.
Published: (2024)
Contrast-Phys: Unsupervised Video-based Remote Physiological Measurement via Spatiotemporal Contrast
by: Sun, Zhaodong, et al.
Published: (2022)
by: Sun, Zhaodong, et al.
Published: (2022)
Contrast-Phys+: Unsupervised and Weakly-supervised Video-based Remote Physiological Measurement via Spatiotemporal Contrast
by: Sun, Zhaodong, et al.
Published: (2023)
by: Sun, Zhaodong, et al.
Published: (2023)
Rethinking Skip Connections: Additive U-Net for Robust and Interpretable Denoising
by: Lakkavalli, Vikram R
Published: (2026)
by: Lakkavalli, Vikram R
Published: (2026)
DEEMO: De-identity Multimodal Emotion Recognition and Reasoning
by: Li, Deng, et al.
Published: (2025)
by: Li, Deng, et al.
Published: (2025)
A Trustworthy Method for Multimodal Emotion Recognition
by: Xue, Junxiao, et al.
Published: (2025)
by: Xue, Junxiao, et al.
Published: (2025)
Leveraging CLIP Encoder for Multimodal Emotion Recognition
by: Song, Yehun, et al.
Published: (2025)
by: Song, Yehun, et al.
Published: (2025)
AGCD-Net: Attention Guided Context Debiasing Network for Emotion Recognition
by: Devi, Varsha, et al.
Published: (2025)
by: Devi, Varsha, et al.
Published: (2025)
Beyond Emotion Recognition: A Multi-Turn Multimodal Emotion Understanding and Reasoning Benchmark
by: Hu, Jinpeng, et al.
Published: (2025)
by: Hu, Jinpeng, et al.
Published: (2025)
A Benchmark for Incremental Micro-expression Recognition
by: Lai, Zhengqin, et al.
Published: (2025)
by: Lai, Zhengqin, et al.
Published: (2025)
Anomaly Triplet-Net: Progress Recognition Model Using Deep Metric Learning Considering Occlusion for Manual Assembly Work
by: Kitsukawa, Takumi, et al.
Published: (2025)
by: Kitsukawa, Takumi, et al.
Published: (2025)
Video Emotion Open-vocabulary Recognition Based on Multimodal Large Language Model
by: Ge, Mengying, et al.
Published: (2024)
by: Ge, Mengying, et al.
Published: (2024)
Complementarity-Supervised Spectral-Band Routing for Multimodal Emotion Recognition
by: Huang, Zhexian, et al.
Published: (2026)
by: Huang, Zhexian, et al.
Published: (2026)
ResEmoteNet: Bridging Accuracy and Loss Reduction in Facial Emotion Recognition
by: Roy, Arnab Kumar, et al.
Published: (2024)
by: Roy, Arnab Kumar, et al.
Published: (2024)
TiCAL:Typicality-Based Consistency-Aware Learning for Multimodal Emotion Recognition
by: Yin, Wen, et al.
Published: (2025)
by: Yin, Wen, et al.
Published: (2025)
TexTAR : Textual Attribute Recognition in Multi-domain and Multi-lingual Document Images
by: Kumar, Rohan, et al.
Published: (2025)
by: Kumar, Rohan, et al.
Published: (2025)
Calibrating Multimodal Consensus for Emotion Recognition
by: Zhong, Guowei, et al.
Published: (2025)
by: Zhong, Guowei, et al.
Published: (2025)
Multimodal Emotion Recognition via Causal-Diffusion Bridge (Affect-Diff)
by: Sanjyal, Ankit
Published: (2026)
by: Sanjyal, Ankit
Published: (2026)
MultiMAE-DER: Multimodal Masked Autoencoder for Dynamic Emotion Recognition
by: Xiang, Peihao, et al.
Published: (2024)
by: Xiang, Peihao, et al.
Published: (2024)
MCN-CL: Multimodal Cross-Attention Network and Contrastive Learning for Multimodal Emotion Recognition
by: Li, Feng, et al.
Published: (2025)
by: Li, Feng, et al.
Published: (2025)
MarkushGrapher: Joint Visual and Textual Recognition of Markush Structures
by: Morin, Lucas, et al.
Published: (2025)
by: Morin, Lucas, et al.
Published: (2025)
DMD-Net: Deep Mesh Denoising Network
by: Gangopadhyay, Aalok, et al.
Published: (2025)
by: Gangopadhyay, Aalok, et al.
Published: (2025)
Digital Twins in Additive Manufacturing: A Systematic Review
by: Ahsan, Md Manjurul, et al.
Published: (2024)
by: Ahsan, Md Manjurul, et al.
Published: (2024)
Navigating the Emotion Tree: Hierarchical Hyperbolic RAG for Multimodal Emotion Recognition
by: Wang, Zeheng, et al.
Published: (2026)
by: Wang, Zeheng, et al.
Published: (2026)
ICANet: A Method of Short Video Emotion Recognition Driven by Multimodal Data
by: Wu, Xuecheng, et al.
Published: (2022)
by: Wu, Xuecheng, et al.
Published: (2022)
Beyond Images: Adaptive Fusion of Visual and Textual Data for Food Classification
by: Mittal, Prateek, et al.
Published: (2023)
by: Mittal, Prateek, et al.
Published: (2023)
Multimodal Prompt Learning with Missing Modalities for Sentiment Analysis and Emotion Recognition
by: Guo, Zirun, et al.
Published: (2024)
by: Guo, Zirun, et al.
Published: (2024)
Multimodal Video Emotion Recognition with Reliable Reasoning Priors
by: Wang, Zhepeng, et al.
Published: (2025)
by: Wang, Zhepeng, et al.
Published: (2025)
Similar Items
-
Interpretable Image Emotion Recognition: A Domain Adaptation Approach Using Facial Expressions
by: Kumar, Puneet, et al.
Published: (2020) -
GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos
by: Kumar, Deepak, et al.
Published: (2026) -
TCCT-Net: Two-Stream Network Architecture for Fast and Efficient Engagement Estimation via Behavioral Feature Signals
by: Vedernikov, Alexander, et al.
Published: (2024) -
Synthesizing Sentiment-Controlled Feedback For Multimodal Text and Image Data
by: Kumar, Puneet, et al.
Published: (2024) -
Textualized and Feature-based Models for Compound Multimodal Emotion Recognition in the Wild
by: Richet, Nicolas, et al.
Published: (2024)