HiQuE: Hierarchical Question Embedding Network for Multimodal Depression Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Jung, Juho, Kang, Chaewon, Yoon, Jeewoo, Kim, Seungbae, Han, Jinyoung |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation
di: Yoon, Hee Suk, et al.
Pubblicazione: (2024)
di: Yoon, Hee Suk, et al.
Pubblicazione: (2024)
SEER: Semantic Enhancement and Emotional Reasoning Network for Multimodal Fake News Detection
di: Zhu, Peican, et al.
Pubblicazione: (2025)
di: Zhu, Peican, et al.
Pubblicazione: (2025)
KEN: Knowledge Augmentation and Emotion Guidance Network for Multimodal Fake News Detection
di: Zhu, Peican, et al.
Pubblicazione: (2025)
di: Zhu, Peican, et al.
Pubblicazione: (2025)
LMVD: A Large-Scale Multimodal Vlog Dataset for Depression Detection in the Wild
di: He, Lang, et al.
Pubblicazione: (2024)
di: He, Lang, et al.
Pubblicazione: (2024)
PSA-MF: Personality-Sentiment Aligned Multi-Level Fusion for Multimodal Sentiment Analysis
di: Xie, Heng, et al.
Pubblicazione: (2025)
di: Xie, Heng, et al.
Pubblicazione: (2025)
Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality
di: Park, Kyu Ri, et al.
Pubblicazione: (2024)
di: Park, Kyu Ri, et al.
Pubblicazione: (2024)
Differential Multimodal Transformers
di: Li, Jerry, et al.
Pubblicazione: (2025)
di: Li, Jerry, et al.
Pubblicazione: (2025)
SynthGuard: An Open Platform for Detecting AI-Generated Multimedia with Multimodal LLMs
di: Desai, Shail, et al.
Pubblicazione: (2025)
di: Desai, Shail, et al.
Pubblicazione: (2025)
Interpretable Multimodal Misinformation Detection with Logic Reasoning
di: Liu, Hui, et al.
Pubblicazione: (2023)
di: Liu, Hui, et al.
Pubblicazione: (2023)
Learning Co-Speech Gesture for Multimodal Aphasia Type Detection
di: Lee, Daeun, et al.
Pubblicazione: (2023)
di: Lee, Daeun, et al.
Pubblicazione: (2023)
BDIQA: A New Dataset for Video Question Answering to Explore Cognitive Reasoning through Theory of Mind
di: Mao, Yuanyuan, et al.
Pubblicazione: (2024)
di: Mao, Yuanyuan, et al.
Pubblicazione: (2024)
Modeling Human Responses to Multimodal AI Content
di: Shen, Zhiqi, et al.
Pubblicazione: (2025)
di: Shen, Zhiqi, et al.
Pubblicazione: (2025)
Tri-Subspaces Disentanglement for Multimodal Sentiment Analysis
di: Meng, Chunlei, et al.
Pubblicazione: (2026)
di: Meng, Chunlei, et al.
Pubblicazione: (2026)
CPCLDETECTOR: Knowledge Enhancement and Alignment Selection for Chinese Patronizing and Condescending Language Detection
di: Yang, Jiaxun, et al.
Pubblicazione: (2025)
di: Yang, Jiaxun, et al.
Pubblicazione: (2025)
A Multimodal Single-Branch Embedding Network for Recommendation in Cold-Start and Missing Modality Scenarios
di: Ganhör, Christian, et al.
Pubblicazione: (2024)
di: Ganhör, Christian, et al.
Pubblicazione: (2024)
WorldGPT: Empowering LLM as Multimodal World Model
di: Ge, Zhiqi, et al.
Pubblicazione: (2024)
di: Ge, Zhiqi, et al.
Pubblicazione: (2024)
Synthesizing Sentiment-Controlled Feedback For Multimodal Text and Image Data
di: Kumar, Puneet, et al.
Pubblicazione: (2024)
di: Kumar, Puneet, et al.
Pubblicazione: (2024)
A Survey on Multimodal Benchmarks: In the Era of Large AI Models
di: Li, Lin, et al.
Pubblicazione: (2024)
di: Li, Lin, et al.
Pubblicazione: (2024)
Enhancing Multimodal Retrieval via Complementary Information Extraction and Alignment
di: Zeng, Delong, et al.
Pubblicazione: (2026)
di: Zeng, Delong, et al.
Pubblicazione: (2026)
AMB-DSGDN: Adaptive Modality-Balanced Dynamic Semantic Graph Differential Network for Multimodal Emotion Recognition
di: Wang, Yunsheng, et al.
Pubblicazione: (2026)
di: Wang, Yunsheng, et al.
Pubblicazione: (2026)
MMVA: Multimodal Matching Based on Valence and Arousal across Images, Music, and Musical Captions
di: Choi, Suhwan, et al.
Pubblicazione: (2025)
di: Choi, Suhwan, et al.
Pubblicazione: (2025)
Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning
di: Cheng, Zebang, et al.
Pubblicazione: (2024)
di: Cheng, Zebang, et al.
Pubblicazione: (2024)
Multimodal Emotion Recognition by Fusing Video Semantic in MOOC Learning Scenarios
di: Zhang, Yuan, et al.
Pubblicazione: (2024)
di: Zhang, Yuan, et al.
Pubblicazione: (2024)
RW-Post: Auditable Evidence-Grounded Multimodal Fact-Checking in the Wild
di: Xu, Danni, et al.
Pubblicazione: (2026)
di: Xu, Danni, et al.
Pubblicazione: (2026)
'No' Matters: Out-of-Distribution Detection in Multimodality Long Dialogue
di: Gao, Rena, et al.
Pubblicazione: (2024)
di: Gao, Rena, et al.
Pubblicazione: (2024)
Has Multimodal Learning Delivered Universal Intelligence in Healthcare? A Comprehensive Survey
di: Lin, Qika, et al.
Pubblicazione: (2024)
di: Lin, Qika, et al.
Pubblicazione: (2024)
QMAVIS: Long Video-Audio Understanding using Fusion of Large Multimodal Models
di: Lin, Zixing, et al.
Pubblicazione: (2026)
di: Lin, Zixing, et al.
Pubblicazione: (2026)
RA-BLIP: Multimodal Adaptive Retrieval-Augmented Bootstrapping Language-Image Pre-training
di: Ding, Muhe, et al.
Pubblicazione: (2024)
di: Ding, Muhe, et al.
Pubblicazione: (2024)
Why We Feel: Breaking Boundaries in Emotional Reasoning with Multimodal Large Language Models
di: Lin, Yuxiang, et al.
Pubblicazione: (2025)
di: Lin, Yuxiang, et al.
Pubblicazione: (2025)
Memory-Centric Embodied Question Answering
di: Zhai, Mingliang, et al.
Pubblicazione: (2025)
di: Zhai, Mingliang, et al.
Pubblicazione: (2025)
A Survey of Generative Categories and Techniques in Multimodal Generative Models
di: Han, Longzhen, et al.
Pubblicazione: (2025)
di: Han, Longzhen, et al.
Pubblicazione: (2025)
MaLoRA: Gated Modality LoRA for Key-Space Alignment in Multimodal LLM Fine-Tuning
di: Zheng, Xinhan, et al.
Pubblicazione: (2025)
di: Zheng, Xinhan, et al.
Pubblicazione: (2025)
LiveK12Bench: Have Large Multimodal Models Truly Conquered High School-level Examinations?
di: Wang, Xiaohan, et al.
Pubblicazione: (2026)
di: Wang, Xiaohan, et al.
Pubblicazione: (2026)
The Dream Within Huang Long Cave: AI-Driven Interactive Narrative for Family Storytelling and Emotional Reflection
di: Huang, Jiayang, et al.
Pubblicazione: (2025)
di: Huang, Jiayang, et al.
Pubblicazione: (2025)
The Synthetic Media Shift: Tracking the Rise, Virality, and Detectability of AI-Generated Multimodal Misinformation
di: Chrysidis, Zacharias, et al.
Pubblicazione: (2026)
di: Chrysidis, Zacharias, et al.
Pubblicazione: (2026)
MM-HSD: Multi-Modal Hate Speech Detection in Videos
di: Céspedes-Sarrias, Berta, et al.
Pubblicazione: (2025)
di: Céspedes-Sarrias, Berta, et al.
Pubblicazione: (2025)
MMTB: Evaluating Terminal Agents on Multimedia-File Tasks
di: Heo, Chiyeong, et al.
Pubblicazione: (2026)
di: Heo, Chiyeong, et al.
Pubblicazione: (2026)
Exposing Cross-Modal Consistency for Fake News Detection in Short-Form Videos
di: Tian, Chong, et al.
Pubblicazione: (2026)
di: Tian, Chong, et al.
Pubblicazione: (2026)
Unsupervised Multimodal Clustering for Semantics Discovery in Multimodal Utterances
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
A Multi-modal Fusion Network for Terrain Perception Based on Illumination Aware
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
Documenti analoghi
-
BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation
di: Yoon, Hee Suk, et al.
Pubblicazione: (2024) -
SEER: Semantic Enhancement and Emotional Reasoning Network for Multimodal Fake News Detection
di: Zhu, Peican, et al.
Pubblicazione: (2025) -
KEN: Knowledge Augmentation and Emotion Guidance Network for Multimodal Fake News Detection
di: Zhu, Peican, et al.
Pubblicazione: (2025) -
LMVD: A Large-Scale Multimodal Vlog Dataset for Depression Detection in the Wild
di: He, Lang, et al.
Pubblicazione: (2024) -
PSA-MF: Personality-Sentiment Aligned Multi-Level Fusion for Multimodal Sentiment Analysis
di: Xie, Heng, et al.
Pubblicazione: (2025)