'No' Matters: Out-of-Distribution Detection in Multimodality Long Dialogue
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Rena, Wu, Xuetong, Luo, Siwen, Han, Caren, Liu, Feng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multimodal Multi-loss Fusion Network for Sentiment Analysis
von: Wu, Zehui, et al.
Veröffentlicht: (2023)
von: Wu, Zehui, et al.
Veröffentlicht: (2023)
DLF: Disentangled-Language-Focused Multimodal Sentiment Analysis
von: Wang, Pan, et al.
Veröffentlicht: (2024)
von: Wang, Pan, et al.
Veröffentlicht: (2024)
Sample then Identify: A General Framework for Risk Control and Assessment in Multimodal Large Language Models
von: Wang, Qingni, et al.
Veröffentlicht: (2024)
von: Wang, Qingni, et al.
Veröffentlicht: (2024)
Doctor Sun: A Bilingual Multimodal Large Language Model for Biomedical AI
von: Xue, Dong, et al.
Veröffentlicht: (2025)
von: Xue, Dong, et al.
Veröffentlicht: (2025)
Multimodal Long Video Modeling Based on Temporal Dynamic Context
von: Hao, Haoran, et al.
Veröffentlicht: (2025)
von: Hao, Haoran, et al.
Veröffentlicht: (2025)
Continual Multimodal Knowledge Graph Construction
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
von: Chen, Xiang, et al.
Veröffentlicht: (2023)
Decoding the Hook: A Multimodal LLM Framework for Analyzing the Hooking Period of Video Ads
von: Zhang, Kunpeng, et al.
Veröffentlicht: (2026)
von: Zhang, Kunpeng, et al.
Veröffentlicht: (2026)
MuPHI: Learning Implicit Multimodal Harm Reasoning via Semantically Grounded Reward Optimization
von: Saha, Anisha, et al.
Veröffentlicht: (2026)
von: Saha, Anisha, et al.
Veröffentlicht: (2026)
Unified Hallucination Detection for Multimodal Large Language Models
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
von: Chen, Xiang, et al.
Veröffentlicht: (2024)
RAP: Retrieval-Augmented Personalization for Multimodal Large Language Models
von: Hao, Haoran, et al.
Veröffentlicht: (2024)
von: Hao, Haoran, et al.
Veröffentlicht: (2024)
Automating Steering for Safe Multimodal Large Language Models
von: Wu, Lyucheng, et al.
Veröffentlicht: (2025)
von: Wu, Lyucheng, et al.
Veröffentlicht: (2025)
A Review of Multimodal Explainable Artificial Intelligence: Past, Present and Future
von: Sun, Shilin, et al.
Veröffentlicht: (2024)
von: Sun, Shilin, et al.
Veröffentlicht: (2024)
BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2024)
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2024)
Unleashing the Power of Imbalanced Modality Information for Multi-modal Knowledge Graph Completion
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
von: Zhang, Yichi, et al.
Veröffentlicht: (2024)
MultiScript30k: Leveraging Multilingual Embeddings to Extend Cross Script Parallel Data
von: Driggers-Ellis, Christopher, et al.
Veröffentlicht: (2025)
von: Driggers-Ellis, Christopher, et al.
Veröffentlicht: (2025)
MultiMedEdit: A Scenario-Aware Benchmark for Evaluating Knowledge Editing in Medical VQA
von: Wen, Shengtao, et al.
Veröffentlicht: (2025)
von: Wen, Shengtao, et al.
Veröffentlicht: (2025)
Reasoning LLMs are Wandering Solution Explorers
von: Lu, Jiahao, et al.
Veröffentlicht: (2025)
von: Lu, Jiahao, et al.
Veröffentlicht: (2025)
ShizhenGPT: Towards Multimodal LLMs for Traditional Chinese Medicine
von: Chen, Junying, et al.
Veröffentlicht: (2025)
von: Chen, Junying, et al.
Veröffentlicht: (2025)
Let Multimodal Embedders Learn When to Augment Query via Adaptive Query Augmentation
von: Kim, Wongyu, et al.
Veröffentlicht: (2025)
von: Kim, Wongyu, et al.
Veröffentlicht: (2025)
Interpretable Multimodal Misinformation Detection with Logic Reasoning
von: Liu, Hui, et al.
Veröffentlicht: (2023)
von: Liu, Hui, et al.
Veröffentlicht: (2023)
Can We Edit Multimodal Large Language Models?
von: Cheng, Siyuan, et al.
Veröffentlicht: (2023)
von: Cheng, Siyuan, et al.
Veröffentlicht: (2023)
NVLM: Open Frontier-Class Multimodal LLMs
von: Dai, Wenliang, et al.
Veröffentlicht: (2024)
von: Dai, Wenliang, et al.
Veröffentlicht: (2024)
OceanPile: A Large-Scale Multimodal Ocean Corpus for Foundation Models
von: Xue, Yida, et al.
Veröffentlicht: (2026)
von: Xue, Yida, et al.
Veröffentlicht: (2026)
SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
Spatial Knowledge Graph-Guided Multimodal Synthesis
von: Xue, Yida, et al.
Veröffentlicht: (2025)
von: Xue, Yida, et al.
Veröffentlicht: (2025)
Unsupervised Multimodal Clustering for Semantics Discovery in Multimodal Utterances
von: Zhang, Hanlei, et al.
Veröffentlicht: (2024)
von: Zhang, Hanlei, et al.
Veröffentlicht: (2024)
MultiMed: Massively Multimodal and Multitask Medical Understanding
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
Multi-level Mixture of Experts for Multimodal Entity Linking
von: Hu, Zhiwei, et al.
Veröffentlicht: (2025)
von: Hu, Zhiwei, et al.
Veröffentlicht: (2025)
SIN-Bench: Tracing Native Evidence Chains in Long-Context Multimodal Scientific Interleaved Literature
von: Ren, Yiming, et al.
Veröffentlicht: (2026)
von: Ren, Yiming, et al.
Veröffentlicht: (2026)
Automatic Time Signature Determination for New Scores Using Lyrics for Latent Rhythmic Structure
von: Liao, Callie C., et al.
Veröffentlicht: (2023)
von: Liao, Callie C., et al.
Veröffentlicht: (2023)
LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model
von: Sun, Yirong, et al.
Veröffentlicht: (2025)
von: Sun, Yirong, et al.
Veröffentlicht: (2025)
Music for All: Representational Bias and Cross-Cultural Adaptability of Music Generation Models
von: Mehta, Atharva, et al.
Veröffentlicht: (2025)
von: Mehta, Atharva, et al.
Veröffentlicht: (2025)
A Multimodal Memes Classification: A Survey and Open Research Issues
von: Afridi, Tariq Habib, et al.
Veröffentlicht: (2020)
von: Afridi, Tariq Habib, et al.
Veröffentlicht: (2020)
LightThinker: Thinking Step-by-Step Compression
von: Zhang, Jintian, et al.
Veröffentlicht: (2025)
von: Zhang, Jintian, et al.
Veröffentlicht: (2025)
Unveiling Covert Toxicity in Multimodal Data via Toxicity Association Graphs: A Graph-Based Metric and Interpretable Detection Framework
von: Wu, Guanzong, et al.
Veröffentlicht: (2026)
von: Wu, Guanzong, et al.
Veröffentlicht: (2026)
LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
von: Zhang, Renrui, et al.
Veröffentlicht: (2023)
von: Zhang, Renrui, et al.
Veröffentlicht: (2023)
LightThinker++: From Reasoning Compression to Memory Management
von: Zhu, Yuqi, et al.
Veröffentlicht: (2026)
von: Zhu, Yuqi, et al.
Veröffentlicht: (2026)
TIGeR: Unifying Text-to-Image Generation and Retrieval with Large Multimodal Models
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
von: Qu, Leigang, et al.
Veröffentlicht: (2024)
ADS-Edit: A Multimodal Knowledge Editing Dataset for Autonomous Driving Systems
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes?
von: Yao, Yang, et al.
Veröffentlicht: (2025)
von: Yao, Yang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multimodal Multi-loss Fusion Network for Sentiment Analysis
von: Wu, Zehui, et al.
Veröffentlicht: (2023) -
DLF: Disentangled-Language-Focused Multimodal Sentiment Analysis
von: Wang, Pan, et al.
Veröffentlicht: (2024) -
Sample then Identify: A General Framework for Risk Control and Assessment in Multimodal Large Language Models
von: Wang, Qingni, et al.
Veröffentlicht: (2024) -
Doctor Sun: A Bilingual Multimodal Large Language Model for Biomedical AI
von: Xue, Dong, et al.
Veröffentlicht: (2025) -
Multimodal Long Video Modeling Based on Temporal Dynamic Context
von: Hao, Haoran, et al.
Veröffentlicht: (2025)