Gespeichert in:
| Hauptverfasser: | Luo, Meng, Li, Bobo, Xu, Shanqing, Zhang, Shize, Chen, Qiuchan, Han, Menglu, Chen, Wenhao, Huang, Yanxiang, Fei, Hao, Lee, Mong-Li, Hsu, Wynne |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.00971 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FormFactory: An Interactive Benchmarking Suite for Multimodal Form-Filling Agents
von: Li, Bobo, et al.
Veröffentlicht: (2025)
von: Li, Bobo, et al.
Veröffentlicht: (2025)
PanoSent: A Panoptic Sextuple Extraction Benchmark for Multimodal Conversational Aspect-based Sentiment Analysis
von: Luo, Meng, et al.
Veröffentlicht: (2024)
von: Luo, Meng, et al.
Veröffentlicht: (2024)
Orthogonal Spatial-temporal Distributional Transfer for 4D Generation
von: Liu, Wei, et al.
Veröffentlicht: (2026)
von: Liu, Wei, et al.
Veröffentlicht: (2026)
Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
von: Fei, Hao, et al.
Veröffentlicht: (2024)
von: Fei, Hao, et al.
Veröffentlicht: (2024)
Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment
von: Li, Bobo, et al.
Veröffentlicht: (2026)
von: Li, Bobo, et al.
Veröffentlicht: (2026)
Faithful Logical Reasoning via Symbolic Chain-of-Thought
von: Xu, Jundong, et al.
Veröffentlicht: (2024)
von: Xu, Jundong, et al.
Veröffentlicht: (2024)
TRUST-VL: An Explainable News Assistant for General Multimodal Misinformation Detection
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
SNIFFER: Multimodal Large Language Model for Explainable Out-of-Context Misinformation Detection
von: Qi, Peng, et al.
Veröffentlicht: (2024)
von: Qi, Peng, et al.
Veröffentlicht: (2024)
Mitigating GenAI-powered Evidence Pollution for Out-of-Context Multimodal Misinformation Detection
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
von: Yan, Zehong, et al.
Veröffentlicht: (2025)
LEAF-Mamba: Local Emphatic and Adaptive Fusion State Space Model for RGB-D Salient Object Detection
von: Wu, Lanhu, et al.
Veröffentlicht: (2025)
von: Wu, Lanhu, et al.
Veröffentlicht: (2025)
Multi-Part Object Representations via Graph Structures and Co-Part Discovery
von: Foo, Alex, et al.
Veröffentlicht: (2025)
von: Foo, Alex, et al.
Veröffentlicht: (2025)
Multi-Modal Continual Learning via Cross-Modality Adapters and Representation Alignment with Knowledge Preservation
von: Chee, Evelyn, et al.
Veröffentlicht: (2025)
von: Chee, Evelyn, et al.
Veröffentlicht: (2025)
From Personas to Talks: Revisiting the Impact of Personas on LLM-Synthesized Emotional Support Conversations
von: Wu, Shenghan, et al.
Veröffentlicht: (2025)
von: Wu, Shenghan, et al.
Veröffentlicht: (2025)
Aristotle: Mastering Logical Reasoning with A Logic-Complete Decompose-Search-Resolve Framework
von: Xu, Jundong, et al.
Veröffentlicht: (2024)
von: Xu, Jundong, et al.
Veröffentlicht: (2024)
NUS-Emo at SemEval-2024 Task 3: Instruction-Tuning LLM for Multimodal Emotion-Cause Analysis in Conversations
von: Luo, Meng, et al.
Veröffentlicht: (2024)
von: Luo, Meng, et al.
Veröffentlicht: (2024)
Evidence-Based Temporal Fact Verification
von: Barik, Anab Maulana, et al.
Veröffentlicht: (2024)
von: Barik, Anab Maulana, et al.
Veröffentlicht: (2024)
ChronoFact: Timeline-based Temporal Fact Verification
von: Barik, Anab Maulana, et al.
Veröffentlicht: (2024)
von: Barik, Anab Maulana, et al.
Veröffentlicht: (2024)
MuSLR: Multimodal Symbolic Logical Reasoning
von: Xu, Jundong, et al.
Veröffentlicht: (2025)
von: Xu, Jundong, et al.
Veröffentlicht: (2025)
Test-Time Adaptation by Causal Trimming
von: Liu, Yingnan, et al.
Veröffentlicht: (2025)
von: Liu, Yingnan, et al.
Veröffentlicht: (2025)
UniM: A Unified Any-to-Any Interleaved Multimodal Benchmark
von: Li, Yanlin, et al.
Veröffentlicht: (2026)
von: Li, Yanlin, et al.
Veröffentlicht: (2026)
LogicReward: Incentivizing LLM Reasoning via Step-Wise Logical Supervision
von: Xu, Jundong, et al.
Veröffentlicht: (2025)
von: Xu, Jundong, et al.
Veröffentlicht: (2025)
Dr.V: A Hierarchical Perception-Temporal-Cognition Framework to Diagnose Video Hallucination by Fine-grained Spatial-Temporal Grounding
von: Luo, Meng, et al.
Veröffentlicht: (2025)
von: Luo, Meng, et al.
Veröffentlicht: (2025)
On the Adaptive Psychological Persuasion of Large Language Models
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
Althea: Human-AI Collaboration for Fact-Checking and Critical Reasoning
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
von: Churina, Svetlana, et al.
Veröffentlicht: (2025)
The Effects of Mindfulness on Shame: Exploring Mediation by Cognitive Flexibility and Self‐Compassion in a Chinese Adult Population
von: Xiaoshuo Zhang, et al.
Veröffentlicht: (2024)
von: Xiaoshuo Zhang, et al.
Veröffentlicht: (2024)
Probing then Editing Response Personality of Large Language Models
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
Towards Robust Out-of-Distribution Generalization Bounds via Sharpness
von: Zou, Yingtian, et al.
Veröffentlicht: (2024)
von: Zou, Yingtian, et al.
Veröffentlicht: (2024)
Cross-Domain Feature Augmentation for Domain Generalization
von: Liu, Yingnan, et al.
Veröffentlicht: (2024)
von: Liu, Yingnan, et al.
Veröffentlicht: (2024)
MultiMind: Enhancing Werewolf Agents with Multimodal Reasoning and Theory of Mind
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
MindMerger: Efficient Boosting LLM Reasoning in non-English Languages
von: Huang, Zixian, et al.
Veröffentlicht: (2024)
von: Huang, Zixian, et al.
Veröffentlicht: (2024)
Watch Out Your Album! On the Inadvertent Privacy Memorization in Multi-Modal Large Language Models
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
QIME: Constructing Interpretable Medical Text Embeddings via Ontology-Grounded Questions
von: Tang, Yixuan, et al.
Veröffentlicht: (2026)
von: Tang, Yixuan, et al.
Veröffentlicht: (2026)
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models
von: Qian, Zhe, et al.
Veröffentlicht: (2026)
von: Qian, Zhe, et al.
Veröffentlicht: (2026)
Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2025)
von: Yeo, Gerard Christopher, et al.
Veröffentlicht: (2025)
COKE: A Cognitive Knowledge Graph for Machine Theory of Mind
von: Wu, Jincenzi, et al.
Veröffentlicht: (2023)
von: Wu, Jincenzi, et al.
Veröffentlicht: (2023)
Dynamic Emotion and Personality Profiling for Multimodal Deception Detection
von: Zheng, Li, et al.
Veröffentlicht: (2026)
von: Zheng, Li, et al.
Veröffentlicht: (2026)
Enhancing Thermal Conductive Properties With Liquid Metal‐Assisted Epoxy Resin Composites
von: Junyan Wang, et al.
Veröffentlicht: (2025)
von: Junyan Wang, et al.
Veröffentlicht: (2025)
When Disagreements Elicit Robustness: Investigating Self-Repair Capabilities under LLM Multi-Agent Disagreements
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
M$^{3}$D: A Multimodal, Multilingual and Multitask Dataset for Grounded Document-level Information Extraction
von: Liu, Jiang, et al.
Veröffentlicht: (2024)
von: Liu, Jiang, et al.
Veröffentlicht: (2024)
Through the Theory of Mind's Eye: Reading Minds with Multimodal Video Large Language Models
von: Chen, Zhawnen, et al.
Veröffentlicht: (2024)
von: Chen, Zhawnen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FormFactory: An Interactive Benchmarking Suite for Multimodal Form-Filling Agents
von: Li, Bobo, et al.
Veröffentlicht: (2025) -
PanoSent: A Panoptic Sextuple Extraction Benchmark for Multimodal Conversational Aspect-based Sentiment Analysis
von: Luo, Meng, et al.
Veröffentlicht: (2024) -
Orthogonal Spatial-temporal Distributional Transfer for 4D Generation
von: Liu, Wei, et al.
Veröffentlicht: (2026) -
Video-of-Thought: Step-by-Step Video Reasoning from Perception to Cognition
von: Fei, Hao, et al.
Veröffentlicht: (2024) -
Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment
von: Li, Bobo, et al.
Veröffentlicht: (2026)