Breaking the Illusion: When Positive Meets Negative in Multimodal Decoding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Yubo, An, Yitong, Yang, Xin, Wuerkaixi, Abudukelimu, Cheng, Xuxin, Xie, Fengying, Jiang, Zhiguo, Liu, Cao, Zeng, Ke, Zhang, Haopeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization
von: Jiang, Yubo, et al.
Veröffentlicht: (2026)
von: Jiang, Yubo, et al.
Veröffentlicht: (2026)
Global Context or Local Detail? Adaptive Visual Grounding for Hallucination Mitigation
von: Jiang, Yubo, et al.
Veröffentlicht: (2026)
von: Jiang, Yubo, et al.
Veröffentlicht: (2026)
Towards Self-Robust LLMs: Intrinsic Prompt Noise Resistance via CoIPO
von: Yang, Xin, et al.
Veröffentlicht: (2026)
von: Yang, Xin, et al.
Veröffentlicht: (2026)
AutothinkRAG: Complexity-Aware Control of Retrieval-Augmented Reasoning for Image-Text Interaction
von: Yang, Jiashu, et al.
Veröffentlicht: (2026)
von: Yang, Jiashu, et al.
Veröffentlicht: (2026)
Pan-cancer Histopathology WSI Pre-training with Position-aware Masked Autoencoder
von: Wu, Kun, et al.
Veröffentlicht: (2024)
von: Wu, Kun, et al.
Veröffentlicht: (2024)
Beyond Similarity: Personalized Federated Recommendation with Composite Aggregation
von: Zhang, Honglei, et al.
Veröffentlicht: (2024)
von: Zhang, Honglei, et al.
Veröffentlicht: (2024)
Decentralized Dynamic Cooperation of Personalized Models for Federated Continual Learning
von: Yang, Danni, et al.
Veröffentlicht: (2025)
von: Yang, Danni, et al.
Veröffentlicht: (2025)
Balancing Similarity and Complementarity for Federated Learning
von: Yan, Kunda, et al.
Veröffentlicht: (2024)
von: Yan, Kunda, et al.
Veröffentlicht: (2024)
When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
Speculative Decoding: Performance or Illusion?
von: Liu, Xiaoxuan, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoxuan, et al.
Veröffentlicht: (2025)
Is New Technology Worth It? Exploring Positive and Negative Effects When New Technology Meets Consumers
von: Philipp Brüggemann, et al.
Veröffentlicht: (2025)
von: Philipp Brüggemann, et al.
Veröffentlicht: (2025)
Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems
von: Zhang, Yuzhe, et al.
Veröffentlicht: (2026)
von: Zhang, Yuzhe, et al.
Veröffentlicht: (2026)
Break the Visual Perception: Adversarial Attacks Targeting Encoded Visual Tokens of Large Vision-Language Models
von: Wang, Yubo, et al.
Veröffentlicht: (2024)
von: Wang, Yubo, et al.
Veröffentlicht: (2024)
Accurate Forgetting for Heterogeneous Federated Continual Learning
von: Wuerkaixi, Abudukelimu, et al.
Veröffentlicht: (2025)
von: Wuerkaixi, Abudukelimu, et al.
Veröffentlicht: (2025)
Breaking the Illusion of Identity in LLM Tooling
von: Miller, Marek
Veröffentlicht: (2026)
von: Miller, Marek
Veröffentlicht: (2026)
When Drafts Evolve: Speculative Decoding Meets Online Learning
von: Qian, Yu-Yang, et al.
Veröffentlicht: (2026)
von: Qian, Yu-Yang, et al.
Veröffentlicht: (2026)
When Graph Neural Network Meets Causality: Opportunities, Methodologies and An Outlook
von: Jiang, Wenzhao, et al.
Veröffentlicht: (2023)
von: Jiang, Wenzhao, et al.
Veröffentlicht: (2023)
Breaking the Illusion: Consensus-Based Generative Mitigation of Adversarial Illusions in Multi-Modal Embeddings
von: Akbarian, Fatemeh, et al.
Veröffentlicht: (2025)
von: Akbarian, Fatemeh, et al.
Veröffentlicht: (2025)
Where the Illusion of Dark Energy Must Break
von: Eric D. Mize
Veröffentlicht: (2026)
von: Eric D. Mize
Veröffentlicht: (2026)
Autoregressive Image Generation with Randomized Parallel Decoding
von: Li, Haopeng, et al.
Veröffentlicht: (2025)
von: Li, Haopeng, et al.
Veröffentlicht: (2025)
DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding
von: Zhou, Jiaming, et al.
Veröffentlicht: (2026)
von: Zhou, Jiaming, et al.
Veröffentlicht: (2026)
FADE: A Task-Agnostic Upsampling Operator for Encoder-Decoder Architectures
von: Lu, Hao, et al.
Veröffentlicht: (2024)
von: Lu, Hao, et al.
Veröffentlicht: (2024)
Prototypical Information Bottlenecking and Disentangling for Multimodal Cancer Survival Prediction
von: Zhang, Yilan, et al.
Veröffentlicht: (2024)
von: Zhang, Yilan, et al.
Veröffentlicht: (2024)
Learning without Isolation: Pathway Protection for Continual Learning
von: Chen, Zhikang, et al.
Veröffentlicht: (2025)
von: Chen, Zhikang, et al.
Veröffentlicht: (2025)
Reflecting Twice before Speaking with Empathy: Self-Reflective Alternating Inference for Empathy-Aware End-to-End Spoken Dialogue
von: Jia, Yuhang, et al.
Veröffentlicht: (2026)
von: Jia, Yuhang, et al.
Veröffentlicht: (2026)
Invariant Representation Guided Multimodal Sentiment Decoding with Sequential Variation Regularization
von: Xu, Guoyang, et al.
Veröffentlicht: (2024)
von: Xu, Guoyang, et al.
Veröffentlicht: (2024)
Coherent Phonon Negative Refraction via Interfacial Momentum Compensation
von: Chen, Hao, et al.
Veröffentlicht: (2025)
von: Chen, Hao, et al.
Veröffentlicht: (2025)
HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding
von: Shi, Mengqi, et al.
Veröffentlicht: (2026)
von: Shi, Mengqi, et al.
Veröffentlicht: (2026)
When MLLMs Meet Compression Distortion: A Coding Paradigm Tailored to MLLMs
von: Liu, Jinming, et al.
Veröffentlicht: (2025)
von: Liu, Jinming, et al.
Veröffentlicht: (2025)
When Trust Collides: Decoding Human-LLM Cooperation Dynamics through the Prisoner's Dilemma
von: Jiang, Guanxuan, et al.
Veröffentlicht: (2025)
von: Jiang, Guanxuan, et al.
Veröffentlicht: (2025)
When Learning Meets Dynamics: Distributed User Connectivity Maximization in UAV-Based Communication Networks
von: Li, Bowei, et al.
Veröffentlicht: (2024)
von: Li, Bowei, et al.
Veröffentlicht: (2024)
Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
von: Hu, Yueqing, et al.
Veröffentlicht: (2026)
von: Hu, Yueqing, et al.
Veröffentlicht: (2026)
The Pomegranate Flower Water Extract Negatively Regulates Melanogenesis by Suppressing MITF Expression and Its Target Enzymes
von: Peng Shu, et al.
Veröffentlicht: (2025)
von: Peng Shu, et al.
Veröffentlicht: (2025)
When Visual Privacy Protection Meets Multimodal Large Language Models
von: Hui, Xiaofei, et al.
Veröffentlicht: (2026)
von: Hui, Xiaofei, et al.
Veröffentlicht: (2026)
Rethinking the adaptive relationship between Encoder Layers and Decoder Layers
von: Song, Yubo
Veröffentlicht: (2024)
von: Song, Yubo
Veröffentlicht: (2024)
A System-View Optimal Additional Active Power Control of Wind Turbines for Grid Frequency Support
von: Zhang, Yubo, et al.
Veröffentlicht: (2026)
von: Zhang, Yubo, et al.
Veröffentlicht: (2026)
Model-Free Fast Frequency Support of Wind Farms for Tracking Optimal Frequency Trajectory
von: Zhang, Yubo, et al.
Veröffentlicht: (2026)
von: Zhang, Yubo, et al.
Veröffentlicht: (2026)
Multimodal Negative Learning
von: Gong, Baoquan, et al.
Veröffentlicht: (2025)
von: Gong, Baoquan, et al.
Veröffentlicht: (2025)
DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
Audio-text Retrieval with Transformer-based Hierarchical Alignment and Disentangled Cross-modal Representation
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization
von: Jiang, Yubo, et al.
Veröffentlicht: (2026) -
Global Context or Local Detail? Adaptive Visual Grounding for Hallucination Mitigation
von: Jiang, Yubo, et al.
Veröffentlicht: (2026) -
Towards Self-Robust LLMs: Intrinsic Prompt Noise Resistance via CoIPO
von: Yang, Xin, et al.
Veröffentlicht: (2026) -
AutothinkRAG: Complexity-Aware Control of Retrieval-Augmented Reasoning for Image-Text Interaction
von: Yang, Jiashu, et al.
Veröffentlicht: (2026) -
Pan-cancer Histopathology WSI Pre-training with Position-aware Masked Autoencoder
von: Wu, Kun, et al.
Veröffentlicht: (2024)