Breaking the Illusion: When Positive Meets Negative in Multimodal Decoding
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Yubo, An, Yitong, Yang, Xin, Wuerkaixi, Abudukelimu, Cheng, Xuxin, Xie, Fengying, Jiang, Zhiguo, Liu, Cao, Zeng, Ke, Zhang, Haopeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization
by: Jiang, Yubo, et al.
Published: (2026)
by: Jiang, Yubo, et al.
Published: (2026)
Global Context or Local Detail? Adaptive Visual Grounding for Hallucination Mitigation
by: Jiang, Yubo, et al.
Published: (2026)
by: Jiang, Yubo, et al.
Published: (2026)
Towards Self-Robust LLMs: Intrinsic Prompt Noise Resistance via CoIPO
by: Yang, Xin, et al.
Published: (2026)
by: Yang, Xin, et al.
Published: (2026)
AutothinkRAG: Complexity-Aware Control of Retrieval-Augmented Reasoning for Image-Text Interaction
by: Yang, Jiashu, et al.
Published: (2026)
by: Yang, Jiashu, et al.
Published: (2026)
Pan-cancer Histopathology WSI Pre-training with Position-aware Masked Autoencoder
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
Beyond Similarity: Personalized Federated Recommendation with Composite Aggregation
by: Zhang, Honglei, et al.
Published: (2024)
by: Zhang, Honglei, et al.
Published: (2024)
Decentralized Dynamic Cooperation of Personalized Models for Federated Continual Learning
by: Yang, Danni, et al.
Published: (2025)
by: Yang, Danni, et al.
Published: (2025)
Balancing Similarity and Complementarity for Federated Learning
by: Yan, Kunda, et al.
Published: (2024)
by: Yan, Kunda, et al.
Published: (2024)
When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training
by: Wang, Haonan, et al.
Published: (2024)
by: Wang, Haonan, et al.
Published: (2024)
Speculative Decoding: Performance or Illusion?
by: Liu, Xiaoxuan, et al.
Published: (2025)
by: Liu, Xiaoxuan, et al.
Published: (2025)
Is New Technology Worth It? Exploring Positive and Negative Effects When New Technology Meets Consumers
by: Philipp Brüggemann, et al.
Published: (2025)
by: Philipp Brüggemann, et al.
Published: (2025)
Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems
by: Zhang, Yuzhe, et al.
Published: (2026)
by: Zhang, Yuzhe, et al.
Published: (2026)
Break the Visual Perception: Adversarial Attacks Targeting Encoded Visual Tokens of Large Vision-Language Models
by: Wang, Yubo, et al.
Published: (2024)
by: Wang, Yubo, et al.
Published: (2024)
Accurate Forgetting for Heterogeneous Federated Continual Learning
by: Wuerkaixi, Abudukelimu, et al.
Published: (2025)
by: Wuerkaixi, Abudukelimu, et al.
Published: (2025)
Breaking the Illusion of Identity in LLM Tooling
by: Miller, Marek
Published: (2026)
by: Miller, Marek
Published: (2026)
When Drafts Evolve: Speculative Decoding Meets Online Learning
by: Qian, Yu-Yang, et al.
Published: (2026)
by: Qian, Yu-Yang, et al.
Published: (2026)
When Graph Neural Network Meets Causality: Opportunities, Methodologies and An Outlook
by: Jiang, Wenzhao, et al.
Published: (2023)
by: Jiang, Wenzhao, et al.
Published: (2023)
Breaking the Illusion: Consensus-Based Generative Mitigation of Adversarial Illusions in Multi-Modal Embeddings
by: Akbarian, Fatemeh, et al.
Published: (2025)
by: Akbarian, Fatemeh, et al.
Published: (2025)
Where the Illusion of Dark Energy Must Break
by: Eric D. Mize
Published: (2026)
by: Eric D. Mize
Published: (2026)
Autoregressive Image Generation with Randomized Parallel Decoding
by: Li, Haopeng, et al.
Published: (2025)
by: Li, Haopeng, et al.
Published: (2025)
DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding
by: Zhou, Jiaming, et al.
Published: (2026)
by: Zhou, Jiaming, et al.
Published: (2026)
FADE: A Task-Agnostic Upsampling Operator for Encoder-Decoder Architectures
by: Lu, Hao, et al.
Published: (2024)
by: Lu, Hao, et al.
Published: (2024)
Prototypical Information Bottlenecking and Disentangling for Multimodal Cancer Survival Prediction
by: Zhang, Yilan, et al.
Published: (2024)
by: Zhang, Yilan, et al.
Published: (2024)
Learning without Isolation: Pathway Protection for Continual Learning
by: Chen, Zhikang, et al.
Published: (2025)
by: Chen, Zhikang, et al.
Published: (2025)
Reflecting Twice before Speaking with Empathy: Self-Reflective Alternating Inference for Empathy-Aware End-to-End Spoken Dialogue
by: Jia, Yuhang, et al.
Published: (2026)
by: Jia, Yuhang, et al.
Published: (2026)
Invariant Representation Guided Multimodal Sentiment Decoding with Sequential Variation Regularization
by: Xu, Guoyang, et al.
Published: (2024)
by: Xu, Guoyang, et al.
Published: (2024)
Coherent Phonon Negative Refraction via Interfacial Momentum Compensation
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding
by: Shi, Mengqi, et al.
Published: (2026)
by: Shi, Mengqi, et al.
Published: (2026)
When MLLMs Meet Compression Distortion: A Coding Paradigm Tailored to MLLMs
by: Liu, Jinming, et al.
Published: (2025)
by: Liu, Jinming, et al.
Published: (2025)
When Trust Collides: Decoding Human-LLM Cooperation Dynamics through the Prisoner's Dilemma
by: Jiang, Guanxuan, et al.
Published: (2025)
by: Jiang, Guanxuan, et al.
Published: (2025)
When Learning Meets Dynamics: Distributed User Connectivity Maximization in UAV-Based Communication Networks
by: Li, Bowei, et al.
Published: (2024)
by: Li, Bowei, et al.
Published: (2024)
Social Catalysts, Not Moral Agents: The Illusion of Alignment in LLM Societies
by: Hu, Yueqing, et al.
Published: (2026)
by: Hu, Yueqing, et al.
Published: (2026)
The Pomegranate Flower Water Extract Negatively Regulates Melanogenesis by Suppressing MITF Expression and Its Target Enzymes
by: Peng Shu, et al.
Published: (2025)
by: Peng Shu, et al.
Published: (2025)
When Visual Privacy Protection Meets Multimodal Large Language Models
by: Hui, Xiaofei, et al.
Published: (2026)
by: Hui, Xiaofei, et al.
Published: (2026)
Rethinking the adaptive relationship between Encoder Layers and Decoder Layers
by: Song, Yubo
Published: (2024)
by: Song, Yubo
Published: (2024)
A System-View Optimal Additional Active Power Control of Wind Turbines for Grid Frequency Support
by: Zhang, Yubo, et al.
Published: (2026)
by: Zhang, Yubo, et al.
Published: (2026)
Model-Free Fast Frequency Support of Wind Farms for Tracking Optimal Frequency Trajectory
by: Zhang, Yubo, et al.
Published: (2026)
by: Zhang, Yubo, et al.
Published: (2026)
Multimodal Negative Learning
by: Gong, Baoquan, et al.
Published: (2025)
by: Gong, Baoquan, et al.
Published: (2025)
DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval
by: Xin, Yifei, et al.
Published: (2024)
by: Xin, Yifei, et al.
Published: (2024)
Audio-text Retrieval with Transformer-based Hierarchical Alignment and Disentangled Cross-modal Representation
by: Xin, Yifei, et al.
Published: (2024)
by: Xin, Yifei, et al.
Published: (2024)
Similar Items
-
V-tableR1: Process-Supervised Multimodal Table Reasoning with Critic-Guided Policy Optimization
by: Jiang, Yubo, et al.
Published: (2026) -
Global Context or Local Detail? Adaptive Visual Grounding for Hallucination Mitigation
by: Jiang, Yubo, et al.
Published: (2026) -
Towards Self-Robust LLMs: Intrinsic Prompt Noise Resistance via CoIPO
by: Yang, Xin, et al.
Published: (2026) -
AutothinkRAG: Complexity-Aware Control of Retrieval-Augmented Reasoning for Image-Text Interaction
by: Yang, Jiashu, et al.
Published: (2026) -
Pan-cancer Histopathology WSI Pre-training with Position-aware Masked Autoencoder
by: Wu, Kun, et al.
Published: (2024)