Breaking the Illusion: Consensus-Based Generative Mitigation of Adversarial Illusions in Multi-Modal Embeddings
Fuente:
arXiv
Saved in:
| Main Authors: | Akbarian, Fatemeh, Baninajjar, Anahita, Zhang, Yingyi, Balashankar, Ananth, Aminifar, Amir |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Formal Local Implication Between Two Neural Networks
by: Baninajjar, Anahita, et al.
Published: (2024)
by: Baninajjar, Anahita, et al.
Published: (2024)
VNN: Verification-Friendly Neural Networks with Hard Robustness Guarantees
by: Baninajjar, Anahita, et al.
Published: (2023)
by: Baninajjar, Anahita, et al.
Published: (2023)
BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes
by: Huang, Baichuan, et al.
Published: (2025)
by: Huang, Baichuan, et al.
Published: (2025)
Adversarial Illusions in Multi-Modal Embeddings
by: Zhang, Tingwei, et al.
Published: (2023)
by: Zhang, Tingwei, et al.
Published: (2023)
Breaking the Illusion: Real-world Challenges for Adversarial Patches in Object Detection
by: Shack, Jakob, et al.
Published: (2024)
by: Shack, Jakob, et al.
Published: (2024)
Break it, Imitate it, Fix it: Robustness by Generating Human-Like Attacks
by: Sinha, Aradhana, et al.
Published: (2023)
by: Sinha, Aradhana, et al.
Published: (2023)
Breaking the Illusion: When Positive Meets Negative in Multimodal Decoding
by: Jiang, Yubo, et al.
Published: (2026)
by: Jiang, Yubo, et al.
Published: (2026)
Privacy-Preserving Edge Federated Learning for Intelligent Mobile-Health Systems
by: Aminifar, Amin, et al.
Published: (2024)
by: Aminifar, Amin, et al.
Published: (2024)
FieldFormer: Locality-Aware Transformers for Spatio-Temporal Modeling on Sparse Sensor Networks
by: Bhardwaj, Ankit, et al.
Published: (2025)
by: Bhardwaj, Ankit, et al.
Published: (2025)
LightFF: Lightweight Inference for Forward-Forward Algorithm
by: Aminifar, Amin, et al.
Published: (2024)
by: Aminifar, Amin, et al.
Published: (2024)
The Illusion of Generalization in Tabular Language Models
by: Gorla, Aditya, et al.
Published: (2026)
by: Gorla, Aditya, et al.
Published: (2026)
R-Diverse: Mitigating Diversity Illusion in Self-Play LLM Training
by: Li, Gengsheng, et al.
Published: (2026)
by: Li, Gengsheng, et al.
Published: (2026)
Systematic Scaling Analysis of Jailbreak Attacks in Large Language Models
by: Wang, Xiangwen, et al.
Published: (2026)
by: Wang, Xiangwen, et al.
Published: (2026)
Interpretability Illusions in the Generalization of Simplified Models
by: Friedman, Dan, et al.
Published: (2023)
by: Friedman, Dan, et al.
Published: (2023)
The Illusion of Stochasticity in LLMs
by: Gu, Xiangming, et al.
Published: (2026)
by: Gu, Xiangming, et al.
Published: (2026)
Are Flat Minima an Illusion?
by: Bennett, Michael Timothy
Published: (2026)
by: Bennett, Michael Timothy
Published: (2026)
The Leaderboard Illusion
by: Singh, Shivalika, et al.
Published: (2025)
by: Singh, Shivalika, et al.
Published: (2025)
The Semantic Illusion: Certified Limits of Embedding-Based Hallucination Detection in RAG Systems
by: Sinha, Debu
Published: (2025)
by: Sinha, Debu
Published: (2025)
The Illusion of Readiness in Health AI
by: Gu, Yu, et al.
Published: (2025)
by: Gu, Yu, et al.
Published: (2025)
Adversarial Reinforcement Learning for Large Language Model Agent Safety
by: Wang, Zizhao, et al.
Published: (2025)
by: Wang, Zizhao, et al.
Published: (2025)
The Illusion of Certainty: Decoupling Capability and Calibration in On-Policy Distillation
by: Zhang, Jiaxin, et al.
Published: (2026)
by: Zhang, Jiaxin, et al.
Published: (2026)
Nested Learning: The Illusion of Deep Learning Architectures
by: Behrouz, Ali, et al.
Published: (2025)
by: Behrouz, Ali, et al.
Published: (2025)
Inference Time Context Sparsity: Illusion or Opportunity?
by: Joshi, Sahil, et al.
Published: (2026)
by: Joshi, Sahil, et al.
Published: (2026)
The Erasure Illusion: Stress-Testing the Generalization of LLM Forgetting Evaluation
by: Jia, Hengrui, et al.
Published: (2025)
by: Jia, Hengrui, et al.
Published: (2025)
An Illusion of Unlearning? Assessing Machine Unlearning Through Internal Representations
by: Gao, Yichen, et al.
Published: (2026)
by: Gao, Yichen, et al.
Published: (2026)
Bridging Reasoning to Learning: Unmasking Illusions using Complexity Out of Distribution Generalization
by: Paqaleh, Mohammad Mahdi Samiei, et al.
Published: (2025)
by: Paqaleh, Mohammad Mahdi Samiei, et al.
Published: (2025)
The Data-Quality Illusion: Rethinking Classifier-Based Quality Filtering for LLM Pretraining
by: Saada, Thiziri Nait, et al.
Published: (2025)
by: Saada, Thiziri Nait, et al.
Published: (2025)
The Narcissus Hypothesis: Descending to the Rung of Illusion
by: Cadei, Riccardo, et al.
Published: (2025)
by: Cadei, Riccardo, et al.
Published: (2025)
Inducing Group Fairness in Prompt-Based Language Model Decisions
by: Atwood, James, et al.
Published: (2024)
by: Atwood, James, et al.
Published: (2024)
The Instinctive Bias: Spurious Images lead to Illusion in MLLMs
by: Han, Tianyang, et al.
Published: (2024)
by: Han, Tianyang, et al.
Published: (2024)
The Illusion of State in State-Space Models
by: Merrill, William, et al.
Published: (2024)
by: Merrill, William, et al.
Published: (2024)
The Illusion of Certainty: Uncertainty Quantification for LLMs Fails under Ambiguity
by: Tomov, Tim, et al.
Published: (2025)
by: Tomov, Tim, et al.
Published: (2025)
LoRA vs Full Fine-tuning: An Illusion of Equivalence
by: Shuttleworth, Reece, et al.
Published: (2024)
by: Shuttleworth, Reece, et al.
Published: (2024)
From Astronomy to Astrology: Testing the Illusion of Zodiac-Based Personality Prediction with Machine Learning
by: Samantaray, Abhinna Sundar, et al.
Published: (2026)
by: Samantaray, Abhinna Sundar, et al.
Published: (2026)
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
by: Janiak, Denis, et al.
Published: (2025)
by: Janiak, Denis, et al.
Published: (2025)
The Illusion of Fit: Spatially Resolved Assessment of Constitutive Model Validity in Elastography and Physics-Based Inverse Problems
by: Scholz, Vincent C., et al.
Published: (2025)
by: Scholz, Vincent C., et al.
Published: (2025)
Exposing the Illusion of Fairness: Auditing Vulnerabilities to Distributional Manipulation Attacks
by: Lafargue, Valentin, et al.
Published: (2025)
by: Lafargue, Valentin, et al.
Published: (2025)
The Illusion of Intervention: Your LLM-Simulated Experiment is an Observational Study
by: Lin, Victoria, et al.
Published: (2026)
by: Lin, Victoria, et al.
Published: (2026)
The Illusion of Superposition? A Principled Analysis of Latent Thinking in Language Models
by: Rizvi-Martel, Michael, et al.
Published: (2026)
by: Rizvi-Martel, Michael, et al.
Published: (2026)
Omni-Modal Dissonance Benchmark: Systematically Breaking Modality Consensus to Probe Robustness and Calibrated Abstention
by: Nazi, Zabir Al, et al.
Published: (2026)
by: Nazi, Zabir Al, et al.
Published: (2026)
Similar Items
-
Formal Local Implication Between Two Neural Networks
by: Baninajjar, Anahita, et al.
Published: (2024) -
VNN: Verification-Friendly Neural Networks with Hard Robustness Guarantees
by: Baninajjar, Anahita, et al.
Published: (2023) -
BEFT: Bias-Efficient Fine-Tuning of Language Models in Low-Data Regimes
by: Huang, Baichuan, et al.
Published: (2025) -
Adversarial Illusions in Multi-Modal Embeddings
by: Zhang, Tingwei, et al.
Published: (2023) -
Breaking the Illusion: Real-world Challenges for Adversarial Patches in Object Detection
by: Shack, Jakob, et al.
Published: (2024)