Robustness of Vision Language Models Against Split-Image Harmful Input Attacks
Fuente:
arXiv
Saved in:
| Main Authors: | Rashid, Md Rafi Ur, Shanto, MD Sadik Hossain, Dasu, Vishnu Asutosh, Mehnaz, Shagufta |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Chain-of-Thought Driven Adversarial Scenario Extrapolation for Robust Language Models
by: Rashid, Md Rafi Ur, et al.
Published: (2025)
by: Rashid, Md Rafi Ur, et al.
Published: (2025)
Gradient-Free Privacy Leakage in Federated Language Models through Selective Weight Tampering
by: Rashid, Md Rafi Ur, et al.
Published: (2023)
by: Rashid, Md Rafi Ur, et al.
Published: (2023)
AttMetNet: Attention-Enhanced Deep Neural Network for Methane Plume Detection in Sentinel-2 Satellite Imagery
by: Ahsan, Rakib, et al.
Published: (2025)
by: Ahsan, Rakib, et al.
Published: (2025)
Second-Order Information Matters: Revisiting Machine Unlearning for Large Language Models
by: Gu, Kang, et al.
Published: (2024)
by: Gu, Kang, et al.
Published: (2024)
Securing Vision-Language Models with a Robust Encoder Against Jailbreak and Adversarial Attacks
by: Hossain, Md Zarif, et al.
Published: (2024)
by: Hossain, Md Zarif, et al.
Published: (2024)
The Art of Saying "Maybe": A Conformal Lens for Uncertainty Benchmarking in VLMs
by: Azad, Asif, et al.
Published: (2025)
by: Azad, Asif, et al.
Published: (2025)
SequentialBreak: Large Language Models Can be Fooled by Embedding Jailbreak Prompts into Sequential Prompt Chains
by: Saiem, Bijoy Ahmed, et al.
Published: (2024)
by: Saiem, Bijoy Ahmed, et al.
Published: (2024)
From Insight to Exploit: Leveraging LLM Collaboration for Adaptive Adversarial Text Generation
by: Sultana, Najrin, et al.
Published: (2025)
by: Sultana, Najrin, et al.
Published: (2025)
Forget to Flourish: Leveraging Machine-Unlearning on Pretrained Language Models for Privacy Leakage
by: Rashid, Md Rafi Ur, et al.
Published: (2024)
by: Rashid, Md Rafi Ur, et al.
Published: (2024)
Attention Pruning: Automated Fairness Repair of Language Models via Surrogate Simulated Annealing
by: Dasu, Vishnu Asutosh, et al.
Published: (2025)
by: Dasu, Vishnu Asutosh, et al.
Published: (2025)
Heimdall: Formally Verified Automated Migration of Legacy eBPF Programs to Rust
by: Dasu, Vishnu Asutosh, et al.
Published: (2026)
by: Dasu, Vishnu Asutosh, et al.
Published: (2026)
Improving Noise Efficiency in Privacy-preserving Dataset Distillation
by: Zheng, Runkai, et al.
Published: (2025)
by: Zheng, Runkai, et al.
Published: (2025)
Chain of Attack: On the Robustness of Vision-Language Models Against Transfer-Based Adversarial Attacks
by: Xie, Peng, et al.
Published: (2024)
by: Xie, Peng, et al.
Published: (2024)
Exploring Adversarial Watermarking in Transformer-Based Models: Transferability and Robustness Against Defense Mechanism for Medical Images
by: Sadik, Rifat, et al.
Published: (2025)
by: Sadik, Rifat, et al.
Published: (2025)
Real-Time Detection and Analysis of Vehicles and Pedestrians using Deep Learning
by: Sadik, Md Nahid, et al.
Published: (2024)
by: Sadik, Md Nahid, et al.
Published: (2024)
Privacy-Preserving Data Deduplication for Enhancing Federated Learning of Language Models (Extended Version)
by: Abadi, Aydin, et al.
Published: (2024)
by: Abadi, Aydin, et al.
Published: (2024)
On the Robustness of GUI Grounding Models Against Image Attacks
by: Zhao, Haoren, et al.
Published: (2025)
by: Zhao, Haoren, et al.
Published: (2025)
Robust Defense Strategies for Multimodal Contrastive Learning: Efficient Fine-tuning Against Backdoor Attacks
by: Hossain, Md. Iqbal, et al.
Published: (2025)
by: Hossain, Md. Iqbal, et al.
Published: (2025)
Rice Leaf Disease Detection: A Comparative Study Between CNN, Transformer and Non-neural Network Architectures
by: Mehnaz, Samia, et al.
Published: (2025)
by: Mehnaz, Samia, et al.
Published: (2025)
TrojVLM: Backdoor Attack Against Vision Language Models
by: Lyu, Weimin, et al.
Published: (2024)
by: Lyu, Weimin, et al.
Published: (2024)
DFCon: Attention-Driven Supervised Contrastive Learning for Robust Deepfake Detection
by: Shanto, MD Sadik Hossain, et al.
Published: (2025)
by: Shanto, MD Sadik Hossain, et al.
Published: (2025)
Preemptive Hallucination Reduction: An Input-Level Approach for Multimodal Language Model
by: Arif, Nokimul Hasan, et al.
Published: (2025)
by: Arif, Nokimul Hasan, et al.
Published: (2025)
Reliable Deep Learning for Small-Scale Classifications: Experiments on Real-World Image Datasets from Bangladesh
by: Suny, Alfe, et al.
Published: (2026)
by: Suny, Alfe, et al.
Published: (2026)
Sim-CLIP: Unsupervised Siamese Adversarial Fine-Tuning for Robust and Semantically-Rich Vision-Language Models
by: Hossain, Md Zarif, et al.
Published: (2024)
by: Hossain, Md Zarif, et al.
Published: (2024)
Speak2Sign3D: A Multi-modal Pipeline for English Speech to American Sign Language Animation
by: Rahman, Kazi Mahathir, et al.
Published: (2025)
by: Rahman, Kazi Mahathir, et al.
Published: (2025)
Steering Away from Harm: An Adaptive Approach to Defending Vision Language Model Against Jailbreaks
by: Wang, Han, et al.
Published: (2024)
by: Wang, Han, et al.
Published: (2024)
Robust Vision-Language Models via Tensor Decomposition: A Defense Against Adversarial Attacks
by: Patel, Het, et al.
Published: (2025)
by: Patel, Het, et al.
Published: (2025)
The Bias of Harmful Label Associations in Vision-Language Models
by: Hazirbas, Caner, et al.
Published: (2024)
by: Hazirbas, Caner, et al.
Published: (2024)
HeBA: Heterogeneous Bottleneck Adapters for Robust Vision-Language Models
by: Islam, Md Jahidul
Published: (2026)
by: Islam, Md Jahidul
Published: (2026)
Impact of Data Duplication on Deep Neural Network-Based Image Classifiers: Robust vs. Standard Models
by: Aghabagherloo, Alireza, et al.
Published: (2025)
by: Aghabagherloo, Alireza, et al.
Published: (2025)
Demonstration of an Adversarial Attack Against a Multimodal Vision Language Model for Pathology Imaging
by: Thota, Poojitha, et al.
Published: (2024)
by: Thota, Poojitha, et al.
Published: (2024)
Improving Robustness of Vision-Language-Action Models by Restoring Corrupted Visual Inputs
by: Orjuela, Daniel Yezid Guarnizo, et al.
Published: (2026)
by: Orjuela, Daniel Yezid Guarnizo, et al.
Published: (2026)
Vision-Language Models for Automated Chest X-ray Interpretation: Leveraging ViT and GPT-2
by: Islam, Md. Rakibul, et al.
Published: (2025)
by: Islam, Md. Rakibul, et al.
Published: (2025)
BioAutoML-NAS: An End-to-End AutoML Framework for Multimodal Insect Classification via Neural Architecture Search on Large-Scale Biodiversity Data
by: Abian, Arefin Ittesafun, et al.
Published: (2025)
by: Abian, Arefin Ittesafun, et al.
Published: (2025)
Beyond Visual Safety: Jailbreaking Multimodal Large Language Models for Harmful Image Generation via Semantic-Agnostic Inputs
by: Yu, Mingyu, et al.
Published: (2026)
by: Yu, Mingyu, et al.
Published: (2026)
PDA: Text-Augmented Defense Framework for Robust Vision-Language Models against Adversarial Image Attacks
by: Xu, Jingning, et al.
Published: (2026)
by: Xu, Jingning, et al.
Published: (2026)
Semantic Shield: Defending Vision-Language Models Against Backdooring and Poisoning via Fine-grained Knowledge Alignment
by: Ishmam, Alvi Md, et al.
Published: (2024)
by: Ishmam, Alvi Md, et al.
Published: (2024)
AdaptoVision: A Multi-Resolution Image Recognition Model for Robust and Scalable Classification
by: Sabrin, Md. Sanaullah Chowdhury Lameya
Published: (2025)
by: Sabrin, Md. Sanaullah Chowdhury Lameya
Published: (2025)
InsideOut: An EfficientNetV2-S Based Deep Learning Framework for Robust Multi-Class Facial Emotion Recognition
by: Farabi, Ahsan, et al.
Published: (2025)
by: Farabi, Ahsan, et al.
Published: (2025)
GNNBleed: Inference Attacks to Unveil Private Edges in Graphs with Realistic Access to GNN Models
by: Song, Zeyu, et al.
Published: (2023)
by: Song, Zeyu, et al.
Published: (2023)
Similar Items
-
Chain-of-Thought Driven Adversarial Scenario Extrapolation for Robust Language Models
by: Rashid, Md Rafi Ur, et al.
Published: (2025) -
Gradient-Free Privacy Leakage in Federated Language Models through Selective Weight Tampering
by: Rashid, Md Rafi Ur, et al.
Published: (2023) -
AttMetNet: Attention-Enhanced Deep Neural Network for Methane Plume Detection in Sentinel-2 Satellite Imagery
by: Ahsan, Rakib, et al.
Published: (2025) -
Second-Order Information Matters: Revisiting Machine Unlearning for Large Language Models
by: Gu, Kang, et al.
Published: (2024) -
Securing Vision-Language Models with a Robust Encoder Against Jailbreak and Adversarial Attacks
by: Hossain, Md Zarif, et al.
Published: (2024)