Self-adaptive Dataset Construction for Real-World Multimodal Safety Scenarios
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qu, Jingen, Li, Lijun, Zhang, Bo, Yan, Yichen, Shao, Jing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Contextual Image Attack: How Visual Context Exposes Multimodal Safety Vulnerabilities
von: Xiong, Yuan, et al.
Veröffentlicht: (2025)
von: Xiong, Yuan, et al.
Veröffentlicht: (2025)
Rethinking Bottlenecks in Safety Fine-Tuning of Vision Language Models
von: Ding, Yi, et al.
Veröffentlicht: (2025)
von: Ding, Yi, et al.
Veröffentlicht: (2025)
Visual Contextual Attack: Jailbreaking MLLMs with Image-Driven Context Injection
von: Miao, Ziqi, et al.
Veröffentlicht: (2025)
von: Miao, Ziqi, et al.
Veröffentlicht: (2025)
VLSBench: Unveiling Visual Leakage in Multimodal Safety
von: Hu, Xuhao, et al.
Veröffentlicht: (2024)
von: Hu, Xuhao, et al.
Veröffentlicht: (2024)
DeepSight: An All-in-One LM Safety Toolkit
von: Zhang, Bo, et al.
Veröffentlicht: (2026)
von: Zhang, Bo, et al.
Veröffentlicht: (2026)
Evolving Contextual Safety in Multi-Modal Large Language Models via Inference-Time Self-Reflective Memory
von: Zhang, Ce, et al.
Veröffentlicht: (2026)
von: Zhang, Ce, et al.
Veröffentlicht: (2026)
MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance
von: Pi, Renjie, et al.
Veröffentlicht: (2024)
von: Pi, Renjie, et al.
Veröffentlicht: (2024)
SafeGRPO: Self-Rewarded Multimodal Safety Alignment via Rule-Governed Policy Optimization
von: Rong, Xuankun, et al.
Veröffentlicht: (2025)
von: Rong, Xuankun, et al.
Veröffentlicht: (2025)
CIS-BA: Continuous Interaction Space Based Backdoor Attack for Object Detection in the Real-World
von: Zhao, Shuxin, et al.
Veröffentlicht: (2025)
von: Zhao, Shuxin, et al.
Veröffentlicht: (2025)
Iteratively Prompting Multimodal LLMs to Reproduce Natural and AI-Generated Images
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety
von: Ma, Xingjun, et al.
Veröffentlicht: (2025)
von: Ma, Xingjun, et al.
Veröffentlicht: (2025)
UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images
von: Qu, Yiting, et al.
Veröffentlicht: (2024)
von: Qu, Yiting, et al.
Veröffentlicht: (2024)
Efficient Backdoor Attacks for Deep Neural Networks in Real-world Scenarios
von: Li, Ziqiang, et al.
Veröffentlicht: (2023)
von: Li, Ziqiang, et al.
Veröffentlicht: (2023)
Dissecting Adversarial Robustness of Multimodal LM Agents
von: Wu, Chen Henry, et al.
Veröffentlicht: (2024)
von: Wu, Chen Henry, et al.
Veröffentlicht: (2024)
Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injection
von: Yang, Wenkui, et al.
Veröffentlicht: (2026)
von: Yang, Wenkui, et al.
Veröffentlicht: (2026)
X-Boundary: Establishing Exact Safety Boundary to Shield LLMs from Multi-Turn Jailbreaks without Compromising Usability
von: Lu, Xiaoya, et al.
Veröffentlicht: (2025)
von: Lu, Xiaoya, et al.
Veröffentlicht: (2025)
SlowBA: An efficiency backdoor attack towards VLM-based GUI agents
von: Li, Junxian, et al.
Veröffentlicht: (2026)
von: Li, Junxian, et al.
Veröffentlicht: (2026)
Benchmarking Large Multimodal Models against Common Corruptions
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
von: Zhang, Jiawei, et al.
Veröffentlicht: (2024)
IAG: Input-aware Backdoor Attack on VLM-based Visual Grounding
von: Li, Junxian, et al.
Veröffentlicht: (2025)
von: Li, Junxian, et al.
Veröffentlicht: (2025)
Image-Based Geolocation Using Large Vision-Language Models
von: Liu, Yi, et al.
Veröffentlicht: (2024)
von: Liu, Yi, et al.
Veröffentlicht: (2024)
When Think-with-Image Meets Safety: What Determines Multimodal Jailbreak Robustness?
von: Tian, Yuan, et al.
Veröffentlicht: (2026)
von: Tian, Yuan, et al.
Veröffentlicht: (2026)
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security
von: Liu, Dongrui, et al.
Veröffentlicht: (2026)
von: Liu, Dongrui, et al.
Veröffentlicht: (2026)
Jailbreaking Attack against Multimodal Large Language Model
von: Niu, Zhenxing, et al.
Veröffentlicht: (2024)
von: Niu, Zhenxing, et al.
Veröffentlicht: (2024)
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
von: Ye, Mang, et al.
Veröffentlicht: (2025)
von: Ye, Mang, et al.
Veröffentlicht: (2025)
Privacy-Preserving Federated Learning with Verifiable Fairness Guarantees
von: Ali, Mohammed Himayath, et al.
Veröffentlicht: (2026)
von: Ali, Mohammed Himayath, et al.
Veröffentlicht: (2026)
Generating Synthetic Data with Formal Privacy Guarantees: State of the Art and the Road Ahead
von: Schlegel, Viktor, et al.
Veröffentlicht: (2025)
von: Schlegel, Viktor, et al.
Veröffentlicht: (2025)
Effective and Efficient Adversarial Detection for Vision-Language Models via A Single Vector
von: Huang, Youcheng, et al.
Veröffentlicht: (2024)
von: Huang, Youcheng, et al.
Veröffentlicht: (2024)
Are Vision-Language Models Safe in the Wild? A Meme-Based Benchmark Study
von: Lee, DongGeon, et al.
Veröffentlicht: (2025)
von: Lee, DongGeon, et al.
Veröffentlicht: (2025)
Doubly-Universal Adversarial Perturbations: Deceiving Vision-Language Models Across Both Images and Text with a Single Perturbation
von: Kim, Hee-Seon, et al.
Veröffentlicht: (2024)
von: Kim, Hee-Seon, et al.
Veröffentlicht: (2024)
BSPA: Exploring Black-box Stealthy Prompt Attacks against Image Generators
von: Tian, Yu, et al.
Veröffentlicht: (2024)
von: Tian, Yu, et al.
Veröffentlicht: (2024)
DAVSP: Safety Alignment for Large Vision-Language Models via Deep Aligned Visual Safety Prompt
von: Zhang, Yitong, et al.
Veröffentlicht: (2025)
von: Zhang, Yitong, et al.
Veröffentlicht: (2025)
PII-VisBench: Evaluating Personally Identifiable Information Safety in Vision Language Models Along a Continuum of Visibility
von: Shahariar, G M, et al.
Veröffentlicht: (2026)
von: Shahariar, G M, et al.
Veröffentlicht: (2026)
The Safety Reminder: A Soft Prompt to Reactivate Delayed Safety Awareness in Vision-Language Models
von: Tang, Peiyuan, et al.
Veröffentlicht: (2025)
von: Tang, Peiyuan, et al.
Veröffentlicht: (2025)
InverTune: Removing Backdoors from Multimodal Contrastive Learning Models via Trigger Inversion and Activation Tuning
von: Sun, Mengyuan, et al.
Veröffentlicht: (2025)
von: Sun, Mengyuan, et al.
Veröffentlicht: (2025)
BodhiPromptShield: Pre-Inference Prompt Mediation for Suppressing Privacy Propagation in LLM/VLM Agents
von: Ma, Bo, et al.
Veröffentlicht: (2026)
von: Ma, Bo, et al.
Veröffentlicht: (2026)
Towards Physical World Backdoor Attacks against Skeleton Action Recognition
von: Zheng, Qichen, et al.
Veröffentlicht: (2024)
von: Zheng, Qichen, et al.
Veröffentlicht: (2024)
Few-Shot Adversarial Prompt Learning on Vision-Language Models
von: Zhou, Yiwei, et al.
Veröffentlicht: (2024)
von: Zhou, Yiwei, et al.
Veröffentlicht: (2024)
Bodhi VLM: Privacy-Alignment Modeling for Hierarchical Visual Representations in Vision Backbones and VLM Encoders via Bottom-Up and Top-Down Feature Search
von: Ma, Bo, et al.
Veröffentlicht: (2026)
von: Ma, Bo, et al.
Veröffentlicht: (2026)
REAEDP: Entropy-Calibrated Differentially Private Data Release with Formal Guarantees and Attack-Based Evaluation
von: Ma, Bo, et al.
Veröffentlicht: (2026)
von: Ma, Bo, et al.
Veröffentlicht: (2026)
SemDP: Semantic-level Differential Privacy Protection for Face Datasets
von: Zhang, Xiaoting, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoting, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Contextual Image Attack: How Visual Context Exposes Multimodal Safety Vulnerabilities
von: Xiong, Yuan, et al.
Veröffentlicht: (2025) -
Rethinking Bottlenecks in Safety Fine-Tuning of Vision Language Models
von: Ding, Yi, et al.
Veröffentlicht: (2025) -
Visual Contextual Attack: Jailbreaking MLLMs with Image-Driven Context Injection
von: Miao, Ziqi, et al.
Veröffentlicht: (2025) -
VLSBench: Unveiling Visual Leakage in Multimodal Safety
von: Hu, Xuhao, et al.
Veröffentlicht: (2024) -
DeepSight: An All-in-One LM Safety Toolkit
von: Zhang, Bo, et al.
Veröffentlicht: (2026)