VLMs Can Aggregate Scattered Training Patches
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Zhanhui, Chen, Lingjie, Yang, Chao, Lu, Chaochao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can VLMs Detect and Localize Fine-Grained AI-Edited Images?
von: Sun, Zhen, et al.
Veröffentlicht: (2025)
von: Sun, Zhen, et al.
Veröffentlicht: (2025)
DeepSight: An All-in-One LM Safety Toolkit
von: Zhang, Bo, et al.
Veröffentlicht: (2026)
von: Zhang, Bo, et al.
Veröffentlicht: (2026)
Beyond Boundaries: A Comprehensive Survey of Transferable Attacks on AI Systems
von: Wang, Guangjing, et al.
Veröffentlicht: (2023)
von: Wang, Guangjing, et al.
Veröffentlicht: (2023)
SafeGen: Mitigating Sexually Explicit Content Generation in Text-to-Image Models
von: Li, Xinfeng, et al.
Veröffentlicht: (2024)
von: Li, Xinfeng, et al.
Veröffentlicht: (2024)
CAAP: Capture-Aware Adversarial Patch Attacks on Palmprint Recognition Models
von: Liu, Renyang, et al.
Veröffentlicht: (2026)
von: Liu, Renyang, et al.
Veröffentlicht: (2026)
Navigation as Attackers Wish? Towards Building Robust Embodied Agents under Federated Learning
von: Zhang, Yunchao, et al.
Veröffentlicht: (2022)
von: Zhang, Yunchao, et al.
Veröffentlicht: (2022)
Rethinking Machine Unlearning in Image Generation Models
von: Liu, Renyang, et al.
Veröffentlicht: (2025)
von: Liu, Renyang, et al.
Veröffentlicht: (2025)
Training-Free Watermarking for Autoregressive Image Generation
von: Tong, Yu, et al.
Veröffentlicht: (2025)
von: Tong, Yu, et al.
Veröffentlicht: (2025)
Rel-Zero: Harnessing Patch-Pair Invariance for Robust Zero-Watermarking Against AI Editing
von: Chen, Pengzhen, et al.
Veröffentlicht: (2026)
von: Chen, Pengzhen, et al.
Veröffentlicht: (2026)
Jailbreaking Safeguarded Text-to-Image Models via Large Language Models
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2025)
von: Jiang, Zhengyuan, et al.
Veröffentlicht: (2025)
Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety
von: Ma, Xingjun, et al.
Veröffentlicht: (2025)
von: Ma, Xingjun, et al.
Veröffentlicht: (2025)
Robustness Analysis against Adversarial Patch Attacks in Fully Unmanned Stores
von: Na, Hyunsik, et al.
Veröffentlicht: (2025)
von: Na, Hyunsik, et al.
Veröffentlicht: (2025)
FakeIDet: Exploring Patches for Privacy-Preserving Fake ID Detection
von: Muñoz-Haro, Javier, et al.
Veröffentlicht: (2025)
von: Muñoz-Haro, Javier, et al.
Veröffentlicht: (2025)
Noise Aggregation Analysis Driven by Small-Noise Injection: Efficient Membership Inference for Diffusion Models
von: Li, Guo, et al.
Veröffentlicht: (2025)
von: Li, Guo, et al.
Veröffentlicht: (2025)
Physical Evaluation of Naturalistic Adversarial Patches for Camera-Based Traffic-Sign Detection
von: D'Urso, Brianna, et al.
Veröffentlicht: (2026)
von: D'Urso, Brianna, et al.
Veröffentlicht: (2026)
PII-VisBench: Evaluating Personally Identifiable Information Safety in Vision Language Models Along a Continuum of Visibility
von: Shahariar, G M, et al.
Veröffentlicht: (2026)
von: Shahariar, G M, et al.
Veröffentlicht: (2026)
VLSBench: Unveiling Visual Leakage in Multimodal Safety
von: Hu, Xuhao, et al.
Veröffentlicht: (2024)
von: Hu, Xuhao, et al.
Veröffentlicht: (2024)
The Curse of Recursion: Training on Generated Data Makes Models Forget
von: Shumailov, Ilia, et al.
Veröffentlicht: (2023)
von: Shumailov, Ilia, et al.
Veröffentlicht: (2023)
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security
von: Liu, Dongrui, et al.
Veröffentlicht: (2026)
von: Liu, Dongrui, et al.
Veröffentlicht: (2026)
ICAS: Detecting Training Data from Autoregressive Image Generative Models
von: Yu, Hongyao, et al.
Veröffentlicht: (2025)
von: Yu, Hongyao, et al.
Veröffentlicht: (2025)
CAPTCHA Solving for Native GUI Agents: Automated Reasoning-Action Data Generation and Self-Corrective Training
von: Chen, Yuxi, et al.
Veröffentlicht: (2026)
von: Chen, Yuxi, et al.
Veröffentlicht: (2026)
Who Can See Through You? Adversarial Shielding Against VLM-Based Attribute Inference Attacks
von: Fan, Yucheng, et al.
Veröffentlicht: (2025)
von: Fan, Yucheng, et al.
Veröffentlicht: (2025)
An Efficient and Multi-private Key Secure Aggregation for Federated Learning
von: Yang, Xue, et al.
Veröffentlicht: (2023)
von: Yang, Xue, et al.
Veröffentlicht: (2023)
Training-Free In-Context Forensic Chain for Image Manipulation Detection and Localization
von: Chen, Rui, et al.
Veröffentlicht: (2025)
von: Chen, Rui, et al.
Veröffentlicht: (2025)
On the Trustworthiness Landscape of State-of-the-art Generative Models: A Survey and Outlook
von: Fan, Mingyuan, et al.
Veröffentlicht: (2023)
von: Fan, Mingyuan, et al.
Veröffentlicht: (2023)
Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances
von: Lu, Shilin, et al.
Veröffentlicht: (2024)
von: Lu, Shilin, et al.
Veröffentlicht: (2024)
Can ChatGPT Perform Image Splicing Detection? A Preliminary Study
von: Nath, Souradip
Veröffentlicht: (2025)
von: Nath, Souradip
Veröffentlicht: (2025)
MOSSBench: Is Your Multimodal Language Model Oversensitive to Safe Queries?
von: Li, Xirui, et al.
Veröffentlicht: (2024)
von: Li, Xirui, et al.
Veröffentlicht: (2024)
X-Boundary: Establishing Exact Safety Boundary to Shield LLMs from Multi-Turn Jailbreaks without Compromising Usability
von: Lu, Xiaoya, et al.
Veröffentlicht: (2025)
von: Lu, Xiaoya, et al.
Veröffentlicht: (2025)
Chain-of-Jailbreak Attack for Image Generation Models via Editing Step by Step
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems
von: Zhang, Yihao, et al.
Veröffentlicht: (2026)
von: Zhang, Yihao, et al.
Veröffentlicht: (2026)
When Think-with-Image Meets Safety: What Determines Multimodal Jailbreak Robustness?
von: Tian, Yuan, et al.
Veröffentlicht: (2026)
von: Tian, Yuan, et al.
Veröffentlicht: (2026)
A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?
von: Chen, Ada, et al.
Veröffentlicht: (2025)
von: Chen, Ada, et al.
Veröffentlicht: (2025)
DIFFender: Diffusion-Based Adversarial Defense against Patch Attacks
von: Kang, Caixin, et al.
Veröffentlicht: (2023)
von: Kang, Caixin, et al.
Veröffentlicht: (2023)
When Understanding Becomes a Risk: Authenticity and Safety Risks in the Emerging Image Generation Paradigm
von: Leng, Ye, et al.
Veröffentlicht: (2026)
von: Leng, Ye, et al.
Veröffentlicht: (2026)
The Safety Reminder: A Soft Prompt to Reactivate Delayed Safety Awareness in Vision-Language Models
von: Tang, Peiyuan, et al.
Veröffentlicht: (2025)
von: Tang, Peiyuan, et al.
Veröffentlicht: (2025)
Real-world Adversarial Defense against Patch Attacks based on Diffusion Model
von: Wei, Xingxing, et al.
Veröffentlicht: (2024)
von: Wei, Xingxing, et al.
Veröffentlicht: (2024)
ArtPerception: ASCII Art-based Jailbreak on LLMs with Recognition Pre-test
von: Yang, Guan-Yan, et al.
Veröffentlicht: (2025)
von: Yang, Guan-Yan, et al.
Veröffentlicht: (2025)
Dormant: Defending against Pose-driven Human Image Animation
von: Zhou, Jiachen, et al.
Veröffentlicht: (2024)
von: Zhou, Jiachen, et al.
Veröffentlicht: (2024)
Membership Inference Attacks against Large Vision-Language Models
von: Li, Zhan, et al.
Veröffentlicht: (2024)
von: Li, Zhan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Can VLMs Detect and Localize Fine-Grained AI-Edited Images?
von: Sun, Zhen, et al.
Veröffentlicht: (2025) -
DeepSight: An All-in-One LM Safety Toolkit
von: Zhang, Bo, et al.
Veröffentlicht: (2026) -
Beyond Boundaries: A Comprehensive Survey of Transferable Attacks on AI Systems
von: Wang, Guangjing, et al.
Veröffentlicht: (2023) -
SafeGen: Mitigating Sexually Explicit Content Generation in Text-to-Image Models
von: Li, Xinfeng, et al.
Veröffentlicht: (2024) -
CAAP: Capture-Aware Adversarial Patch Attacks on Palmprint Recognition Models
von: Liu, Renyang, et al.
Veröffentlicht: (2026)