Modal Aphasia: Can Unified Multimodal Models Describe Images From Memory?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aerni, Michael, Swanson, Joshua, Nikolić, Kristina, Tramèr, Florian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Image Can Bring Your Memory Back: A Novel Multi-Modal Guided Attack against Image Generation Model Unlearning
von: Liu, Renyang, et al.
Veröffentlicht: (2025)
von: Liu, Renyang, et al.
Veröffentlicht: (2025)
Evaluations of Machine Learning Privacy Defenses are Misleading
von: Aerni, Michael, et al.
Veröffentlicht: (2024)
von: Aerni, Michael, et al.
Veröffentlicht: (2024)
Probabilistic Modeling of Jailbreak on Multimodal LLMs: From Quantification to Application
von: Xu, Wenzhuo, et al.
Veröffentlicht: (2025)
von: Xu, Wenzhuo, et al.
Veröffentlicht: (2025)
A Cross-Modal Prompt Injection Attack against Large Vision-Language Models with Image-Only Perturbation
von: Yang, Hao, et al.
Veröffentlicht: (2026)
von: Yang, Hao, et al.
Veröffentlicht: (2026)
Large-scale online deanonymization with LLMs
von: Lermen, Simon, et al.
Veröffentlicht: (2026)
von: Lermen, Simon, et al.
Veröffentlicht: (2026)
Fusion is Not Enough: Single Modal Attacks on Fusion Models for 3D Object Detection
von: Cheng, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Cheng, Zhiyuan, et al.
Veröffentlicht: (2023)
Membership Inference Attacks on Sequence Models
von: Rossi, Lorenzo, et al.
Veröffentlicht: (2025)
von: Rossi, Lorenzo, et al.
Veröffentlicht: (2025)
OmniSafeBench-MM: A Unified Benchmark and Toolbox for Multimodal Jailbreak Attack-Defense Evaluation
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2025)
Unbridled Icarus: A Survey of the Potential Perils of Image Inputs in Multimodal Large Language Model Security
von: Fan, Yihe, et al.
Veröffentlicht: (2024)
von: Fan, Yihe, et al.
Veröffentlicht: (2024)
MMA-Diffusion: MultiModal Attack on Diffusion Models
von: Yang, Yijun, et al.
Veröffentlicht: (2023)
von: Yang, Yijun, et al.
Veröffentlicht: (2023)
SoK: Can Synthetic Images Replace Real Data? A Survey of Utility and Privacy of Synthetic Image Generation
von: Chung, Yunsung, et al.
Veröffentlicht: (2025)
von: Chung, Yunsung, et al.
Veröffentlicht: (2025)
From Evidence to Verdict: An Agent-Based Forensic Framework for AI-Generated Image Detection
von: Liang, Mengfei, et al.
Veröffentlicht: (2025)
von: Liang, Mengfei, et al.
Veröffentlicht: (2025)
Solutions to Deepfakes: Can Camera Hardware, Cryptography, and Deep Learning Verify Real Images?
von: Vilesov, Alexander, et al.
Veröffentlicht: (2024)
von: Vilesov, Alexander, et al.
Veröffentlicht: (2024)
Evolving Contextual Safety in Multi-Modal Large Language Models via Inference-Time Self-Reflective Memory
von: Zhang, Ce, et al.
Veröffentlicht: (2026)
von: Zhang, Ce, et al.
Veröffentlicht: (2026)
Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
von: Ying, Zonghao, et al.
Veröffentlicht: (2024)
FreqCross: A Multi-Modal Frequency-Spatial Fusion Network for Robust Detection of Stable Diffusion 3.5 Generated Images
von: Yang, Guang
Veröffentlicht: (2025)
von: Yang, Guang
Veröffentlicht: (2025)
When Memory Becomes a Vulnerability: Towards Multi-turn Jailbreak Attacks against Text-to-Image Generation Systems
von: Zhao, Shiqian, et al.
Veröffentlicht: (2025)
von: Zhao, Shiqian, et al.
Veröffentlicht: (2025)
Purified and Unified Steganographic Network
von: Li, Guobiao, et al.
Veröffentlicht: (2024)
von: Li, Guobiao, et al.
Veröffentlicht: (2024)
Provably Secure Robust Image Steganography via Cross-Modal Error Correction
von: Qi, Yuang, et al.
Veröffentlicht: (2024)
von: Qi, Yuang, et al.
Veröffentlicht: (2024)
Universally Unfiltered and Unseen:Input-Agnostic Multimodal Jailbreaks against Text-to-Image Model Safeguards
von: Yan, Song, et al.
Veröffentlicht: (2025)
von: Yan, Song, et al.
Veröffentlicht: (2025)
VLATTACK: Multimodal Adversarial Attacks on Vision-Language Tasks via Pre-trained Models
von: Yin, Ziyi, et al.
Veröffentlicht: (2023)
von: Yin, Ziyi, et al.
Veröffentlicht: (2023)
AMB-FHE: Adaptive Multi-biometric Fusion with Fully Homomorphic Encryption
von: Bayer, Florian, et al.
Veröffentlicht: (2025)
von: Bayer, Florian, et al.
Veröffentlicht: (2025)
From Easy to Hard++: Promoting Differentially Private Image Synthesis Through Spatial-Frequency Curriculum
von: Gong, Chen, et al.
Veröffentlicht: (2026)
von: Gong, Chen, et al.
Veröffentlicht: (2026)
Espresso: Robust Concept Filtering in Text-to-Image Models
von: Das, Anudeep, et al.
Veröffentlicht: (2024)
von: Das, Anudeep, et al.
Veröffentlicht: (2024)
Membership Inference Attack Against Masked Image Modeling
von: Li, Zheng, et al.
Veröffentlicht: (2024)
von: Li, Zheng, et al.
Veröffentlicht: (2024)
Where the Devil Hides: Deepfake Detectors Can No Longer Be Trusted
von: Yuan, Shuaiwei, et al.
Veröffentlicht: (2025)
von: Yuan, Shuaiwei, et al.
Veröffentlicht: (2025)
UVL2: A Unified Framework for Video Tampering Localization
von: Pei, Pengfei
Veröffentlicht: (2023)
von: Pei, Pengfei
Veröffentlicht: (2023)
InverTune: Removing Backdoors from Multimodal Contrastive Learning Models via Trigger Inversion and Activation Tuning
von: Sun, Mengyuan, et al.
Veröffentlicht: (2025)
von: Sun, Mengyuan, et al.
Veröffentlicht: (2025)
On the Generation and Mitigation of Harmful Geometry in Image-to-3D Models
von: Liu, Yule, et al.
Veröffentlicht: (2026)
von: Liu, Yule, et al.
Veröffentlicht: (2026)
Federated Learning for Large Models in Medical Imaging: A Comprehensive Review
von: Sun, Mengyu, et al.
Veröffentlicht: (2025)
von: Sun, Mengyu, et al.
Veröffentlicht: (2025)
Token-Level Constraint Boundary Search for Jailbreaking Text-to-Image Models
von: Liu, Jiangtao, et al.
Veröffentlicht: (2025)
von: Liu, Jiangtao, et al.
Veröffentlicht: (2025)
Gungnir: Exploiting Stylistic Features in Images for Backdoor Attacks on Diffusion Models
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
Patronus: Safeguarding Text-to-Image Models against White-Box Adversaries
von: Li, Xinfeng, et al.
Veröffentlicht: (2025)
von: Li, Xinfeng, et al.
Veröffentlicht: (2025)
Neighbor-Aware Localized Concept Erasure in Text-to-Image Diffusion Models
von: Shi, Zhuan, et al.
Veröffentlicht: (2026)
von: Shi, Zhuan, et al.
Veröffentlicht: (2026)
Decomposing Private Image Generation via Coarse-to-Fine Wavelet Modeling
von: Bayrooti, Jasmine, et al.
Veröffentlicht: (2026)
von: Bayrooti, Jasmine, et al.
Veröffentlicht: (2026)
HTS-Attack: Heuristic Token Search for Jailbreaking Text-to-Image Models
von: Gao, Sensen, et al.
Veröffentlicht: (2024)
von: Gao, Sensen, et al.
Veröffentlicht: (2024)
Gaussian Shading: Provable Performance-Lossless Image Watermarking for Diffusion Models
von: Yang, Zijin, et al.
Veröffentlicht: (2024)
von: Yang, Zijin, et al.
Veröffentlicht: (2024)
Stable Signature is Unstable: Removing Image Watermark from Diffusion Models
von: Hu, Yuepeng, et al.
Veröffentlicht: (2024)
von: Hu, Yuepeng, et al.
Veröffentlicht: (2024)
FIDAVL: Fake Image Detection and Attribution using Vision-Language Model
von: Keita, Mamadou, et al.
Veröffentlicht: (2024)
von: Keita, Mamadou, et al.
Veröffentlicht: (2024)
Intriguing Properties of Diffusion Models: An Empirical Study of the Natural Attack Capability in Text-to-Image Generative Models
von: Sato, Takami, et al.
Veröffentlicht: (2023)
von: Sato, Takami, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Image Can Bring Your Memory Back: A Novel Multi-Modal Guided Attack against Image Generation Model Unlearning
von: Liu, Renyang, et al.
Veröffentlicht: (2025) -
Evaluations of Machine Learning Privacy Defenses are Misleading
von: Aerni, Michael, et al.
Veröffentlicht: (2024) -
Probabilistic Modeling of Jailbreak on Multimodal LLMs: From Quantification to Application
von: Xu, Wenzhuo, et al.
Veröffentlicht: (2025) -
A Cross-Modal Prompt Injection Attack against Large Vision-Language Models with Image-Only Perturbation
von: Yang, Hao, et al.
Veröffentlicht: (2026) -
Large-scale online deanonymization with LLMs
von: Lermen, Simon, et al.
Veröffentlicht: (2026)