Image-Perfect Imperfections: Safety, Bias, and Authenticity in the Shadow of Text-To-Image Model Evolution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Yixin, Shen, Yun, Backes, Michael, Zhang, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prompt Stealing Attacks Against Text-to-Image Generation Models
von: Shen, Xinyue, et al.
Veröffentlicht: (2023)
von: Shen, Xinyue, et al.
Veröffentlicht: (2023)
On the Proactive Generation of Unsafe Images From Text-To-Image Models Using Benign Prompts
von: Wu, Yixin, et al.
Veröffentlicht: (2023)
von: Wu, Yixin, et al.
Veröffentlicht: (2023)
The Challenge of Identifying the Origin of Black-Box Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
Voice Jailbreak Attacks Against GPT-4o
von: Shen, Xinyue, et al.
Veröffentlicht: (2024)
von: Shen, Xinyue, et al.
Veröffentlicht: (2024)
Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
von: Wu, Yixin, et al.
Veröffentlicht: (2025)
When GPT Spills the Tea: Comprehensive Assessment of Knowledge File Leakage in GPTs
von: Shen, Xinyue, et al.
Veröffentlicht: (2025)
von: Shen, Xinyue, et al.
Veröffentlicht: (2025)
"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models
von: Shen, Xinyue, et al.
Veröffentlicht: (2023)
von: Shen, Xinyue, et al.
Veröffentlicht: (2023)
HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
von: Shen, Xinyue, et al.
Veröffentlicht: (2025)
von: Shen, Xinyue, et al.
Veröffentlicht: (2025)
MGTBench: Benchmarking Machine-Generated Text Detection
von: He, Xinlei, et al.
Veröffentlicht: (2023)
von: He, Xinlei, et al.
Veröffentlicht: (2023)
Asymmetric Bias in Text-to-Image Generation with Adversarial Attacks
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2023)
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2023)
Composite Backdoor Attacks Against Large Language Models
von: Huang, Hai, et al.
Veröffentlicht: (2023)
von: Huang, Hai, et al.
Veröffentlicht: (2023)
GEO-Detective: Unveiling Location Privacy Risks in Images with LLM Agents
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2025)
Link Stealing Attacks Against Inductive Graph Neural Networks
von: Wu, Yixin, et al.
Veröffentlicht: (2024)
von: Wu, Yixin, et al.
Veröffentlicht: (2024)
Breaking Agents: Compromising Autonomous LLM Agents Through Malfunction Amplification
von: Zhang, Boyang, et al.
Veröffentlicht: (2024)
von: Zhang, Boyang, et al.
Veröffentlicht: (2024)
Understanding Data Importance in Machine Learning Attacks: Does Valuable Data Pose Greater Harm?
von: Wen, Rui, et al.
Veröffentlicht: (2024)
von: Wen, Rui, et al.
Veröffentlicht: (2024)
SoK: Data Reconstruction Attacks Against Machine Learning Models: Definition, Metrics, and Benchmark
von: Wen, Rui, et al.
Veröffentlicht: (2025)
von: Wen, Rui, et al.
Veröffentlicht: (2025)
Instruction Backdoor Attacks Against Customized LLMs
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
Sparse Models, Sparse Safety: Unsafe Routes in Mixture-of-Experts LLMs
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
von: Jiang, Yukun, et al.
Veröffentlicht: (2026)
Vera Verto: Multimodal Hijacking Attack
von: Zhang, Minxing, et al.
Veröffentlicht: (2024)
von: Zhang, Minxing, et al.
Veröffentlicht: (2024)
FameBias: Embedding Manipulation Bias Attack in Text-to-Image Models
von: Roh, Jaechul, et al.
Veröffentlicht: (2024)
von: Roh, Jaechul, et al.
Veröffentlicht: (2024)
Transferable Availability Poisoning Attacks
von: Liu, Yiyong, et al.
Veröffentlicht: (2023)
von: Liu, Yiyong, et al.
Veröffentlicht: (2023)
Excessive Reasoning Attack on Reasoning LLMs
von: Si, Wai Man, et al.
Veröffentlicht: (2025)
von: Si, Wai Man, et al.
Veröffentlicht: (2025)
SOS! Soft Prompt Attack Against Open-Source Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2024)
von: Yang, Ziqing, et al.
Veröffentlicht: (2024)
When Understanding Becomes a Risk: Authenticity and Safety Risks in the Emerging Image Generation Paradigm
von: Leng, Ye, et al.
Veröffentlicht: (2026)
von: Leng, Ye, et al.
Veröffentlicht: (2026)
Model Stealing Attack against Graph Classification with Authenticity, Uncertainty and Diversity
von: Zhu, Zhihao, et al.
Veröffentlicht: (2023)
von: Zhu, Zhihao, et al.
Veröffentlicht: (2023)
Awakening the Hydra: Stabilizing Multi-Concept Backdoor Injection in Text-to-Image Diffusion Models
von: Wang, Kai, et al.
Veröffentlicht: (2026)
von: Wang, Kai, et al.
Veröffentlicht: (2026)
UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images
von: Qu, Yiting, et al.
Veröffentlicht: (2024)
von: Qu, Yiting, et al.
Veröffentlicht: (2024)
Text-to-Image Models Leave Identifiable Signatures: Implications for Leaderboard Security
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
Provably Cost-Sensitive Adversarial Defense via Randomized Smoothing
von: Xin, Yuan, et al.
Veröffentlicht: (2023)
von: Xin, Yuan, et al.
Veröffentlicht: (2023)
Mitigating the Backdoor Effect for Multi-Task Model Merging via Safety-Aware Subspace
von: Yang, Jinluan, et al.
Veröffentlicht: (2024)
von: Yang, Jinluan, et al.
Veröffentlicht: (2024)
Generated Data with Fake Privacy: Hidden Dangers of Fine-tuning Large Language Models on Generated Data
von: Akkus, Atilla, et al.
Veröffentlicht: (2024)
von: Akkus, Atilla, et al.
Veröffentlicht: (2024)
Revisiting Transferable Adversarial Images: Systemization, Evaluation, and New Insights
von: Zhao, Zhengyu, et al.
Veröffentlicht: (2023)
von: Zhao, Zhengyu, et al.
Veröffentlicht: (2023)
Peering Behind the Shield: Guardrail Identification in Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
von: Yang, Ziqing, et al.
Veröffentlicht: (2025)
A Channel-Triggered Backdoor Attack on Wireless Semantic Image Reconstruction
von: Wan, Jialin, et al.
Veröffentlicht: (2025)
von: Wan, Jialin, et al.
Veröffentlicht: (2025)
ShadowGenes: Leveraging Recurring Patterns within Computational Graphs for Model Genealogy
von: Schulz, Kasimir, et al.
Veröffentlicht: (2025)
von: Schulz, Kasimir, et al.
Veröffentlicht: (2025)
Reconstruct Your Previous Conversations! Comprehensively Investigating Privacy Leakage Risks in Conversations with GPT Models
von: Chu, Junjie, et al.
Veröffentlicht: (2024)
von: Chu, Junjie, et al.
Veröffentlicht: (2024)
Channel Prediction-Based Physical Layer Authentication under Consecutive Spoofing Attacks
von: Guo, Yijia, et al.
Veröffentlicht: (2026)
von: Guo, Yijia, et al.
Veröffentlicht: (2026)
Differentially Private SGD Without Clipping Bias: An Error-Feedback Approach
von: Zhang, Xinwei, et al.
Veröffentlicht: (2023)
von: Zhang, Xinwei, et al.
Veröffentlicht: (2023)
Neural Network-Powered Finger-Drawn Biometric Authentication
von: Balkhi, Maan Al, et al.
Veröffentlicht: (2025)
von: Balkhi, Maan Al, et al.
Veröffentlicht: (2025)
Open LLMs are Necessary for Current Private Adaptations and Outperform their Closed Alternatives
von: Hanke, Vincent, et al.
Veröffentlicht: (2024)
von: Hanke, Vincent, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Prompt Stealing Attacks Against Text-to-Image Generation Models
von: Shen, Xinyue, et al.
Veröffentlicht: (2023) -
On the Proactive Generation of Unsafe Images From Text-To-Image Models Using Benign Prompts
von: Wu, Yixin, et al.
Veröffentlicht: (2023) -
The Challenge of Identifying the Origin of Black-Box Large Language Models
von: Yang, Ziqing, et al.
Veröffentlicht: (2025) -
Voice Jailbreak Attacks Against GPT-4o
von: Shen, Xinyue, et al.
Veröffentlicht: (2024) -
Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications
von: Wu, Yixin, et al.
Veröffentlicht: (2025)