Identifying Models Behind Text-to-Image Leaderboards
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Naseh, Ali, Peng, Yuefeng, Suri, Anshuman, Chaudhari, Harsh, Oprea, Alina, Houmansadr, Amir |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Text-to-Image Models Leave Identifiable Signatures: Implications for Leaderboard Security
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
Exploiting Leaderboards for Large-Scale Distribution of Malicious Models
von: Suri, Anshuman, et al.
Veröffentlicht: (2025)
von: Suri, Anshuman, et al.
Veröffentlicht: (2025)
Riddle Me This! Stealthy Membership Inference for Retrieval-Augmented Generation
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
Diffence: Fencing Membership Privacy With Diffusion Models
von: Peng, Yuefeng, et al.
Veröffentlicht: (2023)
von: Peng, Yuefeng, et al.
Veröffentlicht: (2023)
R1dacted: Investigating Local Censorship in DeepSeek's R1 Language Model
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
von: Naseh, Ali, et al.
Veröffentlicht: (2025)
Iteratively Prompting Multimodal LLMs to Reproduce Natural and AI-Generated Images
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
VIDSTAMP: A Temporally-Aware Watermark for Ownership and Integrity in Video Diffusion Models
von: Teymoorianfard, Mohammadreza, et al.
Veröffentlicht: (2025)
von: Teymoorianfard, Mohammadreza, et al.
Veröffentlicht: (2025)
MeanSparse: Post-Training Robustness Enhancement Through Mean-Centered Feature Sparsification
von: Amini, Sajjad, et al.
Veröffentlicht: (2024)
von: Amini, Sajjad, et al.
Veröffentlicht: (2024)
Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem
von: Lin, Shuyi, et al.
Veröffentlicht: (2025)
von: Lin, Shuyi, et al.
Veröffentlicht: (2025)
Backdooring Bias ($B^2$) into Stable Diffusion Models
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
von: Naseh, Ali, et al.
Veröffentlicht: (2024)
SoK: Pitfalls in Evaluating Black-Box Attacks
von: Suya, Fnu, et al.
Veröffentlicht: (2023)
von: Suya, Fnu, et al.
Veröffentlicht: (2023)
OSLO: One-Shot Label-Only Membership Inference Attacks
von: Peng, Yuefeng, et al.
Veröffentlicht: (2024)
von: Peng, Yuefeng, et al.
Veröffentlicht: (2024)
Cascading Adversarial Bias from Injection to Distillation in Language Models
von: Chaudhari, Harsh, et al.
Veröffentlicht: (2025)
von: Chaudhari, Harsh, et al.
Veröffentlicht: (2025)
ArtistAuditor: Auditing Artist Style Pirate in Text-to-Image Generation Models
von: Du, Linkang, et al.
Veröffentlicht: (2025)
von: Du, Linkang, et al.
Veröffentlicht: (2025)
SAGA: A Security Architecture for Governing AI Agentic Systems
von: Syros, Georgios, et al.
Veröffentlicht: (2025)
von: Syros, Georgios, et al.
Veröffentlicht: (2025)
Attacking Autonomous Driving Agents with Adversarial Machine Learning: A Holistic Evaluation with the CARLA Leaderboard
von: Wong, Henry, et al.
Veröffentlicht: (2025)
von: Wong, Henry, et al.
Veröffentlicht: (2025)
Phantom: General Backdoor Attacks on Retrieval Augmented Language Generation
von: Chaudhari, Harsh, et al.
Veröffentlicht: (2024)
von: Chaudhari, Harsh, et al.
Veröffentlicht: (2024)
Throttling Web Agents Using Reasoning Gates
von: Kumar, Abhinav, et al.
Veröffentlicht: (2025)
von: Kumar, Abhinav, et al.
Veröffentlicht: (2025)
UTrace: Poisoning Forensics for Private Collaborative Learning
von: Rose, Evan, et al.
Veröffentlicht: (2024)
von: Rose, Evan, et al.
Veröffentlicht: (2024)
FameBias: Embedding Manipulation Bias Attack in Text-to-Image Models
von: Roh, Jaechul, et al.
Veröffentlicht: (2024)
von: Roh, Jaechul, et al.
Veröffentlicht: (2024)
Reconstruction of Personally Identifiable Information from Supervised Finetuned Models
von: Furukawa, Sae, et al.
Veröffentlicht: (2026)
von: Furukawa, Sae, et al.
Veröffentlicht: (2026)
OverThink: Slowdown Attacks on Reasoning LLMs
von: Kumar, Abhinav, et al.
Veröffentlicht: (2025)
von: Kumar, Abhinav, et al.
Veröffentlicht: (2025)
AI-generated Image Detection: Passive or Watermark?
von: Guo, Moyang, et al.
Veröffentlicht: (2024)
von: Guo, Moyang, et al.
Veröffentlicht: (2024)
DROP: Poison Dilution via Knowledge Distillation for Federated Learning
von: Syros, Georgios, et al.
Veröffentlicht: (2025)
von: Syros, Georgios, et al.
Veröffentlicht: (2025)
Thought-Transfer: Indirect Targeted Poisoning Attacks on Chain-of-Thought Reasoning Models
von: Chaudhari, Harsh, et al.
Veröffentlicht: (2026)
von: Chaudhari, Harsh, et al.
Veröffentlicht: (2026)
Understanding Implosion in Text-to-Image Generative Models
von: Ding, Wenxin, et al.
Veröffentlicht: (2024)
von: Ding, Wenxin, et al.
Veröffentlicht: (2024)
DIAGNOSIS: Detecting Unauthorized Data Usages in Text-to-image Diffusion Models
von: Wang, Zhenting, et al.
Veröffentlicht: (2023)
von: Wang, Zhenting, et al.
Veröffentlicht: (2023)
Diffusion Soup: Model Merging for Text-to-Image Diffusion Models
von: Biggs, Benjamin, et al.
Veröffentlicht: (2024)
von: Biggs, Benjamin, et al.
Veröffentlicht: (2024)
Identifying Physically Realizable Triggers for Backdoored Face Recognition Networks
von: Raj, Ankita, et al.
Veröffentlicht: (2025)
von: Raj, Ankita, et al.
Veröffentlicht: (2025)
Mitigating Sexual Content Generation via Embedding Distortion in Text-conditioned Diffusion Models
von: Ahn, Jaesin, et al.
Veröffentlicht: (2025)
von: Ahn, Jaesin, et al.
Veröffentlicht: (2025)
ToxicTextCLIP: Text-Based Poisoning and Backdoor Attacks on CLIP Pre-training
von: Yao, Xin, et al.
Veröffentlicht: (2025)
von: Yao, Xin, et al.
Veröffentlicht: (2025)
Accurate and Private Diagnosis of Rare Genetic Syndromes from Facial Images with Federated Deep Learning
von: Ünal, Ali Burak, et al.
Veröffentlicht: (2025)
von: Ünal, Ali Burak, et al.
Veröffentlicht: (2025)
Segmentation-free Connectionist Temporal Classification loss based OCR Model for Text Captcha Classification
von: Khatavkar, Vaibhav, et al.
Veröffentlicht: (2024)
von: Khatavkar, Vaibhav, et al.
Veröffentlicht: (2024)
Toward a Principled Framework for Agent Safety Measurement
von: Lin, Shuyi, et al.
Veröffentlicht: (2026)
von: Lin, Shuyi, et al.
Veröffentlicht: (2026)
PrivImage: Differentially Private Synthetic Image Generation using Diffusion Models with Semantic-Aware Pretraining
von: Li, Kecen, et al.
Veröffentlicht: (2023)
von: Li, Kecen, et al.
Veröffentlicht: (2023)
Red-Teaming Text-to-Image Systems by Rule-based Preference Modeling
von: Cao, Yichuan, et al.
Veröffentlicht: (2025)
von: Cao, Yichuan, et al.
Veröffentlicht: (2025)
©Plug-in Authorization for Human Content Copyright Protection in Text-to-Image Model
von: Zhou, Chao, et al.
Veröffentlicht: (2024)
von: Zhou, Chao, et al.
Veröffentlicht: (2024)
An Analysis of Recent Advances in Deepfake Image Detection in an Evolving Threat Landscape
von: Abdullah, Sifat Muhammad, et al.
Veröffentlicht: (2024)
von: Abdullah, Sifat Muhammad, et al.
Veröffentlicht: (2024)
Harnessing the Power of Large Vision Language Models for Synthetic Image Detection
von: Keita, Mamadou, et al.
Veröffentlicht: (2024)
von: Keita, Mamadou, et al.
Veröffentlicht: (2024)
Image Can Bring Your Memory Back: A Novel Multi-Modal Guided Attack against Image Generation Model Unlearning
von: Liu, Renyang, et al.
Veröffentlicht: (2025)
von: Liu, Renyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Text-to-Image Models Leave Identifiable Signatures: Implications for Leaderboard Security
von: Naseh, Ali, et al.
Veröffentlicht: (2025) -
Exploiting Leaderboards for Large-Scale Distribution of Malicious Models
von: Suri, Anshuman, et al.
Veröffentlicht: (2025) -
Riddle Me This! Stealthy Membership Inference for Retrieval-Augmented Generation
von: Naseh, Ali, et al.
Veröffentlicht: (2025) -
Diffence: Fencing Membership Privacy With Diffusion Models
von: Peng, Yuefeng, et al.
Veröffentlicht: (2023) -
R1dacted: Investigating Local Censorship in DeepSeek's R1 Language Model
von: Naseh, Ali, et al.
Veröffentlicht: (2025)