PolyJuice Makes It Real: Black-Box, Universal Red Teaming for Synthetic Image Detectors
Fuente:
arXiv
Salvato in:
| Autori principali: | Dehdashtian, Sepehr, Morshed, Mashrur M., Seidman, Jacob H., Bharaj, Gaurav, Boddeti, Vishnu Naresh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Utility-Fairness Trade-Offs and How to Find Them
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024)
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024)
FairerCLIP: Debiasing CLIP's Zero-Shot Predictions using Functions in RKHSs
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024)
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024)
OASIS Uncovers: High-Quality T2I Models, Same Old Stereotypes
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2025)
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2025)
DiverseFlow: Sample-Efficient Diverse Mode Coverage in Flows
di: Morshed, Mashrur M., et al.
Pubblicazione: (2025)
di: Morshed, Mashrur M., et al.
Pubblicazione: (2025)
The Dark Side of Dataset Scaling: Evaluating Racial Classification in Multimodal Models
di: Birhane, Abeba, et al.
Pubblicazione: (2024)
di: Birhane, Abeba, et al.
Pubblicazione: (2024)
Fairness and Bias Mitigation in Computer Vision: A Survey
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024)
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024)
Incorporating Interventional Independence Improves Robustness against Interventional Distribution Shift
di: Sreekumar, Gautam, et al.
Pubblicazione: (2025)
di: Sreekumar, Gautam, et al.
Pubblicazione: (2025)
CryptoFace: End-to-End Encrypted Face Recognition
di: Ao, Wei, et al.
Pubblicazione: (2025)
di: Ao, Wei, et al.
Pubblicazione: (2025)
Post-hoc Selective Classification for Reliable Synthetic Image Detection
di: Zheng, Kaixiang, et al.
Pubblicazione: (2026)
di: Zheng, Kaixiang, et al.
Pubblicazione: (2026)
Obliviator Reveals the Cost of Nonlinear Guardedness in Concept Erasure
di: Akbari, Ramin, et al.
Pubblicazione: (2026)
di: Akbari, Ramin, et al.
Pubblicazione: (2026)
Compositional World Knowledge leads to High Utility Synthetic data
di: Gaudi, Sachit, et al.
Pubblicazione: (2025)
di: Gaudi, Sachit, et al.
Pubblicazione: (2025)
What Does an Audio Deepfake Detector Focus on? A Study in the Time Domain
di: Grinberg, Petr, et al.
Pubblicazione: (2025)
di: Grinberg, Petr, et al.
Pubblicazione: (2025)
X-Edit: Detecting and Localizing Edits in Images Altered by Text-Guided Diffusion Models
di: Bazyleva, Valentina, et al.
Pubblicazione: (2025)
di: Bazyleva, Valentina, et al.
Pubblicazione: (2025)
Estimating Parameter Fields in Multi-Physics PDEs from Scarce Measurements
di: Li, Xuyang, et al.
Pubblicazione: (2025)
di: Li, Xuyang, et al.
Pubblicazione: (2025)
FaceLift: Semi-supervised 3D Facial Landmark Localization
di: Ferman, David, et al.
Pubblicazione: (2024)
di: Ferman, David, et al.
Pubblicazione: (2024)
Content and Style Aware Audio-Driven Facial Animation
di: Liu, Qingju, et al.
Pubblicazione: (2024)
di: Liu, Qingju, et al.
Pubblicazione: (2024)
A Deep Learning Framework for Three Dimensional Shape Reconstruction from Phaseless Acoustic Scattering Far-field Data
di: Dikbayir, Doga, et al.
Pubblicazione: (2024)
di: Dikbayir, Doga, et al.
Pubblicazione: (2024)
SEAL: Semantic Attention Learning for Long Video Representation
di: Wang, Lan, et al.
Pubblicazione: (2024)
di: Wang, Lan, et al.
Pubblicazione: (2024)
Show Me: Unifying Instructional Image and Video Generation with Diffusion Models
di: Pu, Yujiang, et al.
Pubblicazione: (2025)
di: Pu, Yujiang, et al.
Pubblicazione: (2025)
CoInD: Enabling Logical Compositions in Diffusion Models
di: Gaudi, Sachit, et al.
Pubblicazione: (2025)
di: Gaudi, Sachit, et al.
Pubblicazione: (2025)
Action Reimagined: Text-to-Pose Video Editing for Dynamic Human Actions
di: Wang, Lan, et al.
Pubblicazione: (2024)
di: Wang, Lan, et al.
Pubblicazione: (2024)
A Deep Learning-based Multimodal Depth-Aware Dynamic Hand Gesture Recognition System
di: Mahmud, Hasan, et al.
Pubblicazione: (2021)
di: Mahmud, Hasan, et al.
Pubblicazione: (2021)
Mechanics-Informed Autoencoder Enables Automated Detection and Localization of Unforeseen Structural Damage
di: Li, Xuyang, et al.
Pubblicazione: (2024)
di: Li, Xuyang, et al.
Pubblicazione: (2024)
AutoODD: Agentic Audits via Bayesian Red Teaming in Black-Box Models
di: Martin, Rebecca, et al.
Pubblicazione: (2025)
di: Martin, Rebecca, et al.
Pubblicazione: (2025)
Learn from Real: Reality Defender's Submission to ASVspoof5 Challenge
di: Zhu, Yi, et al.
Pubblicazione: (2024)
di: Zhu, Yi, et al.
Pubblicazione: (2024)
SLIM: Style-Linguistics Mismatch Model for Generalized Audio Deepfake Detection
di: Zhu, Yi, et al.
Pubblicazione: (2024)
di: Zhu, Yi, et al.
Pubblicazione: (2024)
A Data-Driven Diffusion-based Approach for Audio Deepfake Explanations
di: Grinberg, Petr, et al.
Pubblicazione: (2025)
di: Grinberg, Petr, et al.
Pubblicazione: (2025)
Does It Make Sense to Explain a Black Box With Another Black Box?
di: Delaunay, Julien, et al.
Pubblicazione: (2024)
di: Delaunay, Julien, et al.
Pubblicazione: (2024)
Common Sense Reasoning for Deepfake Detection
di: Zhang, Yue, et al.
Pubblicazione: (2024)
di: Zhang, Yue, et al.
Pubblicazione: (2024)
Towards Attention-based Contrastive Learning for Audio Spoof Detection
di: Goel, Chirag, et al.
Pubblicazione: (2024)
di: Goel, Chirag, et al.
Pubblicazione: (2024)
Common-Sense Bias Modeling for Classification Tasks
di: Zhang, Miao, et al.
Pubblicazione: (2024)
di: Zhang, Miao, et al.
Pubblicazione: (2024)
New Zealand's industrial relations act, 1973
di: Joel Seidman
Pubblicazione: (1974)
di: Joel Seidman
Pubblicazione: (1974)
Nueva Zelandia: ley de relaciones de trabajo en 1973
di: Joel Seidman
Pubblicazione: (1974)
di: Joel Seidman
Pubblicazione: (1974)
La loi néo-zélandaise de 1973 sur les relations professionnelles
di: Joel Seidman
Pubblicazione: (1974)
di: Joel Seidman
Pubblicazione: (1974)
CORVUS: Red-Teaming Hallucination Detectors via Internal Signal Camouflage in Large Language Models
di: Min, Nay Myat, et al.
Pubblicazione: (2026)
di: Min, Nay Myat, et al.
Pubblicazione: (2026)
FDLLM: A Dedicated Detector for Black-Box LLMs Fingerprinting
di: Fu, Zhiyuan, et al.
Pubblicazione: (2025)
di: Fu, Zhiyuan, et al.
Pubblicazione: (2025)
Red Teaming AI Red Teaming
di: Majumdar, Subhabrata, et al.
Pubblicazione: (2025)
di: Majumdar, Subhabrata, et al.
Pubblicazione: (2025)
AutoRedTrader: Autonomous Red Teaming of Trading Agents through Synthetic Misinformation Injection
di: Liu, Zhiwei, et al.
Pubblicazione: (2026)
di: Liu, Zhiwei, et al.
Pubblicazione: (2026)
An Evaluation of Explanation Methods for Black-Box Detectors of Machine-Generated Text
di: Schoenegger, Loris, et al.
Pubblicazione: (2024)
di: Schoenegger, Loris, et al.
Pubblicazione: (2024)
Enhancing Privacy in Face Analytics Using Fully Homomorphic Encryption
di: Yalavarthi, Bharat, et al.
Pubblicazione: (2024)
di: Yalavarthi, Bharat, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Utility-Fairness Trade-Offs and How to Find Them
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024) -
FairerCLIP: Debiasing CLIP's Zero-Shot Predictions using Functions in RKHSs
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2024) -
OASIS Uncovers: High-Quality T2I Models, Same Old Stereotypes
di: Dehdashtian, Sepehr, et al.
Pubblicazione: (2025) -
DiverseFlow: Sample-Efficient Diverse Mode Coverage in Flows
di: Morshed, Mashrur M., et al.
Pubblicazione: (2025) -
The Dark Side of Dataset Scaling: Evaluating Racial Classification in Multimodal Models
di: Birhane, Abeba, et al.
Pubblicazione: (2024)