Towards Understanding Unsafe Video Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Pang, Yan, Xiong, Aiping, Zhang, Yang, Wang, Tianhao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VGMShield: Mitigating Misuse of Video Generative Models
por: Pang, Yan, et al.
Publicado: (2024)
por: Pang, Yan, et al.
Publicado: (2024)
SAGE: Exploring the Boundaries of Unsafe Concept Domain with Semantic-Augment Erasing
por: Zhu, Hongguang, et al.
Publicado: (2025)
por: Zhu, Hongguang, et al.
Publicado: (2025)
Beyond the Safety Tax: Mitigating Unsafe Text-to-Image Generation via External Safety Rectification
por: Meng, Xiangtao, et al.
Publicado: (2025)
por: Meng, Xiangtao, et al.
Publicado: (2025)
PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models
por: Yuan, Lingzhi, et al.
Publicado: (2025)
por: Yuan, Lingzhi, et al.
Publicado: (2025)
HomeSafe-Bench: Evaluating Vision-Language Models on Unsafe Action Detection for Embodied Agents in Household Scenarios
por: Pu, Jiayue, et al.
Publicado: (2026)
por: Pu, Jiayue, et al.
Publicado: (2026)
GuardTrace-VL: Detecting Unsafe Multimodel Reasoning via Iterative Safety Supervision
por: Xiang, Yuxiao, et al.
Publicado: (2025)
por: Xiang, Yuxiao, et al.
Publicado: (2025)
GUIGuard-Bench: Toward a General Evaluation for Privacy-Preserving GUI Agents
por: Wang, Yanxi, et al.
Publicado: (2026)
por: Wang, Yanxi, et al.
Publicado: (2026)
When Understanding Becomes a Risk: Authenticity and Safety Risks in the Emerging Image Generation Paradigm
por: Leng, Ye, et al.
Publicado: (2026)
por: Leng, Ye, et al.
Publicado: (2026)
Rethinking Data Protection in the (Generative) Artificial Intelligence Era
por: Li, Yiming, et al.
Publicado: (2025)
por: Li, Yiming, et al.
Publicado: (2025)
The Adversarial AI-Art: Understanding, Generation, Detection, and Benchmarking
por: Li, Yuying, et al.
Publicado: (2024)
por: Li, Yuying, et al.
Publicado: (2024)
CGI-DM: Digital Copyright Authentication for Diffusion Models via Contrasting Gradient Inversion
por: Wu, Xiaoyu, et al.
Publicado: (2024)
por: Wu, Xiaoyu, et al.
Publicado: (2024)
Combating Falsification of Speech Videos with Live Optical Signatures (Extended Version)
por: Schwartz, Hadleigh, et al.
Publicado: (2025)
por: Schwartz, Hadleigh, et al.
Publicado: (2025)
GenAI Confessions: Black-box Membership Inference for Generative Image Models
por: Bohacek, Matyas, et al.
Publicado: (2025)
por: Bohacek, Matyas, et al.
Publicado: (2025)
VideoEraser: Concept Erasure in Text-to-Video Diffusion Models
por: Xu, Naen, et al.
Publicado: (2025)
por: Xu, Naen, et al.
Publicado: (2025)
CAPTCHA Solving for Native GUI Agents: Automated Reasoning-Action Data Generation and Self-Corrective Training
por: Chen, Yuxi, et al.
Publicado: (2026)
por: Chen, Yuxi, et al.
Publicado: (2026)
Adversarial Nibbler: An Open Red-Teaming Method for Identifying Diverse Harms in Text-to-Image Generation
por: Quaye, Jessica, et al.
Publicado: (2024)
por: Quaye, Jessica, et al.
Publicado: (2024)
PromptLA: Towards Integrity Verification of Black-box Text-to-Image Diffusion Models
por: Zhang, Zhuomeng, et al.
Publicado: (2024)
por: Zhang, Zhuomeng, et al.
Publicado: (2024)
Towards Understanding Dual BN In Hybrid Adversarial Training
por: Zhang, Chenshuang, et al.
Publicado: (2024)
por: Zhang, Chenshuang, et al.
Publicado: (2024)
SKeDA: A Generative Watermarking Framework for Text-to-video Diffusion Models
por: Yang, Yang, et al.
Publicado: (2026)
por: Yang, Yang, et al.
Publicado: (2026)
FraudBench: A Multimodal Benchmark for Detecting AI-Generated Fraudulent Refund Evidence
por: Yan, Xinyu, et al.
Publicado: (2026)
por: Yan, Xinyu, et al.
Publicado: (2026)
FedPalm: A General Federated Learning Framework for Closed- and Open-Set Palmprint Verification
por: Yang, Ziyuan, et al.
Publicado: (2025)
por: Yang, Ziyuan, et al.
Publicado: (2025)
Towards Dataset Copyright Evasion Attack against Personalized Text-to-Image Diffusion Models
por: Gao, Kuofeng, et al.
Publicado: (2025)
por: Gao, Kuofeng, et al.
Publicado: (2025)
Training-Free Watermarking for Autoregressive Image Generation
por: Tong, Yu, et al.
Publicado: (2025)
por: Tong, Yu, et al.
Publicado: (2025)
DREAM: Scalable Red Teaming for Text-to-Image Generative Systems via Distribution Modeling
por: Li, Boheng, et al.
Publicado: (2025)
por: Li, Boheng, et al.
Publicado: (2025)
Invisible Image Watermarks Are Provably Removable Using Generative AI
por: Zhao, Xuandong, et al.
Publicado: (2023)
por: Zhao, Xuandong, et al.
Publicado: (2023)
Beyond Known Fakes: Generalized Detection of AI-Generated Images via Post-hoc Distribution Alignment
por: Wang, Li, et al.
Publicado: (2025)
por: Wang, Li, et al.
Publicado: (2025)
GOTCHA: Real-Time Video Deepfake Detection via Challenge-Response
por: Mittal, Govind, et al.
Publicado: (2022)
por: Mittal, Govind, et al.
Publicado: (2022)
Unveiling the Potential: Harnessing Deep Metric Learning to Circumvent Video Streaming Encryption
por: Gansekoele, Arwin, et al.
Publicado: (2024)
por: Gansekoele, Arwin, et al.
Publicado: (2024)
Shaking the Fake: Detecting Deepfake Videos in Real Time via Active Probes
por: Xie, Zhixin, et al.
Publicado: (2024)
por: Xie, Zhixin, et al.
Publicado: (2024)
The Structural Safety Generalization Problem
por: Broomfield, Julius, et al.
Publicado: (2025)
por: Broomfield, Julius, et al.
Publicado: (2025)
CPR: Retrieval Augmented Generation for Copyright Protection
por: Golatkar, Aditya, et al.
Publicado: (2024)
por: Golatkar, Aditya, et al.
Publicado: (2024)
Towards Context-Aware Image Anonymization with Multi-Agent Reasoning
por: Aufschläger, Robert, et al.
Publicado: (2026)
por: Aufschläger, Robert, et al.
Publicado: (2026)
SafeGen: Mitigating Sexually Explicit Content Generation in Text-to-Image Models
por: Li, Xinfeng, et al.
Publicado: (2024)
por: Li, Xinfeng, et al.
Publicado: (2024)
SafeScreen: A Safety-First Screening Framework for Personalized Video Retrieval for Vulnerable Users
por: Zhao, Wenzheng, et al.
Publicado: (2026)
por: Zhao, Wenzheng, et al.
Publicado: (2026)
Towards A Correct Usage of Cryptography in Semantic Watermarks for Diffusion Models
por: Thietke, Jonas, et al.
Publicado: (2025)
por: Thietke, Jonas, et al.
Publicado: (2025)
Navigation as Attackers Wish? Towards Building Robust Embodied Agents under Federated Learning
por: Zhang, Yunchao, et al.
Publicado: (2022)
por: Zhang, Yunchao, et al.
Publicado: (2022)
ICAS: Detecting Training Data from Autoregressive Image Generative Models
por: Yu, Hongyao, et al.
Publicado: (2025)
por: Yu, Hongyao, et al.
Publicado: (2025)
CatchBackdoor: Backdoor Detection via Critical Trojan Neural Path Fuzzing
por: Jin, Haibo, et al.
Publicado: (2021)
por: Jin, Haibo, et al.
Publicado: (2021)
Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
por: Glukhov, David, et al.
Publicado: (2024)
por: Glukhov, David, et al.
Publicado: (2024)
Imitation Game for Adversarial Disillusion with Chain-of-Thought Reasoning in Generative AI
por: Chang, Ching-Chun, et al.
Publicado: (2025)
por: Chang, Ching-Chun, et al.
Publicado: (2025)
Ejemplares similares
-
VGMShield: Mitigating Misuse of Video Generative Models
por: Pang, Yan, et al.
Publicado: (2024) -
SAGE: Exploring the Boundaries of Unsafe Concept Domain with Semantic-Augment Erasing
por: Zhu, Hongguang, et al.
Publicado: (2025) -
Beyond the Safety Tax: Mitigating Unsafe Text-to-Image Generation via External Safety Rectification
por: Meng, Xiangtao, et al.
Publicado: (2025) -
PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models
por: Yuan, Lingzhi, et al.
Publicado: (2025) -
HomeSafe-Bench: Evaluating Vision-Language Models on Unsafe Action Detection for Embodied Agents in Household Scenarios
por: Pu, Jiayue, et al.
Publicado: (2026)