SafeScreen: A Safety-First Screening Framework for Personalized Video Retrieval for Vulnerable Users
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhao, Wenzheng, Gadiputi, Madhava Kalyan, Yuan, Fengpei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Wukong Framework for Not Safe For Work Detection in Text-to-Image systems
por: Liu, Mingrui, et al.
Publicado: (2025)
por: Liu, Mingrui, et al.
Publicado: (2025)
Beyond Vulnerabilities: A Survey of Adversarial Attacks as Both Threats and Defenses in Computer Vision Systems
por: Guo, Zhongliang, et al.
Publicado: (2025)
por: Guo, Zhongliang, et al.
Publicado: (2025)
SafeGuider: Robust and Practical Content Safety Control for Text-to-Image Models
por: Qi, Peigui, et al.
Publicado: (2025)
por: Qi, Peigui, et al.
Publicado: (2025)
GuardTrace-VL: Detecting Unsafe Multimodel Reasoning via Iterative Safety Supervision
por: Xiang, Yuxiao, et al.
Publicado: (2025)
por: Xiang, Yuxiao, et al.
Publicado: (2025)
SteerDiff: Steering towards Safe Text-to-Image Diffusion Models
por: Zhang, Hongxiang, et al.
Publicado: (2024)
por: Zhang, Hongxiang, et al.
Publicado: (2024)
SafeVision: Efficient Image Guardrail with Robust Policy Adherence and Explainability
por: Xu, Peiyang, et al.
Publicado: (2025)
por: Xu, Peiyang, et al.
Publicado: (2025)
The Structural Safety Generalization Problem
por: Broomfield, Julius, et al.
Publicado: (2025)
por: Broomfield, Julius, et al.
Publicado: (2025)
Beyond the Safety Tax: Mitigating Unsafe Text-to-Image Generation via External Safety Rectification
por: Meng, Xiangtao, et al.
Publicado: (2025)
por: Meng, Xiangtao, et al.
Publicado: (2025)
X-SG$^2$S: Safe and Generalizable Gaussian Splatting with X-dimensional Watermarks
por: Cheng, Zihang, et al.
Publicado: (2025)
por: Cheng, Zihang, et al.
Publicado: (2025)
VideoEraser: Concept Erasure in Text-to-Video Diffusion Models
por: Xu, Naen, et al.
Publicado: (2025)
por: Xu, Naen, et al.
Publicado: (2025)
CPR: Retrieval Augmented Generation for Copyright Protection
por: Golatkar, Aditya, et al.
Publicado: (2024)
por: Golatkar, Aditya, et al.
Publicado: (2024)
HomeSafe-Bench: Evaluating Vision-Language Models on Unsafe Action Detection for Embodied Agents in Household Scenarios
por: Pu, Jiayue, et al.
Publicado: (2026)
por: Pu, Jiayue, et al.
Publicado: (2026)
Rethinking and Red-Teaming Protective Perturbation in Personalized Diffusion Models
por: Liu, Yixin, et al.
Publicado: (2024)
por: Liu, Yixin, et al.
Publicado: (2024)
IdentityGuard: Context-Aware Restriction and Provenance for Personalized Synthesis
por: Zhang, Lingyun, et al.
Publicado: (2026)
por: Zhang, Lingyun, et al.
Publicado: (2026)
An Experimental Study of Trojan Vulnerabilities in UAV Autonomous Landing
por: Ahmari, Reza, et al.
Publicado: (2025)
por: Ahmari, Reza, et al.
Publicado: (2025)
Watertox: The Art of Simplicity in Universal Attacks A Cross-Model Framework for Robust Adversarial Generation
por: Gao, Zhenghao, et al.
Publicado: (2024)
por: Gao, Zhenghao, et al.
Publicado: (2024)
VLM-Guard: Safeguarding Vision-Language Models via Fulfilling Safety Alignment Gap
por: Liu, Qin, et al.
Publicado: (2025)
por: Liu, Qin, et al.
Publicado: (2025)
Towards Dataset Copyright Evasion Attack against Personalized Text-to-Image Diffusion Models
por: Gao, Kuofeng, et al.
Publicado: (2025)
por: Gao, Kuofeng, et al.
Publicado: (2025)
When Understanding Becomes a Risk: Authenticity and Safety Risks in the Emerging Image Generation Paradigm
por: Leng, Ye, et al.
Publicado: (2026)
por: Leng, Ye, et al.
Publicado: (2026)
T2I-RiskyPrompt: A Benchmark for Safety Evaluation, Attack, and Defense on Text-to-Image Model
por: Zhang, Chenyu, et al.
Publicado: (2025)
por: Zhang, Chenyu, et al.
Publicado: (2025)
Physically Realizable Natural-Looking Clothing Textures Evade Person Detectors via 3D Modeling
por: Hu, Zhanhao, et al.
Publicado: (2023)
por: Hu, Zhanhao, et al.
Publicado: (2023)
Value-Aligned Prompt Moderation via Zero-Shot Agentic Rewriting for Safe Image Generation
por: Zhao, Xin, et al.
Publicado: (2025)
por: Zhao, Xin, et al.
Publicado: (2025)
Combating Falsification of Speech Videos with Live Optical Signatures (Extended Version)
por: Schwartz, Hadleigh, et al.
Publicado: (2025)
por: Schwartz, Hadleigh, et al.
Publicado: (2025)
GOTCHA: Real-Time Video Deepfake Detection via Challenge-Response
por: Mittal, Govind, et al.
Publicado: (2022)
por: Mittal, Govind, et al.
Publicado: (2022)
PII-VisBench: Evaluating Personally Identifiable Information Safety in Vision Language Models Along a Continuum of Visibility
por: Shahariar, G M, et al.
Publicado: (2026)
por: Shahariar, G M, et al.
Publicado: (2026)
Unveiling the Potential: Harnessing Deep Metric Learning to Circumvent Video Streaming Encryption
por: Gansekoele, Arwin, et al.
Publicado: (2024)
por: Gansekoele, Arwin, et al.
Publicado: (2024)
Shaking the Fake: Detecting Deepfake Videos in Real Time via Active Probes
por: Xie, Zhixin, et al.
Publicado: (2024)
por: Xie, Zhixin, et al.
Publicado: (2024)
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
por: Li, Xiao, et al.
Publicado: (2026)
por: Li, Xiao, et al.
Publicado: (2026)
Vulnerability analysis of captcha using Deep learning
por: Walia, Jaskaran Singh, et al.
Publicado: (2023)
por: Walia, Jaskaran Singh, et al.
Publicado: (2023)
SKeDA: A Generative Watermarking Framework for Text-to-video Diffusion Models
por: Yang, Yang, et al.
Publicado: (2026)
por: Yang, Yang, et al.
Publicado: (2026)
Representation Magnitude has a Liability to Privacy Vulnerability
por: Fang, Xingli, et al.
Publicado: (2024)
por: Fang, Xingli, et al.
Publicado: (2024)
Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety
por: Ma, Xingjun, et al.
Publicado: (2025)
por: Ma, Xingjun, et al.
Publicado: (2025)
Architectural Neural Backdoors from First Principles
por: Langford, Harry, et al.
Publicado: (2024)
por: Langford, Harry, et al.
Publicado: (2024)
FedPalm: A General Federated Learning Framework for Closed- and Open-Set Palmprint Verification
por: Yang, Ziyuan, et al.
Publicado: (2025)
por: Yang, Ziyuan, et al.
Publicado: (2025)
Improving Adversarial Training using Vulnerability-Aware Perturbation Budget
por: Fakorede, Olukorede, et al.
Publicado: (2024)
por: Fakorede, Olukorede, et al.
Publicado: (2024)
CipherDM: Secure Three-Party Inference for Diffusion Model Sampling
por: Zhao, Xin, et al.
Publicado: (2024)
por: Zhao, Xin, et al.
Publicado: (2024)
Refusing Safe Prompts for Multi-modal Large Language Models
por: Shao, Zedian, et al.
Publicado: (2024)
por: Shao, Zedian, et al.
Publicado: (2024)
Region-Guided Attack on the Segment Anything Model (SAM)
por: Liu, Xiaoliang, et al.
Publicado: (2024)
por: Liu, Xiaoliang, et al.
Publicado: (2024)
Antelope: Potent and Concealed Jailbreak Attack Strategy
por: Zhao, Xin, et al.
Publicado: (2024)
por: Zhao, Xin, et al.
Publicado: (2024)
Spot Risks Before Speaking! Unraveling Safety Attention Heads in Large Vision-Language Models
por: Zheng, Ziwei, et al.
Publicado: (2025)
por: Zheng, Ziwei, et al.
Publicado: (2025)
Ejemplares similares
-
Wukong Framework for Not Safe For Work Detection in Text-to-Image systems
por: Liu, Mingrui, et al.
Publicado: (2025) -
Beyond Vulnerabilities: A Survey of Adversarial Attacks as Both Threats and Defenses in Computer Vision Systems
por: Guo, Zhongliang, et al.
Publicado: (2025) -
SafeGuider: Robust and Practical Content Safety Control for Text-to-Image Models
por: Qi, Peigui, et al.
Publicado: (2025) -
GuardTrace-VL: Detecting Unsafe Multimodel Reasoning via Iterative Safety Supervision
por: Xiang, Yuxiao, et al.
Publicado: (2025) -
SteerDiff: Steering towards Safe Text-to-Image Diffusion Models
por: Zhang, Hongxiang, et al.
Publicado: (2024)