SafeWatch: An Efficient Safety-Policy Following Video Guardrail Model with Transparent Explanations
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Zhaorun, Pinto, Francesco, Pan, Minzhou, Li, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SafeVision: Efficient Image Guardrail with Robust Policy Adherence and Explainability
by: Xu, Peiyang, et al.
Published: (2025)
by: Xu, Peiyang, et al.
Published: (2025)
Beyond Training: Dynamic Token Merging for Zero-Shot Video Understanding
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
VERITAS: Verification and Explanation of Realness in Images for Transparency in AI Systems
by: Srivastava, Aadi, et al.
Published: (2025)
by: Srivastava, Aadi, et al.
Published: (2025)
VMDT: Decoding the Trustworthiness of Video Foundation Models
by: Potter, Yujin, et al.
Published: (2025)
by: Potter, Yujin, et al.
Published: (2025)
Show, Don't Tell: Detecting Novel Objects by Watching Human Videos
by: Akl, James, et al.
Published: (2026)
by: Akl, James, et al.
Published: (2026)
Breaking Free: How to Hack Safety Guardrails in Black-Box Diffusion Models!
by: Kotyan, Shashank, et al.
Published: (2024)
by: Kotyan, Shashank, et al.
Published: (2024)
HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding
by: Chen, Zhaorun, et al.
Published: (2024)
by: Chen, Zhaorun, et al.
Published: (2024)
GRAPE: Generalizing Robot Policy via Preference Alignment
by: Zhang, Zijian, et al.
Published: (2024)
by: Zhang, Zijian, et al.
Published: (2024)
LD-ViCE: Latent Diffusion Model for Video Counterfactual Explanations
by: Varshney, Payal, et al.
Published: (2025)
by: Varshney, Payal, et al.
Published: (2025)
SafeAug: Safety-Critical Driving Data Augmentation from Naturalistic Datasets
by: Mo, Zhaobin, et al.
Published: (2025)
by: Mo, Zhaobin, et al.
Published: (2025)
Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding
by: Fang, Yixiong, et al.
Published: (2024)
by: Fang, Yixiong, et al.
Published: (2024)
SafeLens: Deliberate and Efficient Video Guardrails with Fast-and-Slow Screening
by: Nahin, Shahriar Kabir, et al.
Published: (2026)
by: Nahin, Shahriar Kabir, et al.
Published: (2026)
MEGL: Multimodal Explanation-Guided Learning
by: Zhang, Yifei, et al.
Published: (2024)
by: Zhang, Yifei, et al.
Published: (2024)
Right Predictions, Misleading Explanations: On the Vulnerability of Vision-Language Model Explanations
by: Babadi, Narges, et al.
Published: (2026)
by: Babadi, Narges, et al.
Published: (2026)
UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models
by: Wang, Jiaqi, et al.
Published: (2026)
by: Wang, Jiaqi, et al.
Published: (2026)
Open-Source Acceleration of Stable-Diffusion.cpp Deployable on All Devices
by: Ng, Jingxu, et al.
Published: (2024)
by: Ng, Jingxu, et al.
Published: (2024)
Pyramidal Flow Matching for Efficient Video Generative Modeling
by: Jin, Yang, et al.
Published: (2024)
by: Jin, Yang, et al.
Published: (2024)
LightCache: Memory-Efficient, Training-Free Acceleration for Video Generation
by: Xiao, Yang, et al.
Published: (2025)
by: Xiao, Yang, et al.
Published: (2025)
Enhancing Multimodal Large Language Models for Safety-Critical Driving Video Analysis
by: Trinci, Tomaso, et al.
Published: (2026)
by: Trinci, Tomaso, et al.
Published: (2026)
Studying How to Efficiently and Effectively Guide Models with Explanations
by: Rao, Sukrut, et al.
Published: (2023)
by: Rao, Sukrut, et al.
Published: (2023)
Safe and Reliable Diffusion Models via Subspace Projection
by: Chen, Huiqiang, et al.
Published: (2025)
by: Chen, Huiqiang, et al.
Published: (2025)
Rethinking JEPA: Compute-Efficient Video SSL with Frozen Teachers
by: Li, Xianhang, et al.
Published: (2025)
by: Li, Xianhang, et al.
Published: (2025)
Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
by: He, Haoran, et al.
Published: (2024)
by: He, Haoran, et al.
Published: (2024)
Learning by Watching: A Review of Video-based Learning Approaches for Robot Manipulation
by: Eze, Chrisantus, et al.
Published: (2024)
by: Eze, Chrisantus, et al.
Published: (2024)
AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security
by: Liu, Dongrui, et al.
Published: (2026)
by: Liu, Dongrui, et al.
Published: (2026)
iWatchRoad: Scalable Detection and Geospatial Visualization of Potholes for Smart Cities
by: Sahoo, Rishi Raj, et al.
Published: (2025)
by: Sahoo, Rishi Raj, et al.
Published: (2025)
How to Squeeze An Explanation Out of Your Model
by: Roxo, Tiago, et al.
Published: (2024)
by: Roxo, Tiago, et al.
Published: (2024)
Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition
by: Yu, Sihyun, et al.
Published: (2024)
by: Yu, Sihyun, et al.
Published: (2024)
Learning an Efficient Optimizer via Hybrid-Policy Sub-Trajectory Balance
by: Guan, Yunchuan, et al.
Published: (2025)
by: Guan, Yunchuan, et al.
Published: (2025)
iWatchRoadv2: Pothole Detection, Geospatial Mapping, and Intelligent Road Governance
by: Sahoo, Rishi Raj, et al.
Published: (2025)
by: Sahoo, Rishi Raj, et al.
Published: (2025)
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model
by: Wang, Xiyao, et al.
Published: (2025)
by: Wang, Xiyao, et al.
Published: (2025)
MonarchRT: Efficient Attention for Real-Time Video Generation
by: Agarwal, Krish, et al.
Published: (2026)
by: Agarwal, Krish, et al.
Published: (2026)
VidGuard-R1: AI-Generated Video Detection and Explanation via Reasoning MLLMs and RL
by: Park, Kyoungjun, et al.
Published: (2025)
by: Park, Kyoungjun, et al.
Published: (2025)
Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection
by: Zhang, Shuhai, et al.
Published: (2025)
by: Zhang, Shuhai, et al.
Published: (2025)
Video to Video Generative Adversarial Network for Few-shot Learning Based on Policy Gradient
by: Ma, Yintai, et al.
Published: (2024)
by: Ma, Yintai, et al.
Published: (2024)
Towards a Transparent and Interpretable AI Model for Medical Image Classifications
by: Wen, Binbin, et al.
Published: (2025)
by: Wen, Binbin, et al.
Published: (2025)
RankCLIP: Ranking-Consistent Language-Image Pretraining
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
Architecture-Aware Explanation Auditing for Industrial Visual Inspection
by: Jia, Sibo, et al.
Published: (2026)
by: Jia, Sibo, et al.
Published: (2026)
Global Renewables Watch: A Temporal Dataset of Solar and Wind Energy Derived from Satellite Imagery
by: Robinson, Caleb, et al.
Published: (2025)
by: Robinson, Caleb, et al.
Published: (2025)
On the Generalization and Causal Explanation in Self-Supervised Learning
by: Qiang, Wenwen, et al.
Published: (2024)
by: Qiang, Wenwen, et al.
Published: (2024)
Similar Items
-
SafeVision: Efficient Image Guardrail with Robust Policy Adherence and Explainability
by: Xu, Peiyang, et al.
Published: (2025) -
Beyond Training: Dynamic Token Merging for Zero-Shot Video Understanding
by: Zhang, Yiming, et al.
Published: (2024) -
VERITAS: Verification and Explanation of Realness in Images for Transparency in AI Systems
by: Srivastava, Aadi, et al.
Published: (2025) -
VMDT: Decoding the Trustworthiness of Video Foundation Models
by: Potter, Yujin, et al.
Published: (2025) -
Show, Don't Tell: Detecting Novel Objects by Watching Human Videos
by: Akl, James, et al.
Published: (2026)