DeepKnown-Guard: A Proprietary Model-Based Safety Response Framework for AI Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Qi, Xu, Jianjun, Wei, Pingtao, Li, Jiu, Zhao, Peiqiang, Shi, Jiwei, Zhang, Xuan, Yang, Yanhui, Hui, Xiaodong, Xu, Peng, Shao, Wenqin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Comparative Evaluation of AI Agent Security Guardrails
by: Li, Qi, et al.
Published: (2026)
by: Li, Qi, et al.
Published: (2026)
Lemma on logarithmic derivative over directed manifolds
by: Lin, Peiqiang
Published: (2025)
by: Lin, Peiqiang
Published: (2025)
ResGuard: Enhancing Robustness Against Known Original Attacks in Deep Watermarking
by: Wang, Hanyi, et al.
Published: (2026)
by: Wang, Hanyi, et al.
Published: (2026)
Black-Box Skill Stealing Attack from Proprietary LLM Agents: An Empirical Study
by: Wang, Zihan, et al.
Published: (2026)
by: Wang, Zihan, et al.
Published: (2026)
DialogGuard: Multi-Agent Psychosocial Safety Evaluation of Sensitive LLM Responses
by: Luo, Han, et al.
Published: (2025)
by: Luo, Han, et al.
Published: (2025)
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
by: Xu, Peng, et al.
Published: (2024)
by: Xu, Peng, et al.
Published: (2024)
AgentGuard: Runtime Verification of AI Agents
by: Koohestani, Roham
Published: (2025)
by: Koohestani, Roham
Published: (2025)
GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning
by: Xiang, Zhen, et al.
Published: (2024)
by: Xiang, Zhen, et al.
Published: (2024)
The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants
by: Zhang, Yiqun, et al.
Published: (2025)
by: Zhang, Yiqun, et al.
Published: (2025)
Online Federation For Mixtures of Proprietary Agents with Black-Box Encoders
by: Yang, Xuwei, et al.
Published: (2025)
by: Yang, Xuwei, et al.
Published: (2025)
SciNet: Evaluating AI Agents in Relation-Aware Scientific Literature Retrieval
by: Shao, Chenyang, et al.
Published: (2025)
by: Shao, Chenyang, et al.
Published: (2025)
BehaviorGuard: Online Backdoor Defense for Deep Reinforcement Learning
by: Yu, Yinbo, et al.
Published: (2026)
by: Yu, Yinbo, et al.
Published: (2026)
ProbGuard: Probabilistic Runtime Monitoring for LLM Agent Safety
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Trusting What You Cannot See: Auditable Fine-Tuning and Inference for Proprietary AI
by: Jin, Heng, et al.
Published: (2026)
by: Jin, Heng, et al.
Published: (2026)
Linking Multitasking to Creative Process Engagement Through Psychological Detachment: Temporal Leadership as a Moderator
by: Jianfeng Yang, et al.
Published: (2025)
by: Jianfeng Yang, et al.
Published: (2025)
Ecosystem Carbon Fluxes Exhibit Thermal Response Thresholds at Which Carbon–Climate Feedback Changes
by: Xiaoni Xu, et al.
Published: (2025)
by: Xiaoni Xu, et al.
Published: (2025)
Electrospinning Fiber Membrane‐Derived Gel Polymer Electrolytes with High Mechanical Strength and Low Swelling Effect for High‐Safety Lithium Metal Batteries
by: Peng Wang, et al.
Published: (2024)
by: Peng Wang, et al.
Published: (2024)
Generic Guard AI in Stealth Game with Composite Potential Fields
by: Xu, Kaijie, et al.
Published: (2025)
by: Xu, Kaijie, et al.
Published: (2025)
PoseGuard: Pose-Guided Generation with Safety Guardrails
by: Wang, Kongxin, et al.
Published: (2025)
by: Wang, Kongxin, et al.
Published: (2025)
AgentGuard: Repurposing Agentic Orchestrator for Safety Evaluation of Tool Orchestration
by: Chen, Jizhou, et al.
Published: (2025)
by: Chen, Jizhou, et al.
Published: (2025)
Poly-Guard: Massive Multi-Domain Safety Policy-Grounded Guardrail Dataset
by: Kang, Mintong, et al.
Published: (2025)
by: Kang, Mintong, et al.
Published: (2025)
OmniGuard: Hybrid Manipulation Localization via Augmented Versatile Deep Image Watermarking
by: Zhang, Xuanyu, et al.
Published: (2024)
by: Zhang, Xuanyu, et al.
Published: (2024)
Interfacial Engineering of Hierarchical Iron Oxysulfide Integrated MoS 2 Heterostructures for Enhanced Oxygen Evolution Electrocatalysis
by: Xiaoli Shi, et al.
Published: (2025)
by: Xiaoli Shi, et al.
Published: (2025)
FlakyGuard: Automatically Fixing Flaky Tests at Industry Scale
by: Li, Chengpeng, et al.
Published: (2025)
by: Li, Chengpeng, et al.
Published: (2025)
GuardAD: Safeguarding Autonomous Driving MLLMs via Markovian Safety Logic
by: Zhang, Tianyuan, et al.
Published: (2026)
by: Zhang, Tianyuan, et al.
Published: (2026)
Questionnaire Responses Do not Capture the Safety of AI Agents
by: Hellrigel-Holderbaum, Max, et al.
Published: (2026)
by: Hellrigel-Holderbaum, Max, et al.
Published: (2026)
WebAgentGuard: A Reasoning-Driven Guard Model for Detecting Prompt Injection Attacks in Web Agents
by: Chen, Yulin, et al.
Published: (2026)
by: Chen, Yulin, et al.
Published: (2026)
A Sound‐Absorbing Metamaterial With Tree‐Inspired Bionic Helmholtz Resonators
by: Li Bo Wang, et al.
Published: (2025)
by: Li Bo Wang, et al.
Published: (2025)
Is a team only as strong as its weakest link? Quantifying the short-board effect with AI Agents
by: Xu, Xin, et al.
Published: (2026)
by: Xu, Xin, et al.
Published: (2026)
CACA Agent: Capability Collaboration based AI Agent
by: Xu, Peng, et al.
Published: (2024)
by: Xu, Peng, et al.
Published: (2024)
CareGuardAI: Context-Aware Multi-Agent Guardrails for Clinical Safety & Hallucination Mitigation in Patient-Facing LLMs
by: Nasarian, Elham, et al.
Published: (2026)
by: Nasarian, Elham, et al.
Published: (2026)
PerfGuard: A Performance-Aware Agent for Visual Content Generation
by: Chen, Zhipeng, et al.
Published: (2026)
by: Chen, Zhipeng, et al.
Published: (2026)
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
by: Miculicich, Lesly, et al.
Published: (2025)
by: Miculicich, Lesly, et al.
Published: (2025)
GrandGuard: Taxonomy, Benchmark, and Safeguards for Elderly-Chatbot Interaction Safety
by: Fan, Changxuan, et al.
Published: (2026)
by: Fan, Changxuan, et al.
Published: (2026)
Geography and Gains of Doctoral Mobility: Origin, Training Site and Employment Location Among Recent PhDs From Chinese Universities
by: Haotian Xu, et al.
Published: (2026)
by: Haotian Xu, et al.
Published: (2026)
State Oversight of the Private and Proprietary Sector.
by: Chaloux, Bruce N.
Published: (1985)
by: Chaloux, Bruce N.
Published: (1985)
Proprietary Rights in Data Bases and Software.
by: Levina, Arthur J.
Published: (1986)
by: Levina, Arthur J.
Published: (1986)
Taiwan Safety Benchmark and Breeze Guard: Toward Trustworthy AI for Taiwanese Mandarin
by: Hsu, Po-Chun, et al.
Published: (2026)
by: Hsu, Po-Chun, et al.
Published: (2026)
X-Guard: Multilingual Guard Agent for Content Moderation
by: Upadhayay, Bibek, et al.
Published: (2025)
by: Upadhayay, Bibek, et al.
Published: (2025)
HarmonyGuard: Toward Safety and Utility in Web Agents via Adaptive Policy Enhancement and Dual-Objective Optimization
by: Chen, Yurun, et al.
Published: (2025)
by: Chen, Yurun, et al.
Published: (2025)
Similar Items
-
A Comparative Evaluation of AI Agent Security Guardrails
by: Li, Qi, et al.
Published: (2026) -
Lemma on logarithmic derivative over directed manifolds
by: Lin, Peiqiang
Published: (2025) -
ResGuard: Enhancing Robustness Against Known Original Attacks in Deep Watermarking
by: Wang, Hanyi, et al.
Published: (2026) -
Black-Box Skill Stealing Attack from Proprietary LLM Agents: An Empirical Study
by: Wang, Zihan, et al.
Published: (2026) -
DialogGuard: Multi-Agent Psychosocial Safety Evaluation of Sensitive LLM Responses
by: Luo, Han, et al.
Published: (2025)