PSG-Agent: Personality-Aware Safety Guardrail for LLM-based Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Yaozu, Guo, Jizhou, Li, Dongyuan, Zou, Henry Peng, Huang, Wei-Chieh, Chen, Yankai, Wang, Zhen, Zhang, Weizhi, Li, Yangning, Zhang, Meng, Jiang, Renhe, Yu, Philip S. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deep Research with Open-Domain Evaluation and Multi-Stage Guardrails for Safety
by: Huang, Wei-Chieh, et al.
Published: (2025)
by: Huang, Wei-Chieh, et al.
Published: (2025)
Multi-Agent Autonomous Driving Systems with Large Language Models: A Survey of Recent Advances
by: Wu, Yaozu, et al.
Published: (2025)
by: Wu, Yaozu, et al.
Published: (2025)
A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
Node Role-Guided LLMs for Dynamic Graph Clustering
by: Li, Dongyuan, et al.
Published: (2026)
by: Li, Dongyuan, et al.
Published: (2026)
Embracing Trustworthy Brain-Agent Collaboration as Paradigm Extension for Intelligent Assistive Technologies
by: Chen, Yankai, et al.
Published: (2025)
by: Chen, Yankai, et al.
Published: (2025)
Diffusion Models in 3D Vision: A Survey
by: Wang, Zhen, et al.
Published: (2024)
by: Wang, Zhen, et al.
Published: (2024)
SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety
by: Liu, Zhe, et al.
Published: (2026)
by: Liu, Zhe, et al.
Published: (2026)
Towards Agentic RAG with Deep Reasoning: A Survey of RAG-Reasoning Systems in LLMs
by: Li, Yangning, et al.
Published: (2025)
by: Li, Yangning, et al.
Published: (2025)
A Survey on Deep Active Learning: Recent Advances and New Frontiers
by: Li, Dongyuan, et al.
Published: (2024)
by: Li, Dongyuan, et al.
Published: (2024)
GAM: Hierarchical Graph-based Agentic Memory for LLM Agents
by: Wu, Zhaofen, et al.
Published: (2026)
by: Wu, Zhaofen, et al.
Published: (2026)
Taming Recommendation Bias with Causal Intervention on Evolving Personal Popularity
by: Tan, Shiyin, et al.
Published: (2025)
by: Tan, Shiyin, et al.
Published: (2025)
When Users Change Their Mind: Evaluating Interruptible Agents in Long-Horizon Web Navigation
by: Zou, Henry Peng, et al.
Published: (2026)
by: Zou, Henry Peng, et al.
Published: (2026)
MemoryCD: Benchmarking Long-Context User Memory of LLM Agents for Lifelong Cross-Domain Personalization
by: Zhang, Weizhi, et al.
Published: (2026)
by: Zhang, Weizhi, et al.
Published: (2026)
TestNUC: Enhancing Test-Time Computing Approaches and Scaling through Neighboring Unlabeled Data Consistency
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
TimeClaw: A Time-Series AI Agent with Exploratory Execution Learning
by: Liu, Hangchen, et al.
Published: (2026)
by: Liu, Hangchen, et al.
Published: (2026)
AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security
by: Liu, Dongrui, et al.
Published: (2026)
by: Liu, Dongrui, et al.
Published: (2026)
Scaling Laws for Many-Shot In-Context Learning with Self-Generated Annotations
by: Gu, Zhengyao, et al.
Published: (2025)
by: Gu, Zhengyao, et al.
Published: (2025)
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
by: Li, Yangning, et al.
Published: (2025)
by: Li, Yangning, et al.
Published: (2025)
Integrating LLM and Diffusion-Based Agents for Social Simulation
by: Li, Xinyi, et al.
Published: (2025)
by: Li, Xinyi, et al.
Published: (2025)
Event-Aware Prompt Learning for Dynamic Graphs
by: Yu, Xingtong, et al.
Published: (2025)
by: Yu, Xingtong, et al.
Published: (2025)
Community-Invariant Graph Contrastive Learning
by: Tan, Shiyin, et al.
Published: (2024)
by: Tan, Shiyin, et al.
Published: (2024)
From Web Search towards Agentic Deep Research: Incentivizing Search with Reasoning Agents
by: Zhang, Weizhi, et al.
Published: (2025)
by: Zhang, Weizhi, et al.
Published: (2025)
A Comparative Evaluation of AI Agent Security Guardrails
by: Li, Qi, et al.
Published: (2026)
by: Li, Qi, et al.
Published: (2026)
Provably Secure Agent Guardrail
by: Wu, Benlong, et al.
Published: (2026)
by: Wu, Benlong, et al.
Published: (2026)
TabGen-ICL: Residual-Aware In-Context Example Selection for Tabular Data Generation
by: Fang, Liancheng, et al.
Published: (2025)
by: Fang, Liancheng, et al.
Published: (2025)
Routing Channel-Patch Dependencies in Time Series Forecasting with Graph Spectral Decomposition
by: Li, Dongyuan, et al.
Published: (2026)
by: Li, Dongyuan, et al.
Published: (2026)
PersonaAgent with GraphRAG: Community-Aware Knowledge Graphs for Personalized LLM
by: Liang, Siqi, et al.
Published: (2025)
by: Liang, Siqi, et al.
Published: (2025)
AgentGuard: Repurposing Agentic Orchestrator for Safety Evaluation of Tool Orchestration
by: Chen, Jizhou, et al.
Published: (2025)
by: Chen, Jizhou, et al.
Published: (2025)
CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification
by: Zhang, Hanrong, et al.
Published: (2026)
by: Zhang, Hanrong, et al.
Published: (2026)
Large Language Models as Urban Residents: An LLM Agent Framework for Personal Mobility Generation
by: Wang, Jiawei, et al.
Published: (2024)
by: Wang, Jiawei, et al.
Published: (2024)
NeuroFilter: Privacy Guardrails for Conversational LLM Agents
by: Das, Saswat, et al.
Published: (2026)
by: Das, Saswat, et al.
Published: (2026)
A Unified Retrieval Framework with Document Ranking and EDU Filtering for Multi-document Summarization
by: Tan, Shiyin, et al.
Published: (2025)
by: Tan, Shiyin, et al.
Published: (2025)
Revisiting Dynamic Graph Clustering via Matrix Factorization
by: Li, Dongyuan, et al.
Published: (2025)
by: Li, Dongyuan, et al.
Published: (2025)
AgentBelt: Deterministic Runtime Guardrails for Execution-Level Security in LLM Agents
by: Anonymous
Published: (2026)
by: Anonymous
Published: (2026)
Safety Guardrails for LLM-Enabled Robots
by: Ravichandran, Zachary, et al.
Published: (2025)
by: Ravichandran, Zachary, et al.
Published: (2025)
R-Judge: Benchmarking Safety Risk Awareness for LLM Agents
by: Yuan, Tongxin, et al.
Published: (2024)
by: Yuan, Tongxin, et al.
Published: (2024)
Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents
by: Li, Junkai, et al.
Published: (2024)
by: Li, Junkai, et al.
Published: (2024)
AGrail: A Lifelong Agent Guardrail with Effective and Adaptive Safety Detection
by: Luo, Weidi, et al.
Published: (2025)
by: Luo, Weidi, et al.
Published: (2025)
Agent-SafetyBench: Evaluating the Safety of LLM Agents
by: Zhang, Zhexin, et al.
Published: (2024)
by: Zhang, Zhexin, et al.
Published: (2024)
Similar Items
-
Deep Research with Open-Domain Evaluation and Multi-Stage Guardrails for Safety
by: Huang, Wei-Chieh, et al.
Published: (2025) -
Multi-Agent Autonomous Driving Systems with Large Language Models: A Survey of Recent Advances
by: Wu, Yaozu, et al.
Published: (2025) -
A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy
by: Zou, Henry Peng, et al.
Published: (2025) -
LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey
by: Zou, Henry Peng, et al.
Published: (2025) -
Node Role-Guided LLMs for Dynamic Graph Clustering
by: Li, Dongyuan, et al.
Published: (2026)