ShieldNet: Network-Level Guardrails against Emerging Supply-Chain Injections in Agentic Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Yuan, Zhuowen, Chen, Zhaorun, Xiang, Zhen, Bastian, Nathaniel D., Hashemi, Seyyed Hadi, Xiao, Chaowei, Guo, Wenbo, Li, Bo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SafeVision: Efficient Image Guardrail with Robust Policy Adherence and Explainability
por: Xu, Peiyang, et al.
Publicado: (2025)
por: Xu, Peiyang, et al.
Publicado: (2025)
AgentPoison: Red-teaming LLM Agents via Poisoning Memory or Knowledge Bases
por: Chen, Zhaorun, et al.
Publicado: (2024)
por: Chen, Zhaorun, et al.
Publicado: (2024)
RigorLLM: Resilient Guardrails for Large Language Models against Undesired Content
por: Yuan, Zhuowen, et al.
Publicado: (2024)
por: Yuan, Zhuowen, et al.
Publicado: (2024)
System-Level Defense against Indirect Prompt Injection Attacks: An Information Flow Control Perspective
por: Wu, Fangzhou, et al.
Publicado: (2024)
por: Wu, Fangzhou, et al.
Publicado: (2024)
ShieldAgent: Shielding Agents via Verifiable Safety Policy Reasoning
por: Chen, Zhaorun, et al.
Publicado: (2025)
por: Chen, Zhaorun, et al.
Publicado: (2025)
SafeWatch: An Efficient Safety-Policy Following Video Guardrail Model with Transparent Explanations
por: Chen, Zhaorun, et al.
Publicado: (2024)
por: Chen, Zhaorun, et al.
Publicado: (2024)
LPG: Balancing Efficiency and Policy Reasoning in Latent Policy Guardrails
por: Li, Nanxi, et al.
Publicado: (2026)
por: Li, Nanxi, et al.
Publicado: (2026)
ML-Bench&Guard: Policy-Grounded Multilingual Safety Benchmark and Guardrail for Large Language Models
por: Zhao, Yunhan, et al.
Publicado: (2026)
por: Zhao, Yunhan, et al.
Publicado: (2026)
Automatic and Universal Prompt Injection Attacks against Large Language Models
por: Liu, Xiaogeng, et al.
Publicado: (2024)
por: Liu, Xiaogeng, et al.
Publicado: (2024)
Mitigating Indirect Prompt Injection via Instruction-Following Intent Analysis
por: Kang, Mintong, et al.
Publicado: (2025)
por: Kang, Mintong, et al.
Publicado: (2025)
ReasAlign: Reasoning Enhanced Safety Alignment against Prompt Injection Attack
por: Li, Hao, et al.
Publicado: (2026)
por: Li, Hao, et al.
Publicado: (2026)
Analgesic Efficacy of Bromelain and Bromelain Plus Turmeric for Pain Control After Orthodontic Separator Placement: A Triple‐Blind Randomized Clinical Trial
por: Shabnam Ajami, et al.
Publicado: (2025)
por: Shabnam Ajami, et al.
Publicado: (2025)
Roadmap to Precision 3D Printing of Cellulose: Rheology‐Guided Formulation, Fidelity Assessment, and Application Horizons (Adv. Mater. Technol. 8/2026)
por: Majed Amini, et al.
Publicado: (2026)
por: Majed Amini, et al.
Publicado: (2026)
Structural Design for EMI Shielding: From Underlying Mechanisms to Common Pitfalls
por: Ali Akbar Isari, et al.
Publicado: (2024)
por: Ali Akbar Isari, et al.
Publicado: (2024)
OneShield -- the Next Generation of LLM Guardrails
por: DeLuca, Chad, et al.
Publicado: (2025)
por: DeLuca, Chad, et al.
Publicado: (2025)
A Survey of Visual Attention Models
por: Seyyed Mohammad Reza Hashemi
Publicado: (2015)
por: Seyyed Mohammad Reza Hashemi
Publicado: (2015)
Architecting Secure AI Agents: Perspectives on System-Level Defenses Against Indirect Prompt Injection Attacks
por: Xiang, Chong, et al.
Publicado: (2026)
por: Xiang, Chong, et al.
Publicado: (2026)
Poly-Guard: Massive Multi-Domain Safety Policy-Grounded Guardrail Dataset
por: Kang, Mintong, et al.
Publicado: (2025)
por: Kang, Mintong, et al.
Publicado: (2025)
Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems
por: Hackett, William, et al.
Publicado: (2025)
por: Hackett, William, et al.
Publicado: (2025)
AGrail: A Lifelong Agent Guardrail with Effective and Adaptive Safety Detection
por: Luo, Weidi, et al.
Publicado: (2025)
por: Luo, Weidi, et al.
Publicado: (2025)
AdaShield: Safeguarding Multimodal Large Language Models from Structure-based Attack via Adaptive Shield Prompting
por: Wang, Yu, et al.
Publicado: (2024)
por: Wang, Yu, et al.
Publicado: (2024)
Peering Behind the Shield: Guardrail Identification in Large Language Models
por: Yang, Ziqing, et al.
Publicado: (2025)
por: Yang, Ziqing, et al.
Publicado: (2025)
Emerging Trends and Challenges in Supply Chain Management and Sustainability
por: Isabel Castillo‐Pérez, et al.
Publicado: (2025)
por: Isabel Castillo‐Pérez, et al.
Publicado: (2025)
Supply Chain Network Extraction and Entity Classification Leveraging Large Language Models
por: Liu, Tong, et al.
Publicado: (2024)
por: Liu, Tong, et al.
Publicado: (2024)
FATH: Authentication-based Test-time Defense against Indirect Prompt Injection Attacks
por: Wang, Jiongxiao, et al.
Publicado: (2024)
por: Wang, Jiongxiao, et al.
Publicado: (2024)
Evaluation of Attribution Bias in Generator-Aware Retrieval-Augmented Large Language Models
por: Abolghasemi, Amin, et al.
Publicado: (2024)
por: Abolghasemi, Amin, et al.
Publicado: (2024)
Agentic AI Sustainability Assessment for Supply Chain Document Insights
por: Gosmar, Diego, et al.
Publicado: (2025)
por: Gosmar, Diego, et al.
Publicado: (2025)
Formal Analysis and Supply Chain Security for Agentic AI Skills
por: Bhardwaj, Varun Pratap
Publicado: (2026)
por: Bhardwaj, Varun Pratap
Publicado: (2026)
ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models
por: Yu, Chung-En Johnny, et al.
Publicado: (2025)
por: Yu, Chung-En Johnny, et al.
Publicado: (2025)
EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
por: Liao, Zeyi, et al.
Publicado: (2024)
por: Liao, Zeyi, et al.
Publicado: (2024)
Generalized Holographic Reduced Representations
por: Yeung, Calvin, et al.
Publicado: (2024)
por: Yeung, Calvin, et al.
Publicado: (2024)
AI Supply Chains: An Emerging Ecosystem of AI Actors, Products, and Services
por: Hopkins, Aspen, et al.
Publicado: (2025)
por: Hopkins, Aspen, et al.
Publicado: (2025)
Dynamic Self-Assessment of Supply Chains Performance: an Emerging Market Approach
por: M. Cedillo-Campos
Publicado: (2013)
por: M. Cedillo-Campos
Publicado: (2013)
ClusComp: A Simple Paradigm for Model Compression and Efficient Finetuning
por: Liao, Baohao, et al.
Publicado: (2025)
por: Liao, Baohao, et al.
Publicado: (2025)
Unilogit: Robust Machine Unlearning for LLMs Using Uniform-Target Self-Distillation
por: Vasilev, Stefan, et al.
Publicado: (2025)
por: Vasilev, Stefan, et al.
Publicado: (2025)
Automating Supply Chain Disruption Monitoring via an Agentic AI Approach
por: AlMahri, Sara, et al.
Publicado: (2026)
por: AlMahri, Sara, et al.
Publicado: (2026)
LLM Scalability Risk for Agentic-AI and Model Supply Chain Security
por: Ahi, Kiarash, et al.
Publicado: (2026)
por: Ahi, Kiarash, et al.
Publicado: (2026)
Novel Blockchain-based Protocols for Electronic Voting and Auctions
por: Lin, Zhaorun
Publicado: (2025)
por: Lin, Zhaorun
Publicado: (2025)
TrojFM: Resource-efficient Backdoor Attacks against Very Large Foundation Models
por: Nie, Yuzhou., et al.
Publicado: (2024)
por: Nie, Yuzhou., et al.
Publicado: (2024)
Noise Injection Systemically Degrades Large Language Model Safety Guardrails
por: Shahani, Prithviraj Singh, et al.
Publicado: (2025)
por: Shahani, Prithviraj Singh, et al.
Publicado: (2025)
Ejemplares similares
-
SafeVision: Efficient Image Guardrail with Robust Policy Adherence and Explainability
por: Xu, Peiyang, et al.
Publicado: (2025) -
AgentPoison: Red-teaming LLM Agents via Poisoning Memory or Knowledge Bases
por: Chen, Zhaorun, et al.
Publicado: (2024) -
RigorLLM: Resilient Guardrails for Large Language Models against Undesired Content
por: Yuan, Zhuowen, et al.
Publicado: (2024) -
System-Level Defense against Indirect Prompt Injection Attacks: An Information Flow Control Perspective
por: Wu, Fangzhou, et al.
Publicado: (2024) -
ShieldAgent: Shielding Agents via Verifiable Safety Policy Reasoning
por: Chen, Zhaorun, et al.
Publicado: (2025)