Privacy Guard & Token Parsimony by Prompt and Context Handling and LLM Routing
Fuente:
arXiv
Guardado en:
| Autor principal: | Langiu, Alessio |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents
por: Xiao, Wenjie, et al.
Publicado: (2026)
por: Xiao, Wenjie, et al.
Publicado: (2026)
Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
por: Hasan, Md. Mehedi, et al.
Publicado: (2025)
por: Hasan, Md. Mehedi, et al.
Publicado: (2025)
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
por: Lin, Lixing, et al.
Publicado: (2026)
por: Lin, Lixing, et al.
Publicado: (2026)
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
por: Zhao, Wei, et al.
Publicado: (2026)
por: Zhao, Wei, et al.
Publicado: (2026)
Token-Efficient Prompt Injection Attack: Provoking Cessation in LLM Reasoning via Adaptive Token Compression
por: Cui, Yu, et al.
Publicado: (2025)
por: Cui, Yu, et al.
Publicado: (2025)
CourtGuard: A Local, Multiagent Prompt Injection Classifier
por: Wu, Isaac, et al.
Publicado: (2025)
por: Wu, Isaac, et al.
Publicado: (2025)
Privacy-Preserving LLMs Routing
por: Wu, Xidong, et al.
Publicado: (2026)
por: Wu, Xidong, et al.
Publicado: (2026)
RTBAS: Defending LLM Agents Against Prompt Injection and Privacy Leakage
por: Zhong, Peter Yong, et al.
Publicado: (2025)
por: Zhong, Peter Yong, et al.
Publicado: (2025)
SnapGuard: Lightweight Prompt Injection Detection for Screenshot-Based Web Agents
por: Du, Mengyao, et al.
Publicado: (2026)
por: Du, Mengyao, et al.
Publicado: (2026)
X-Guard: Multilingual Guard Agent for Content Moderation
por: Upadhayay, Bibek, et al.
Publicado: (2025)
por: Upadhayay, Bibek, et al.
Publicado: (2025)
QGuard:Question-based Zero-shot Guard for Multi-modal LLM Safety
por: Lee, Taegyeong, et al.
Publicado: (2025)
por: Lee, Taegyeong, et al.
Publicado: (2025)
A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
por: Wei, Qianshan, et al.
Publicado: (2025)
por: Wei, Qianshan, et al.
Publicado: (2025)
MultiPhishGuard: An Explainable and Adaptive Multi-Agent LLM System for Phishing Email Detection
por: Xue, Yinuo, et al.
Publicado: (2025)
por: Xue, Yinuo, et al.
Publicado: (2025)
MCP-Guard: A Multi-Stage Defense-in-Depth Framework for Securing Model Context Protocol in Agentic AI
por: Xing, Wenpeng, et al.
Publicado: (2025)
por: Xing, Wenpeng, et al.
Publicado: (2025)
Life-Cycle Routing Vulnerabilities of LLM Router
por: Lin, Qiqi, et al.
Publicado: (2025)
por: Lin, Qiqi, et al.
Publicado: (2025)
Breaking the Protocol: Security Analysis of the Model Context Protocol Specification and Prompt Injection Vulnerabilities in Tool-Integrated LLM Agents
por: Maloyan, Narek, et al.
Publicado: (2026)
por: Maloyan, Narek, et al.
Publicado: (2026)
Privacy in Action: Towards Realistic Privacy Mitigation and Evaluation for LLM-Powered Agents
por: Wang, Shouju, et al.
Publicado: (2025)
por: Wang, Shouju, et al.
Publicado: (2025)
AEGIS : Automated Co-Evolutionary Framework for Guarding Prompt Injections Schema
por: Liu, Ting-Chun, et al.
Publicado: (2025)
por: Liu, Ting-Chun, et al.
Publicado: (2025)
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt
por: Wang, Zhilong, et al.
Publicado: (2025)
por: Wang, Zhilong, et al.
Publicado: (2025)
Unveiling Privacy Risks in LLM Agent Memory
por: Wang, Bo, et al.
Publicado: (2025)
por: Wang, Bo, et al.
Publicado: (2025)
Smart Privacy Policy Assistant: An LLM-Powered System for Transparent and Actionable Privacy Notices
por: Kalvakuntla, Sriharshini, et al.
Publicado: (2026)
por: Kalvakuntla, Sriharshini, et al.
Publicado: (2026)
Safety-Oriented Routing Analysis of Mixtral MoE Under Benign and Harmful Prompts
por: Siddiky, Md Nurul Absar
Publicado: (2026)
por: Siddiky, Md Nurul Absar
Publicado: (2026)
GuardReasoner: Towards Reasoning-based LLM Safeguards
por: Liu, Yue, et al.
Publicado: (2025)
por: Liu, Yue, et al.
Publicado: (2025)
LLM Access Shield: Domain-Specific LLM Framework for Privacy Policy Compliance
por: Wang, Yu, et al.
Publicado: (2025)
por: Wang, Yu, et al.
Publicado: (2025)
CoT-Guard: Small Models for Strong Monitoring
por: Diwan, Nirav, et al.
Publicado: (2026)
por: Diwan, Nirav, et al.
Publicado: (2026)
InjecGuard: Benchmarking and Mitigating Over-defense in Prompt Injection Guardrail Models
por: Li, Hao, et al.
Publicado: (2024)
por: Li, Hao, et al.
Publicado: (2024)
No Free Lunch Theorem for Privacy-Preserving LLM Inference
por: Zhang, Xiaojin, et al.
Publicado: (2024)
por: Zhang, Xiaojin, et al.
Publicado: (2024)
LLM-Guided Prompt Evolution for Password Guessing
por: Mazin, Vladimir A., et al.
Publicado: (2026)
por: Mazin, Vladimir A., et al.
Publicado: (2026)
Guarding Your Conversations: Privacy Gatekeepers for Secure Interactions with Cloud-Based AI Models
por: Uzor, GodsGift, et al.
Publicado: (2025)
por: Uzor, GodsGift, et al.
Publicado: (2025)
Casper: Prompt Sanitization for Protecting User Privacy in Web-Based Large Language Models
por: Chong, Chun Jie, et al.
Publicado: (2024)
por: Chong, Chun Jie, et al.
Publicado: (2024)
DRIP: Defending Prompt Injection via Token-wise Representation Editing and Residual Instruction Fusion
por: Liu, Ruofan, et al.
Publicado: (2025)
por: Liu, Ruofan, et al.
Publicado: (2025)
Listening Alone, Understanding Together: Collaborative Context Recovery for Privacy-Aware AI
por: Srivastava, Tanmay, et al.
Publicado: (2026)
por: Srivastava, Tanmay, et al.
Publicado: (2026)
DistillGuard: Evaluating Defenses Against LLM Knowledge Distillation
por: Jiang, Bo
Publicado: (2026)
por: Jiang, Bo
Publicado: (2026)
AIRGuard: Guarding Agent Actions with Runtime Authority Control
por: Qin, Suliu, et al.
Publicado: (2026)
por: Qin, Suliu, et al.
Publicado: (2026)
Towards Confidential and Efficient LLM Inference with Dual Privacy Protection
por: Yu, Honglan, et al.
Publicado: (2025)
por: Yu, Honglan, et al.
Publicado: (2025)
LLM-PBE: Assessing Data Privacy in Large Language Models
por: Li, Qinbin, et al.
Publicado: (2024)
por: Li, Qinbin, et al.
Publicado: (2024)
MoJE: Mixture of Jailbreak Experts, Naive Tabular Classifiers as Guard for Prompt Attacks
por: Cornacchia, Giandomenico, et al.
Publicado: (2024)
por: Cornacchia, Giandomenico, et al.
Publicado: (2024)
UTF:Undertrained Tokens as Fingerprints A Novel Approach to LLM Identification
por: Cai, Jiacheng, et al.
Publicado: (2024)
por: Cai, Jiacheng, et al.
Publicado: (2024)
Signed-Prompt: A New Approach to Prevent Prompt Injection Attacks Against LLM-Integrated Applications
por: Suo, Xuchen
Publicado: (2024)
por: Suo, Xuchen
Publicado: (2024)
GuardReasoner-VL: Safeguarding VLMs via Reinforced Reasoning
por: Liu, Yue, et al.
Publicado: (2025)
por: Liu, Yue, et al.
Publicado: (2025)
Ejemplares similares
-
RouteGuard: Internal-Signal Detection of Skill Poisoning in LLM Agents
por: Xiao, Wenjie, et al.
Publicado: (2026) -
Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
por: Hasan, Md. Mehedi, et al.
Publicado: (2025) -
Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection
por: Lin, Lixing, et al.
Publicado: (2026) -
ClawGuard: A Runtime Security Framework for Tool-Augmented LLM Agents Against Indirect Prompt Injection
por: Zhao, Wei, et al.
Publicado: (2026) -
Token-Efficient Prompt Injection Attack: Provoking Cessation in LLM Reasoning via Adaptive Token Compression
por: Cui, Yu, et al.
Publicado: (2025)