PICO: Secure Transformers via Robust Prompt Isolation and Cybersecurity Oversight
Fuente:
arXiv
Salvato in:
| Autori principali: | Goertzel, Ben, Yibelo, Paulos |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SecPE: Secure Prompt Ensembling for Private and Robust Large Language Models
di: Zhang, Jiawen, et al.
Pubblicazione: (2025)
di: Zhang, Jiawen, et al.
Pubblicazione: (2025)
Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations
di: Shafranskyi, Daniil, et al.
Pubblicazione: (2026)
di: Shafranskyi, Daniil, et al.
Pubblicazione: (2026)
From Legacy to Standard: LLM-Assisted Transformation of Cybersecurity Playbooks into CACAO Format
di: Gurabi, Mehdi Akbari, et al.
Pubblicazione: (2025)
di: Gurabi, Mehdi Akbari, et al.
Pubblicazione: (2025)
Privacy-Aware RAG: Secure and Isolated Knowledge Retrieval
di: Zhou, Pengcheng, et al.
Pubblicazione: (2025)
di: Zhou, Pengcheng, et al.
Pubblicazione: (2025)
On the (In-)Security of the Shuffling Defense in the Transformer Secure Inference
di: Li, Zhengyi, et al.
Pubblicazione: (2026)
di: Li, Zhengyi, et al.
Pubblicazione: (2026)
Integrative Approaches in Cybersecurity and AI
di: Omar, Marwan
Pubblicazione: (2024)
di: Omar, Marwan
Pubblicazione: (2024)
Securing AI Agents Against Prompt Injection Attacks
di: Ramakrishnan, Badrinath, et al.
Pubblicazione: (2025)
di: Ramakrishnan, Badrinath, et al.
Pubblicazione: (2025)
DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
Encrypted Prompt: Securing LLM Applications Against Unauthorized Actions
di: Chan, Shih-Han
Pubblicazione: (2025)
di: Chan, Shih-Han
Pubblicazione: (2025)
Enhancing Security of AI-Based Code Synthesis with GitHub Copilot via Cheap and Efficient Prompt-Engineering
di: Res, Jakub, et al.
Pubblicazione: (2024)
di: Res, Jakub, et al.
Pubblicazione: (2024)
SecureBERT 2.0: Advanced Language Model for Cybersecurity Intelligence
di: Aghaei, Ehsan, et al.
Pubblicazione: (2025)
di: Aghaei, Ehsan, et al.
Pubblicazione: (2025)
SoK: Taxonomy and Evaluation of Prompt Security in Large Language Models
di: Hong, Hanbin, et al.
Pubblicazione: (2025)
di: Hong, Hanbin, et al.
Pubblicazione: (2025)
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
di: Evtimov, Ivan, et al.
Pubblicazione: (2025)
di: Evtimov, Ivan, et al.
Pubblicazione: (2025)
Reinforcement Learning for Automated Cybersecurity Penetration Testing
di: López-Montero, Daniel, et al.
Pubblicazione: (2025)
di: López-Montero, Daniel, et al.
Pubblicazione: (2025)
Dynamic Risk Assessments for Offensive Cybersecurity Agents
di: Wei, Boyi, et al.
Pubblicazione: (2025)
di: Wei, Boyi, et al.
Pubblicazione: (2025)
A Survey on Offensive AI Within Cybersecurity
di: Girhepuje, Sahil, et al.
Pubblicazione: (2024)
di: Girhepuje, Sahil, et al.
Pubblicazione: (2024)
A Survey of Large Language Models in Cybersecurity
di: da Silva, Gabriel de Jesus Coelho, et al.
Pubblicazione: (2024)
di: da Silva, Gabriel de Jesus Coelho, et al.
Pubblicazione: (2024)
Empirical Analysis of Adversarial Robustness and Explainability Drift in Cybersecurity Classifiers
di: Rajhans, Mona, et al.
Pubblicazione: (2026)
di: Rajhans, Mona, et al.
Pubblicazione: (2026)
Is Your Prompt Poisoning Code? Defect Induction Rates and Security Mitigation Strategies
di: Wang, Bin, et al.
Pubblicazione: (2025)
di: Wang, Bin, et al.
Pubblicazione: (2025)
Joint Optimization of Prompt Security and System Performance in Edge-Cloud LLM Systems
di: Huang, Haiyang, et al.
Pubblicazione: (2025)
di: Huang, Haiyang, et al.
Pubblicazione: (2025)
Demo: SGCode: A Flexible Prompt-Optimizing System for Secure Generation of Code
di: Ton, Khiem, et al.
Pubblicazione: (2024)
di: Ton, Khiem, et al.
Pubblicazione: (2024)
The Granularity Mismatch in Agent Security: Argument-Level Provenance Solves Enforcement and Isolates the LLM Reasoning Bottleneck
di: Fan, Linfeng, et al.
Pubblicazione: (2026)
di: Fan, Linfeng, et al.
Pubblicazione: (2026)
Security-First AI: Foundations for Robust and Trustworthy Systems
di: Tallam, Krti
Pubblicazione: (2025)
di: Tallam, Krti
Pubblicazione: (2025)
Semantic Encryption: Secure and Effective Interaction with Cloud-based Large Language Models via Semantic Transformation
di: Chen, Dong, et al.
Pubblicazione: (2025)
di: Chen, Dong, et al.
Pubblicazione: (2025)
The Adaptive Arms Race: Redefining Robustness in AI Security
di: Tsingenopoulos, Ilias, et al.
Pubblicazione: (2023)
di: Tsingenopoulos, Ilias, et al.
Pubblicazione: (2023)
Nimbus: Secure and Efficient Two-Party Inference for Transformers
di: Li, Zhengyi, et al.
Pubblicazione: (2024)
di: Li, Zhengyi, et al.
Pubblicazione: (2024)
Know Thy Enemy: Securing LLMs Against Prompt Injection via Diverse Data Synthesis and Instruction-Level Chain-of-Thought Learning
di: Chang, Zhiyuan, et al.
Pubblicazione: (2026)
di: Chang, Zhiyuan, et al.
Pubblicazione: (2026)
Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks
di: Chen, Sizhe, et al.
Pubblicazione: (2025)
di: Chen, Sizhe, et al.
Pubblicazione: (2025)
Reasoning Under Threat: Symbolic and Neural Techniques for Cybersecurity Verification
di: Veronica, Sarah
Pubblicazione: (2025)
di: Veronica, Sarah
Pubblicazione: (2025)
Neuro-Symbolic AI for Cybersecurity: State of the Art, Challenges, and Opportunities
di: Hakim, Safayat Bin, et al.
Pubblicazione: (2025)
di: Hakim, Safayat Bin, et al.
Pubblicazione: (2025)
Cognitive Cybersecurity for Artificial Intelligence: Guardrail Engineering with CCS-7
di: Aydin, Yuksel
Pubblicazione: (2025)
di: Aydin, Yuksel
Pubblicazione: (2025)
From Texts to Shields: Convergence of Large Language Models and Cybersecurity
di: Li, Tao, et al.
Pubblicazione: (2025)
di: Li, Tao, et al.
Pubblicazione: (2025)
Explainable but Vulnerable: Adversarial Attacks on XAI Explanation in Cybersecurity Applications
di: Mia, Maraz, et al.
Pubblicazione: (2025)
di: Mia, Maraz, et al.
Pubblicazione: (2025)
CyberCertBench: Evaluating LLMs in Cybersecurity Certification Knowledge
di: Keppler, Gustav, et al.
Pubblicazione: (2026)
di: Keppler, Gustav, et al.
Pubblicazione: (2026)
When LLMs Meet Cybersecurity: A Systematic Literature Review
di: Zhang, Jie, et al.
Pubblicazione: (2024)
di: Zhang, Jie, et al.
Pubblicazione: (2024)
CyberEvolver: Structured Self-Evolution for Cybersecurity Agents On the Fly
di: Fan, Yihe, et al.
Pubblicazione: (2026)
di: Fan, Yihe, et al.
Pubblicazione: (2026)
Improving Generalization on Cybersecurity Tasks with Multi-Modal Contrastive Learning
di: Huang, Jianan, et al.
Pubblicazione: (2026)
di: Huang, Jianan, et al.
Pubblicazione: (2026)
CIPHER: Cybersecurity Intelligent Penetration-testing Helper for Ethical Researcher
di: Pratama, Derry, et al.
Pubblicazione: (2024)
di: Pratama, Derry, et al.
Pubblicazione: (2024)
Structured Security Auditing and Robustness Enhancement for Untrusted Agent Skills
di: Lv, Lijia, et al.
Pubblicazione: (2026)
di: Lv, Lijia, et al.
Pubblicazione: (2026)
F2A: An Innovative Approach for Prompt Injection by Utilizing Feign Security Detection Agents
di: Ren, Yupeng
Pubblicazione: (2024)
di: Ren, Yupeng
Pubblicazione: (2024)
Documenti analoghi
-
SecPE: Secure Prompt Ensembling for Private and Robust Large Language Models
di: Zhang, Jiawen, et al.
Pubblicazione: (2025) -
Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations
di: Shafranskyi, Daniil, et al.
Pubblicazione: (2026) -
From Legacy to Standard: LLM-Assisted Transformation of Cybersecurity Playbooks into CACAO Format
di: Gurabi, Mehdi Akbari, et al.
Pubblicazione: (2025) -
Privacy-Aware RAG: Secure and Isolated Knowledge Retrieval
di: Zhou, Pengcheng, et al.
Pubblicazione: (2025) -
On the (In-)Security of the Shuffling Defense in the Transformer Secure Inference
di: Li, Zhengyi, et al.
Pubblicazione: (2026)