SecInfer: Preventing Prompt Injection via Inference-time Scaling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yupei, Wang, Yanting, Jia, Yuqi, Jia, Jinyuan, Gong, Neil Zhenqiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DataSentinel: A Game-Theoretic Detection of Prompt Injection Attacks
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
A Critical Evaluation of Defenses against Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
PromptLocate: Localizing Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
von: Jia, Yuqi, et al.
Veröffentlicht: (2025)
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
von: Liu, Yupei, et al.
Veröffentlicht: (2023)
von: Liu, Yupei, et al.
Veröffentlicht: (2023)
TrojanDec: Data-free Detection of Trojan Inputs in Self-supervised Learning
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
von: Liu, Yupei, et al.
Veröffentlicht: (2025)
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
von: Liu, Yinuo, et al.
Veröffentlicht: (2025)
Evaluating LLM-based Personal Information Extraction and Countermeasures
von: Liu, Yupei, et al.
Veröffentlicht: (2024)
von: Liu, Yupei, et al.
Veröffentlicht: (2024)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
von: Geng, Runpeng, et al.
Veröffentlicht: (2025)
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
von: Zou, Wei, et al.
Veröffentlicht: (2025)
von: Zou, Wei, et al.
Veröffentlicht: (2025)
ObliInjection: Order-Oblivious Prompt Injection Attack to LLM Agents with Multi-source Data
von: Wang, Reachal, et al.
Veröffentlicht: (2025)
von: Wang, Reachal, et al.
Veröffentlicht: (2025)
PIArena: A Platform for Prompt Injection Evaluation
von: Geng, Runpeng, et al.
Veröffentlicht: (2026)
von: Geng, Runpeng, et al.
Veröffentlicht: (2026)
Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening
von: Zhang, Mohan, et al.
Veröffentlicht: (2026)
von: Zhang, Mohan, et al.
Veröffentlicht: (2026)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
von: Shao, Zedian, et al.
Veröffentlicht: (2024)
von: Shao, Zedian, et al.
Veröffentlicht: (2024)
Optimization-based Prompt Injection Attack to LLM-as-a-Judge
von: Shi, Jiawen, et al.
Veröffentlicht: (2024)
von: Shi, Jiawen, et al.
Veröffentlicht: (2024)
Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection
von: Shao, Zedian, et al.
Veröffentlicht: (2026)
von: Shao, Zedian, et al.
Veröffentlicht: (2026)
PromptArmor: Simple yet Effective Prompt Injection Defenses
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
von: Shi, Tianneng, et al.
Veröffentlicht: (2025)
AgentWatcher: A Rule-based Prompt Injection Monitor
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)
FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
TracLLM: A Generic Framework for Attributing Long Context LLMs
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
von: Wang, Yanting, et al.
Veröffentlicht: (2025)
PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses
von: Yin, Chenlong, et al.
Veröffentlicht: (2026)
von: Yin, Chenlong, et al.
Veröffentlicht: (2026)
AlignSentinel: Alignment-Aware Detection of Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2026)
von: Jia, Yuqi, et al.
Veröffentlicht: (2026)
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
von: Wang, Xilong, et al.
Veröffentlicht: (2026)
von: Wang, Xilong, et al.
Veröffentlicht: (2026)
Signed-Prompt: A New Approach to Prevent Prompt Injection Attacks Against LLM-Integrated Applications
von: Suo, Xuchen
Veröffentlicht: (2024)
von: Suo, Xuchen
Veröffentlicht: (2024)
EnsembleSHAP: Faithful and Certifiably Robust Attribution for Random Subspace Method
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
von: Wang, Yanting, et al.
Veröffentlicht: (2026)
SecPE: Secure Prompt Ensembling for Private and Robust Large Language Models
von: Zhang, Jiawen, et al.
Veröffentlicht: (2025)
von: Zhang, Jiawen, et al.
Veröffentlicht: (2025)
ICON: Indirect Prompt Injection Defense for Agents based on Inference-Time Correction
von: Wang, Che, et al.
Veröffentlicht: (2026)
von: Wang, Che, et al.
Veröffentlicht: (2026)
Toward Trustworthy Agentic AI: A Multimodal Framework for Preventing Prompt Injection Attacks
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2025)
von: Syed, Toqeer Ali, et al.
Veröffentlicht: (2025)
Competitive Advantage Attacks to Decentralized Federated Learning
von: Jia, Yuqi, et al.
Veröffentlicht: (2023)
von: Jia, Yuqi, et al.
Veröffentlicht: (2023)
SecMoE: Communication-Efficient Secure MoE Inference via Select-Then-Compute
von: Shen, Bowen, et al.
Veröffentlicht: (2026)
von: Shen, Bowen, et al.
Veröffentlicht: (2026)
SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents
von: Jia, Xiaojun, et al.
Veröffentlicht: (2026)
von: Jia, Xiaojun, et al.
Veröffentlicht: (2026)
Know Thy Enemy: Securing LLMs Against Prompt Injection via Diverse Data Synthesis and Instruction-Level Chain-of-Thought Learning
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Chang, Zhiyuan, et al.
Veröffentlicht: (2026)
To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt
von: Wang, Zhilong, et al.
Veröffentlicht: (2025)
von: Wang, Zhilong, et al.
Veröffentlicht: (2025)
Refusing Safe Prompts for Multi-modal Large Language Models
von: Shao, Zedian, et al.
Veröffentlicht: (2024)
von: Shao, Zedian, et al.
Veröffentlicht: (2024)
Bypassing Prompt Injection Detectors through Evasive Injections
von: Rahman, Md Jahedur, et al.
Veröffentlicht: (2026)
von: Rahman, Md Jahedur, et al.
Veröffentlicht: (2026)
How Vulnerable Are AI Agents to Indirect Prompt Injections? Insights from a Large-Scale Public Competition
von: Dziemian, Mateusz, et al.
Veröffentlicht: (2026)
von: Dziemian, Mateusz, et al.
Veröffentlicht: (2026)
Defeating Prompt Injections by Design
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2025)
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2025)
FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts
von: Gong, Yichen, et al.
Veröffentlicht: (2023)
von: Gong, Yichen, et al.
Veröffentlicht: (2023)
SoK: On Gradient Leakage in Federated Learning
von: Du, Jiacheng, et al.
Veröffentlicht: (2024)
von: Du, Jiacheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DataSentinel: A Game-Theoretic Detection of Prompt Injection Attacks
von: Liu, Yupei, et al.
Veröffentlicht: (2025) -
A Critical Evaluation of Defenses against Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2025) -
PromptLocate: Localizing Prompt Injection Attacks
von: Jia, Yuqi, et al.
Veröffentlicht: (2025) -
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
von: Liu, Yupei, et al.
Veröffentlicht: (2023) -
TrojanDec: Data-free Detection of Trojan Inputs in Self-supervised Learning
von: Liu, Yupei, et al.
Veröffentlicht: (2025)