Cognitive Overload Attack:Prompt Injection for Long Context
Fuente:
arXiv
Salvato in:
| Autori principali: | Upadhayay, Bibek, Behzadan, Vahid, Karbasi, Amin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs
di: Upadhayay, Bibek, et al.
Pubblicazione: (2024)
di: Upadhayay, Bibek, et al.
Pubblicazione: (2024)
TaCo: Enhancing Cross-Lingual Transfer for Low-Resource Languages in LLMs through Translation-Assisted Chain-of-Thought Processes
di: Upadhayay, Bibek, et al.
Pubblicazione: (2023)
di: Upadhayay, Bibek, et al.
Pubblicazione: (2023)
X-Guard: Multilingual Guard Agent for Content Moderation
di: Upadhayay, Bibek, et al.
Pubblicazione: (2025)
di: Upadhayay, Bibek, et al.
Pubblicazione: (2025)
When Chain-of-Thought Backfires: Evaluating Prompt Sensitivity in Medical Language Models
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
Prompt Injection Attacks in Defended Systems
di: Khomsky, Daniil, et al.
Pubblicazione: (2024)
di: Khomsky, Daniil, et al.
Pubblicazione: (2024)
Cognitive Overload: Jailbreaking Large Language Models with Overloaded Logical Thinking
di: Xu, Nan, et al.
Pubblicazione: (2023)
di: Xu, Nan, et al.
Pubblicazione: (2023)
Learning Task Representations from In-Context Learning
di: Saglam, Baturay, et al.
Pubblicazione: (2025)
di: Saglam, Baturay, et al.
Pubblicazione: (2025)
PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization
di: Geng, Runpeng, et al.
Pubblicazione: (2025)
di: Geng, Runpeng, et al.
Pubblicazione: (2025)
Defense against Prompt Injection Attacks via Mixture of Encodings
di: Zhang, Ruiyi, et al.
Pubblicazione: (2025)
di: Zhang, Ruiyi, et al.
Pubblicazione: (2025)
Dialogue Injection Attack: Jailbreaking LLMs through Context Manipulation
di: Meng, Wenlong, et al.
Pubblicazione: (2025)
di: Meng, Wenlong, et al.
Pubblicazione: (2025)
Investigating the Vulnerability of LLM-as-a-Judge Architectures to Prompt-Injection Attacks
di: Maloyan, Narek, et al.
Pubblicazione: (2025)
di: Maloyan, Narek, et al.
Pubblicazione: (2025)
Is Your Prompt Safe? Investigating Prompt Injection Attacks Against Open-Source LLMs
di: Wang, Jiawen, et al.
Pubblicazione: (2025)
di: Wang, Jiawen, et al.
Pubblicazione: (2025)
Defending Against Indirect Prompt Injection Attacks With Spotlighting
di: Hines, Keegan, et al.
Pubblicazione: (2024)
di: Hines, Keegan, et al.
Pubblicazione: (2024)
Mechanistically Guided LoRA Improves Paraphrase Consistency in Medical Vision-Language Models
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
Promise of Data-Driven Modeling and Decision Support for Precision Oncology and Theranostics
di: Sadanandan, Binesh, et al.
Pubblicazione: (2025)
di: Sadanandan, Binesh, et al.
Pubblicazione: (2025)
Consistent but Dangerous: Per-Sample Safety Classification Reveals False Reliability in Medical Vision-Language Models
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
Formalizing and Benchmarking Prompt Injection Attacks and Defenses
di: Liu, Yupei, et al.
Pubblicazione: (2023)
di: Liu, Yupei, et al.
Pubblicazione: (2023)
CAPTURE: Context-Aware Prompt Injection Testing and Robustness Enhancement
di: Kholkar, Gauri, et al.
Pubblicazione: (2025)
di: Kholkar, Gauri, et al.
Pubblicazione: (2025)
Scaling Behavior of Machine Translation with Large Language Models under Prompt Injection Attacks
di: Sun, Zhifan, et al.
Pubblicazione: (2024)
di: Sun, Zhifan, et al.
Pubblicazione: (2024)
Multilingual Hidden Prompt Injection Attacks on LLM-Based Academic Reviewing
di: Theocharopoulos, Panagiotis, et al.
Pubblicazione: (2025)
di: Theocharopoulos, Panagiotis, et al.
Pubblicazione: (2025)
WebInject: Prompt Injection Attack to Web Agents
di: Wang, Xilong, et al.
Pubblicazione: (2025)
di: Wang, Xilong, et al.
Pubblicazione: (2025)
PSF-Med: Measuring and Explaining Paraphrase Sensitivity in Medical Vision Language Models
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026)
VSF-Med:A Vulnerability Scoring Framework for Medical Vision-Language Models
di: Sadanandan, Binesh, et al.
Pubblicazione: (2025)
di: Sadanandan, Binesh, et al.
Pubblicazione: (2025)
Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
di: Mehrotra, Anay, et al.
Pubblicazione: (2023)
di: Mehrotra, Anay, et al.
Pubblicazione: (2023)
An Early Categorization of Prompt Injection Attacks on Large Language Models
di: Rossi, Sippo, et al.
Pubblicazione: (2024)
di: Rossi, Sippo, et al.
Pubblicazione: (2024)
(Im)possibility of Automated Hallucination Detection in Large Language Models
di: Karbasi, Amin, et al.
Pubblicazione: (2025)
di: Karbasi, Amin, et al.
Pubblicazione: (2025)
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language Models
di: Yi, Jingwei, et al.
Pubblicazione: (2023)
di: Yi, Jingwei, et al.
Pubblicazione: (2023)
Adversarial Attacks on LLM-as-a-Judge Systems: Insights from Prompt Injections
di: Maloyan, Narek, et al.
Pubblicazione: (2025)
di: Maloyan, Narek, et al.
Pubblicazione: (2025)
MPIB: A Benchmark for Medical Prompt Injection Attacks and Clinical Safety in LLMs
di: Lee, Junhyeok, et al.
Pubblicazione: (2026)
di: Lee, Junhyeok, et al.
Pubblicazione: (2026)
Harnessing Task Overload for Scalable Jailbreak Attacks on Large Language Models
di: Dong, Yiting, et al.
Pubblicazione: (2024)
di: Dong, Yiting, et al.
Pubblicazione: (2024)
Goal-guided Generative Prompt Injection Attack on Large Language Models
di: Zhang, Chong, et al.
Pubblicazione: (2024)
di: Zhang, Chong, et al.
Pubblicazione: (2024)
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
di: Wang, Xilong, et al.
Pubblicazione: (2026)
di: Wang, Xilong, et al.
Pubblicazione: (2026)
Securing Large Language Models (LLMs) from Prompt Injection Attacks
di: Suri, Omar Farooq Khan, et al.
Pubblicazione: (2025)
di: Suri, Omar Farooq Khan, et al.
Pubblicazione: (2025)
Led to Mislead: Adversarial Content Injection for Attacks on Neural Ranking Models
di: Bigdeli, Amin, et al.
Pubblicazione: (2026)
di: Bigdeli, Amin, et al.
Pubblicazione: (2026)
Visual Contextual Attack: Jailbreaking MLLMs with Image-Driven Context Injection
di: Miao, Ziqi, et al.
Pubblicazione: (2025)
di: Miao, Ziqi, et al.
Pubblicazione: (2025)
Fine-tuned Large Language Models (LLMs): Improved Prompt Injection Attacks Detection
di: Rahman, Md Abdur, et al.
Pubblicazione: (2024)
di: Rahman, Md Abdur, et al.
Pubblicazione: (2024)
FATH: Authentication-based Test-time Defense against Indirect Prompt Injection Attacks
di: Wang, Jiongxiao, et al.
Pubblicazione: (2024)
di: Wang, Jiongxiao, et al.
Pubblicazione: (2024)
Enhancing Prompt Injection Attacks to LLMs via Poisoning Alignment
di: Shao, Zedian, et al.
Pubblicazione: (2024)
di: Shao, Zedian, et al.
Pubblicazione: (2024)
HalluciNot: Hallucination Detection Through Context and Common Knowledge Verification
di: Paudel, Bibek, et al.
Pubblicazione: (2025)
di: Paudel, Bibek, et al.
Pubblicazione: (2025)
Human-Interpretable Adversarial Prompt Attack on Large Language Models with Situational Context
di: Das, Nilanjana, et al.
Pubblicazione: (2024)
di: Das, Nilanjana, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs
di: Upadhayay, Bibek, et al.
Pubblicazione: (2024) -
TaCo: Enhancing Cross-Lingual Transfer for Low-Resource Languages in LLMs through Translation-Assisted Chain-of-Thought Processes
di: Upadhayay, Bibek, et al.
Pubblicazione: (2023) -
X-Guard: Multilingual Guard Agent for Content Moderation
di: Upadhayay, Bibek, et al.
Pubblicazione: (2025) -
When Chain-of-Thought Backfires: Evaluating Prompt Sensitivity in Medical Language Models
di: Sadanandan, Binesh, et al.
Pubblicazione: (2026) -
Prompt Injection Attacks in Defended Systems
di: Khomsky, Daniil, et al.
Pubblicazione: (2024)