Misleading Large Language Models used (or misused) in Scientific Peer-Reviewing via Hidden Prompt-Injection Attacks
Fuente:
arXiv
Saved in:
| Main Authors: | Collu, Matteo Gioele, Salviati, Umberto, Confalonieri, Roberto, Conti, Mauro, Apruzzese, Giovanni |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations
by: Collu, Matteo Gioele, et al.
Published: (2026)
by: Collu, Matteo Gioele, et al.
Published: (2026)
The Hidden Threat in Plain Text: Attacking RAG Data Loaders
by: Castagnaro, Alberto, et al.
Published: (2025)
by: Castagnaro, Alberto, et al.
Published: (2025)
Exploiting AI for Attacks: On the Interplay between Adversarial AI and Offensive AI
by: Schröer, Saskia Laura, et al.
Published: (2025)
by: Schröer, Saskia Laura, et al.
Published: (2025)
"Are Adversarial Phishing Webpages a Threat in Reality?" Understanding the Users' Perception of Adversarial Webpages
by: Yuan, Ying, et al.
Published: (2024)
by: Yuan, Ying, et al.
Published: (2024)
SoK: Reshaping Research on Network Intrusion Detection Systems
by: Apruzzese, Giovanni
Published: (2026)
by: Apruzzese, Giovanni
Published: (2026)
I can't recognize (yet): Delayed Rendering to Defeat Visual Phishing Detectors
by: Yuan, Ying, et al.
Published: (2026)
by: Yuan, Ying, et al.
Published: (2026)
E-PhishGen: Unlocking Novel Research in Phishing Email Detection
by: Pajola, Luca, et al.
Published: (2025)
by: Pajola, Luca, et al.
Published: (2025)
Dr. Jekyll and Mr. Hyde: Two Faces of LLMs
by: Collu, Matteo Gioele, et al.
Published: (2023)
by: Collu, Matteo Gioele, et al.
Published: (2023)
Offensive AI: Enhancing Directory Brute-forcing Attack with the Use of Language Models
by: Castagnaro, Alberto, et al.
Published: (2024)
by: Castagnaro, Alberto, et al.
Published: (2024)
AttackEval: A Systematic Empirical Study of Prompt Injection Attack Effectiveness Against Large Language Models
by: Wang, Jackson
Published: (2026)
by: Wang, Jackson
Published: (2026)
Leaky Batteries: A Novel Set of Side-Channel Attacks on Electric Vehicles
by: Marchiori, Francesco, et al.
Published: (2025)
by: Marchiori, Francesco, et al.
Published: (2025)
Can SOC Operators Explain their Decisions while Triaging Alarms? A Real-World Study
by: Moosmann, Jessica, et al.
Published: (2026)
by: Moosmann, Jessica, et al.
Published: (2026)
The Ephemeral Threat: Assessing the Security of Algorithmic Trading Systems powered by Deep Learning
by: Rizvani, Advije, et al.
Published: (2025)
by: Rizvani, Advije, et al.
Published: (2025)
Department-Specific Security Awareness Campaigns: A Cross-Organizational Study of HR and Accounting
by: Pfister, Matthias, et al.
Published: (2025)
by: Pfister, Matthias, et al.
Published: (2025)
The Hidden Dangers of Public Serverless Repositories: An Empirical Security Assessment
by: Marin, Eduard, et al.
Published: (2025)
by: Marin, Eduard, et al.
Published: (2025)
PromptShield: Deployable Detection for Prompt Injection Attacks
by: Jacob, Dennis, et al.
Published: (2025)
by: Jacob, Dennis, et al.
Published: (2025)
When Your Reviewer is an LLM: Biases, Divergence, and Prompt Injection Risks in Peer Review
by: Zhu, Changjia, et al.
Published: (2025)
by: Zhu, Changjia, et al.
Published: (2025)
QUACK! Making the (Rubber) Ducky Talk: A Systematic Study of Keystroke Dynamics for HID Injection Detection
by: Lotto, Alessandro, et al.
Published: (2026)
by: Lotto, Alessandro, et al.
Published: (2026)
Prompt Injection Attacks on Large Language Models in Oncology
by: Clusmann, Jan, et al.
Published: (2024)
by: Clusmann, Jan, et al.
Published: (2024)
A Novel Evaluation Framework for Assessing Resilience Against Prompt Injection Attacks in Large Language Models
by: Yip, Daniel Wankit, et al.
Published: (2024)
by: Yip, Daniel Wankit, et al.
Published: (2024)
Defense Against Prompt Injection Attack by Leveraging Attack Techniques
by: Chen, Yulin, et al.
Published: (2024)
by: Chen, Yulin, et al.
Published: (2024)
The Vulnerability of LLM Rankers to Prompt Injection Attacks
by: Yin, Yu, et al.
Published: (2026)
by: Yin, Yu, et al.
Published: (2026)
Inference Attacks on Encrypted Online Voting via Traffic Analysis
by: Belousova, Anastasiia, et al.
Published: (2025)
by: Belousova, Anastasiia, et al.
Published: (2025)
TopicAttack: An Indirect Prompt Injection Attack via Topic Transition
by: Chen, Yulin, et al.
Published: (2025)
by: Chen, Yulin, et al.
Published: (2025)
Towards a Self-rescuing System for UAVs Under GNSS Attack
by: Rigoni, Giulio, et al.
Published: (2024)
by: Rigoni, Giulio, et al.
Published: (2024)
An Early Categorization of Prompt Injection Attacks on Large Language Models
by: Rossi, Sippo, et al.
Published: (2024)
by: Rossi, Sippo, et al.
Published: (2024)
KeTS: Kernel-based Trust Segmentation against Model Poisoning Attacks
by: Gangwal, Ankit, et al.
Published: (2025)
by: Gangwal, Ankit, et al.
Published: (2025)
Prompt Injection Attack to Tool Selection in LLM Agents
by: Shi, Jiawen, et al.
Published: (2025)
by: Shi, Jiawen, et al.
Published: (2025)
PINA: Prompt Injection Attack against Navigation Agents
by: Liu, Jiani, et al.
Published: (2026)
by: Liu, Jiani, et al.
Published: (2026)
Can Indirect Prompt Injection Attacks Be Detected and Removed?
by: Chen, Yulin, et al.
Published: (2025)
by: Chen, Yulin, et al.
Published: (2025)
Through the Static: Demystifying Malware Visualization via Explainability
by: Brosolo, Matteo, et al.
Published: (2025)
by: Brosolo, Matteo, et al.
Published: (2025)
The Road Less Traveled: Investigating Robustness and Explainability in CNN Malware Detection
by: Brosolo, Matteo, et al.
Published: (2025)
by: Brosolo, Matteo, et al.
Published: (2025)
System Prompt Poisoning: Persistent Attacks on Large Language Models Beyond User Injection
by: Li, Zongze, et al.
Published: (2025)
by: Li, Zongze, et al.
Published: (2025)
Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
by: Clop, Cody, et al.
Published: (2024)
by: Clop, Cody, et al.
Published: (2024)
PromptLocate: Localizing Prompt Injection Attacks
by: Jia, Yuqi, et al.
Published: (2025)
by: Jia, Yuqi, et al.
Published: (2025)
Work-in-Progress: Crash Course: Can (Under Attack) Autonomous Driving Beat Human Drivers?
by: Marchiori, Francesco, et al.
Published: (2024)
by: Marchiori, Francesco, et al.
Published: (2024)
The Impact of Emerging Phishing Threats: Assessing Quishing and LLM-generated Phishing Emails against Organizations
by: Weinz, Marie, et al.
Published: (2025)
by: Weinz, Marie, et al.
Published: (2025)
Backdoor-Powered Prompt Injection Attacks Nullify Defense Methods
by: Chen, Yulin, et al.
Published: (2025)
by: Chen, Yulin, et al.
Published: (2025)
AlignSentinel: Alignment-Aware Detection of Prompt Injection Attacks
by: Jia, Yuqi, et al.
Published: (2026)
by: Jia, Yuqi, et al.
Published: (2026)
Adversarial News and Lost Profits: Manipulating Headlines in LLM-Driven Algorithmic Trading
by: Rizvani, Advije, et al.
Published: (2026)
by: Rizvani, Advije, et al.
Published: (2026)
Similar Items
-
Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations
by: Collu, Matteo Gioele, et al.
Published: (2026) -
The Hidden Threat in Plain Text: Attacking RAG Data Loaders
by: Castagnaro, Alberto, et al.
Published: (2025) -
Exploiting AI for Attacks: On the Interplay between Adversarial AI and Offensive AI
by: Schröer, Saskia Laura, et al.
Published: (2025) -
"Are Adversarial Phishing Webpages a Threat in Reality?" Understanding the Users' Perception of Adversarial Webpages
by: Yuan, Ying, et al.
Published: (2024) -
SoK: Reshaping Research on Network Intrusion Detection Systems
by: Apruzzese, Giovanni
Published: (2026)