Tricking LLM-Based NPCs into Spilling Secrets
Fuente:
arXiv
Saved in:
| Main Authors: | Shiomi, Kyohei, Lian, Zhuotao, Nakanishi, Toru, Kitasuka, Teruaki |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prompt-in-Content Attacks: Exploiting Uploaded Inputs to Hijack LLM Behavior
by: Lian, Zhuotao, et al.
Published: (2025)
by: Lian, Zhuotao, et al.
Published: (2025)
Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs
by: Xu, Zhao, et al.
Published: (2024)
by: Xu, Zhao, et al.
Published: (2024)
Secret Stealing Attacks on Local LLM Fine-Tuning through Supply-Chain Model Code Backdoors
by: Li, Zi, et al.
Published: (2026)
by: Li, Zi, et al.
Published: (2026)
SafeSearch: Automated Red-Teaming of LLM-Based Search Agents
by: Dong, Jianshuo, et al.
Published: (2025)
by: Dong, Jianshuo, et al.
Published: (2025)
Narrow Secret Loyalty Dodges Black-Box Audits
by: Lamerton, Alfie, et al.
Published: (2026)
by: Lamerton, Alfie, et al.
Published: (2026)
SUDP: Secret-Use Delegation Protocol for Agentic Systems
by: Yu, Xiaohang, et al.
Published: (2026)
by: Yu, Xiaohang, et al.
Published: (2026)
Hidden Thoughts Are Not Secret: Reasoning Trace Exposure in LLMs
by: Lu, Yu-An, et al.
Published: (2026)
by: Lu, Yu-An, et al.
Published: (2026)
Context-Aware Hierarchical Learning: A Two-Step Paradigm towards Safer LLMs
by: Ma, Tengyun, et al.
Published: (2025)
by: Ma, Tengyun, et al.
Published: (2025)
Towards Fine-Grained Webpage Fingerprinting at Scale
by: Zhao, Xiyuan, et al.
Published: (2024)
by: Zhao, Xiyuan, et al.
Published: (2024)
When Speculation Spills Secrets: Side Channels via Speculative Decoding In LLMs
by: Wei, Jiankun, et al.
Published: (2024)
by: Wei, Jiankun, et al.
Published: (2024)
Understanding Secret Leakage Risks in Code LLMs: A Tokenization Perspective
by: Chen, Meifang, et al.
Published: (2026)
by: Chen, Meifang, et al.
Published: (2026)
CapSeal: Capability-Sealed Secret Mediation for Secure Agent Execution
by: Jin, Shutong, et al.
Published: (2026)
by: Jin, Shutong, et al.
Published: (2026)
Towards Robust Multi-tab Website Fingerprinting
by: Deng, Xinhao, et al.
Published: (2025)
by: Deng, Xinhao, et al.
Published: (2025)
Secret Collusion among AI Agents: Multi-Agent Deception via Steganography
by: Motwani, Sumeet Ramesh, et al.
Published: (2024)
by: Motwani, Sumeet Ramesh, et al.
Published: (2024)
When Agents Handle Secrets: A Survey of Confidential Computing for Agentic AI
by: Forough, Javad, et al.
Published: (2026)
by: Forough, Javad, et al.
Published: (2026)
A Middle Path for On-Premises LLM Deployment: Preserving Privacy Without Sacrificing Model Confidentiality
by: Huang, Hanbo, et al.
Published: (2024)
by: Huang, Hanbo, et al.
Published: (2024)
Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing
by: Holtzman, Ari, et al.
Published: (2026)
by: Holtzman, Ari, et al.
Published: (2026)
Enhancing Scalability of Metric Differential Privacy via Secret Dataset Partitioning and Benders Decomposition
by: Qiu, Chenxi
Published: (2024)
by: Qiu, Chenxi
Published: (2024)
Explainable Adversarial Learning Framework on Physical Layer Secret Keys Combating Malicious Reconfigurable Intelligent Surface
by: Wei, Zhuangkun, et al.
Published: (2024)
by: Wei, Zhuangkun, et al.
Published: (2024)
From Data Leak to Secret Misses: The Impact of Data Leakage on Secret Detection Models
by: Soltaniani, Farnaz, et al.
Published: (2026)
by: Soltaniani, Farnaz, et al.
Published: (2026)
LLM-Powered Intent-Based Categorization of Phishing Emails
by: Eilertsen, Even, et al.
Published: (2025)
by: Eilertsen, Even, et al.
Published: (2025)
Information Security Based on LLM Approaches: A Review
by: Gong, Chang, et al.
Published: (2025)
by: Gong, Chang, et al.
Published: (2025)
Targeted Bit-Flip Attacks on LLM-Based Agents
by: Wang, Jialai, et al.
Published: (2026)
by: Wang, Jialai, et al.
Published: (2026)
Leveraging LLM to Strengthen ML-Based Cross-Site Scripting Detection
by: Miczek, Dennis, et al.
Published: (2025)
by: Miczek, Dennis, et al.
Published: (2025)
DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Spill The Beans: Exploiting CPU Cache Side-Channels to Leak Tokens from Large Language Models
by: Adiletta, Andrew, et al.
Published: (2025)
by: Adiletta, Andrew, et al.
Published: (2025)
A-MemGuard: A Proactive Defense Framework for LLM-Based Agent Memory
by: Wei, Qianshan, et al.
Published: (2025)
by: Wei, Qianshan, et al.
Published: (2025)
AttackPilot: Autonomous Inference Attacks Against ML Services With LLM-Based Agents
by: Wu, Yixin, et al.
Published: (2025)
by: Wu, Yixin, et al.
Published: (2025)
PrivacyPeek: Auditing What LLM-Based Agents Acquire, Not Just What They Say
by: Zhang, Mingxuan, et al.
Published: (2026)
by: Zhang, Mingxuan, et al.
Published: (2026)
SmartLLMSentry: A Comprehensive LLM Based Smart Contract Vulnerability Detection Framework
by: Zaazaa, Oualid, et al.
Published: (2024)
by: Zaazaa, Oualid, et al.
Published: (2024)
Llama-3.1-FoundationAI-SecurityLLM-Base-8B Technical Report
by: Kassianik, Paul, et al.
Published: (2025)
by: Kassianik, Paul, et al.
Published: (2025)
Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations
by: Shafranskyi, Daniil, et al.
Published: (2026)
by: Shafranskyi, Daniil, et al.
Published: (2026)
PathSeeker: Exploring LLM Security Vulnerabilities with a Reinforcement Learning-Based Jailbreak Approach
by: Lin, Zhihao, et al.
Published: (2024)
by: Lin, Zhihao, et al.
Published: (2024)
RedacBench: Can AI Erase Your Secrets?
by: Jeon, Hyunjun, et al.
Published: (2026)
by: Jeon, Hyunjun, et al.
Published: (2026)
Contextualized Privacy Defense for LLM Agents
by: Wen, Yule, et al.
Published: (2026)
by: Wen, Yule, et al.
Published: (2026)
Advancing LLM-Based Security Automation with Customized Group Relative Policy Optimization for Zero-Touch Networks
by: Cao, Xinye, et al.
Published: (2025)
by: Cao, Xinye, et al.
Published: (2025)
WebWeaver: Breaking Topology Confidentiality in LLM Multi-Agent Systems with Stealthy Context-Based Inference
by: Xiong, Zixun, et al.
Published: (2026)
by: Xiong, Zixun, et al.
Published: (2026)
Benchmarking LLM-Based Static Analysis for Secure Smart Contract Development: Reliability, Limitations, and Potential Hybrid Solutions
by: Susan, Stefan-Claudiu, et al.
Published: (2026)
by: Susan, Stefan-Claudiu, et al.
Published: (2026)
CheatAgent: Attacking LLM-Empowered Recommender Systems via LLM Agent
by: Ning, Liang-bo, et al.
Published: (2025)
by: Ning, Liang-bo, et al.
Published: (2025)
LLM Access Shield: Domain-Specific LLM Framework for Privacy Policy Compliance
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
Similar Items
-
Prompt-in-Content Attacks: Exploiting Uploaded Inputs to Hijack LLM Behavior
by: Lian, Zhuotao, et al.
Published: (2025) -
Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs
by: Xu, Zhao, et al.
Published: (2024) -
Secret Stealing Attacks on Local LLM Fine-Tuning through Supply-Chain Model Code Backdoors
by: Li, Zi, et al.
Published: (2026) -
SafeSearch: Automated Red-Teaming of LLM-Based Search Agents
by: Dong, Jianshuo, et al.
Published: (2025) -
Narrow Secret Loyalty Dodges Black-Box Audits
by: Lamerton, Alfie, et al.
Published: (2026)