PII-Compass: Guiding LLM training data extraction prompts towards the target PII via grounding
Fuente:
arXiv
Saved in:
| Main Authors: | Nakka, Krishna Kanth, Frikha, Ahmed, Mendes, Ricardo, Jiang, Xue, Zhou, Xuebing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PII Jailbreaking in LLMs via Activation Steering Reveals Personal Information Leakage
by: Nakka, Krishna Kanth, et al.
Published: (2025)
by: Nakka, Krishna Kanth, et al.
Published: (2025)
PII-Scope: A Comprehensive Study on Training Data PII Extraction Attacks in LLMs
by: Nakka, Krishna Kanth, et al.
Published: (2024)
by: Nakka, Krishna Kanth, et al.
Published: (2024)
IncogniText: Privacy-enhancing Conditional Text Anonymization via LLM-based Private Attribute Randomization
by: Frikha, Ahmed, et al.
Published: (2024)
by: Frikha, Ahmed, et al.
Published: (2024)
ObfuscaTune: Obfuscated Offsite Fine-tuning and Inference of Proprietary LLMs on Private Datasets
by: Frikha, Ahmed, et al.
Published: (2024)
by: Frikha, Ahmed, et al.
Published: (2024)
WebPII: Benchmarking Visual PII Detection for Computer-Use Agents
by: Zhao, Nathan
Published: (2026)
by: Zhao, Nathan
Published: (2026)
Measuring the Accuracy and Effectiveness of PII Removal Services
by: He, Jiahui, et al.
Published: (2025)
by: He, Jiahui, et al.
Published: (2025)
Membership Inference for Contrastive Pre-training Models with Text-only PII Queries
by: Cheng, Ruoxi, et al.
Published: (2026)
by: Cheng, Ruoxi, et al.
Published: (2026)
Merger-as-a-Stealer: Stealing Targeted PII from Aligned LLMs with Model Merging
by: Lu, Lin, et al.
Published: (2025)
by: Lu, Lin, et al.
Published: (2025)
PRvL: Quantifying the Capabilities and Risks of Large Language Models for PII Redaction
by: Garza, Leon, et al.
Published: (2025)
by: Garza, Leon, et al.
Published: (2025)
CAPID: Context-Aware PII Detection for Question-Answering Systems
by: Ponomarenko, Mariia, et al.
Published: (2026)
by: Ponomarenko, Mariia, et al.
Published: (2026)
Discovering Universal Activation Directions for PII Leakage in Language Models
by: Marchyok, Leo, et al.
Published: (2026)
by: Marchyok, Leo, et al.
Published: (2026)
Adaptive PII Mitigation Framework for Large Language Models
by: Asthana, Shubhi, et al.
Published: (2025)
by: Asthana, Shubhi, et al.
Published: (2025)
PIIGuard: Mitigating PII Harvesting under Adversarial Sanitization
by: Liu, Mingshuo, et al.
Published: (2026)
by: Liu, Mingshuo, et al.
Published: (2026)
SEAL-Tag: Self-Tag Evidence Aggregation with Probabilistic Circuits for PII-Safe Retrieval-Augmented Generation
by: Xie, Jin, et al.
Published: (2026)
by: Xie, Jin, et al.
Published: (2026)
PrivacyScalpel: Enhancing LLM Privacy via Interpretable Feature Intervention with Sparse Autoencoders
by: Frikha, Ahmed, et al.
Published: (2025)
by: Frikha, Ahmed, et al.
Published: (2025)
SecureGate: Learning When to Reveal PII Safely via Token-Gated Dual-Adapters for Federated LLMs
by: Shaaban, Mohamed, et al.
Published: (2026)
by: Shaaban, Mohamed, et al.
Published: (2026)
PATCH: Mitigating PII Leakage in Language Models with Privacy-Aware Targeted Circuit PatcHing
by: Hughes, Anthony, et al.
Published: (2025)
by: Hughes, Anthony, et al.
Published: (2025)
PII-Bench: Evaluating Query-Aware Privacy Protection Systems
by: Shen, Hao, et al.
Published: (2025)
by: Shen, Hao, et al.
Published: (2025)
Model Inversion Attacks on Llama 3: Extracting PII from Large Language Models
by: Sivashanmugam, Sathesh P.
Published: (2025)
by: Sivashanmugam, Sathesh P.
Published: (2025)
AnonLFI 2.0: Extensible Architecture for PII Pseudonymization in CSIRTs with OCR and Technical Recognizers
by: Kapelinski, Cristhian, et al.
Published: (2025)
by: Kapelinski, Cristhian, et al.
Published: (2025)
"I Strongly Suspect This Website Is a Scam": Benchmarking PII Leakage and Detection without Defense in Autonomous Web Agents
by: Roy, Soham, et al.
Published: (2026)
by: Roy, Soham, et al.
Published: (2026)
UnPII: Unlearning Personally Identifiable Information with Quantifiable Exposure Risk
by: Jeon, Intae, et al.
Published: (2026)
by: Jeon, Intae, et al.
Published: (2026)
A Hardware-Anchored Privacy Middleware for PII Sharing Across Heterogeneous Embedded Consumer Devices
by: Sabbineni, Aditya, et al.
Published: (2026)
by: Sabbineni, Aditya, et al.
Published: (2026)
VisualLeakBench: Auditing the Fragility of Large Vision-Language Models against PII Leakage and Social Engineering
by: Wang, Youting, et al.
Published: (2026)
by: Wang, Youting, et al.
Published: (2026)
BitBypass: A New Direction in Jailbreaking Aligned Large Language Models with Bitstream Camouflage
by: Nakka, Kalyan, et al.
Published: (2025)
by: Nakka, Kalyan, et al.
Published: (2025)
PII-VisBench: Evaluating Personally Identifiable Information Safety in Vision Language Models Along a Continuum of Visibility
by: Shahariar, G M, et al.
Published: (2026)
by: Shahariar, G M, et al.
Published: (2026)
Quantum Key Distribution Routing Protocol in Quantum Networks: Overview and Challenges
by: Kumar, Pankaj, et al.
Published: (2024)
by: Kumar, Pankaj, et al.
Published: (2024)
A Machine Learning-Based Framework for Assessing Cryptographic Indistinguishability of Lightweight Block Ciphers
by: Dani, Jimmy, et al.
Published: (2024)
by: Dani, Jimmy, et al.
Published: (2024)
Is On-Device AI Broken and Exploitable? Assessing the Trust and Ethics in Small Language Models
by: Nakka, Kalyan, et al.
Published: (2024)
by: Nakka, Kalyan, et al.
Published: (2024)
LiteLMGuard: Seamless and Lightweight On-Device Prompt Filtering for Safeguarding Small Language Models against Quantization-induced Risks and Vulnerabilities
by: Nakka, Kalyan, et al.
Published: (2025)
by: Nakka, Kalyan, et al.
Published: (2025)
Dynamic Quantum Key Distribution for Microgrids with Distributed Error Correction
by: Rath, Suman, et al.
Published: (2024)
by: Rath, Suman, et al.
Published: (2024)
Innovative tokenisation of structured data for LLM training
by: Karim, Kayvan, et al.
Published: (2025)
by: Karim, Kayvan, et al.
Published: (2025)
Hardening x402: PII-Safe Agentic Payments via Pre-Execution Metadata Filtering
by: Stantchev, Vladimir
Published: (2026)
by: Stantchev, Vladimir
Published: (2026)
A Comprehensive Content Verification System for ensuring Digital Integrity in the Age of Deep Fakes
by: Kaja, RaviKanth
Published: (2024)
by: Kaja, RaviKanth
Published: (2024)
Feedback-Guided Extraction of Knowledge Base from Retrieval-Augmented LLM Applications
by: Jiang, Changyue, et al.
Published: (2024)
by: Jiang, Changyue, et al.
Published: (2024)
A LINDDUN-based Privacy Threat Modeling Framework for GenAI
by: Liao, Qianying, et al.
Published: (2026)
by: Liao, Qianying, et al.
Published: (2026)
TeleAI-Safety: A comprehensive LLM jailbreaking benchmark towards attacks, defenses, and evaluations
by: Chen, Xiuyuan, et al.
Published: (2025)
by: Chen, Xiuyuan, et al.
Published: (2025)
Comparison of feature extraction tools for network traffic data
by: Lypa, Borys, et al.
Published: (2025)
by: Lypa, Borys, et al.
Published: (2025)
Reconstructing training data from document understanding models
by: Dentan, Jérémie, et al.
Published: (2024)
by: Dentan, Jérémie, et al.
Published: (2024)
Don't Trust Your Upstream: Exploiting LLM Multi-Agent System via Topology-Guided Adversarial Propagation
by: Liang, Ruichao, et al.
Published: (2025)
by: Liang, Ruichao, et al.
Published: (2025)
Similar Items
-
PII Jailbreaking in LLMs via Activation Steering Reveals Personal Information Leakage
by: Nakka, Krishna Kanth, et al.
Published: (2025) -
PII-Scope: A Comprehensive Study on Training Data PII Extraction Attacks in LLMs
by: Nakka, Krishna Kanth, et al.
Published: (2024) -
IncogniText: Privacy-enhancing Conditional Text Anonymization via LLM-based Private Attribute Randomization
by: Frikha, Ahmed, et al.
Published: (2024) -
ObfuscaTune: Obfuscated Offsite Fine-tuning and Inference of Proprietary LLMs on Private Datasets
by: Frikha, Ahmed, et al.
Published: (2024) -
WebPII: Benchmarking Visual PII Detection for Computer-Use Agents
by: Zhao, Nathan
Published: (2026)