Towards Contextual Sensitive Data Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Telkamp, Liang, Hulsebos, Madelon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TARGET: Benchmarking Table Retrieval for Generative Tasks
by: Ji, Xingyu, et al.
Published: (2025)
by: Ji, Xingyu, et al.
Published: (2025)
Fine-Grained Table Retrieval Through the Lens of Complex Queries
by: Kosiuk, Wojciech, et al.
Published: (2026)
by: Kosiuk, Wojciech, et al.
Published: (2026)
Token-wise Influential Training Data Retrieval for Large Language Models
by: Lin, Huawei, et al.
Published: (2024)
by: Lin, Huawei, et al.
Published: (2024)
PII-Bench: Evaluating Query-Aware Privacy Protection Systems
by: Shen, Hao, et al.
Published: (2025)
by: Shen, Hao, et al.
Published: (2025)
TITAN: Graph-Executable Reasoning for Cyber Threat Intelligence
by: Simoni, Marco, et al.
Published: (2025)
by: Simoni, Marco, et al.
Published: (2025)
DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box Retrieval-Augmented Generation
by: Gong, Yuyang, et al.
Published: (2026)
by: Gong, Yuyang, et al.
Published: (2026)
Improving User Privacy in Personalized Generation: Client-Side Retrieval-Augmented Modification of Server-Side Generated Speculations
by: Salemi, Alireza, et al.
Published: (2026)
by: Salemi, Alireza, et al.
Published: (2026)
One Single Hub Text Breaks CLIP: Identifying Vulnerabilities in Cross-Modal Encoders via Hubness
by: Deguchi, Hiroyuki, et al.
Published: (2026)
by: Deguchi, Hiroyuki, et al.
Published: (2026)
SoK: Agentic Retrieval-Augmented Generation (RAG): Taxonomy, Architectures, Evaluation, and Research Directions
by: Mishra, Saroj, et al.
Published: (2026)
by: Mishra, Saroj, et al.
Published: (2026)
ChineseHarm-Bench: A Chinese Harmful Content Detection Benchmark
by: Liu, Kangwei, et al.
Published: (2025)
by: Liu, Kangwei, et al.
Published: (2025)
Towards Copyright Protection for Knowledge Bases of Retrieval-augmented Language Models via Reasoning
by: Guo, Junfeng, et al.
Published: (2025)
by: Guo, Junfeng, et al.
Published: (2025)
TableGuard -- Securing Structured & Unstructured Data
by: Sharma, Anantha, et al.
Published: (2024)
by: Sharma, Anantha, et al.
Published: (2024)
Nautilus Compass: Black-box Persona Drift Detection for Production LLM Agents
by: Wang, Chunxiao
Published: (2026)
by: Wang, Chunxiao
Published: (2026)
Are We Asking the Right Questions? On Ambiguity in Natural Language Queries for Tabular Data Analysis
by: Gomm, Daniel, et al.
Published: (2025)
by: Gomm, Daniel, et al.
Published: (2025)
RAID: An In-Training Defense against Attribute Inference Attacks in Recommender Systems
by: Feng, Xiaohua, et al.
Published: (2025)
by: Feng, Xiaohua, et al.
Published: (2025)
RCVaR: an Economic Approach to Estimate Cyberattacks Costs using Data from Industry Reports
by: Franco, Muriel Figueredo, et al.
Published: (2023)
by: Franco, Muriel Figueredo, et al.
Published: (2023)
Riddle Me This! Stealthy Membership Inference for Retrieval-Augmented Generation
by: Naseh, Ali, et al.
Published: (2025)
by: Naseh, Ali, et al.
Published: (2025)
SQuaD: The Software Quality Dataset
by: Robredo, Mikel, et al.
Published: (2025)
by: Robredo, Mikel, et al.
Published: (2025)
Membership Inference Attacks on LLM-based Recommender Systems
by: He, Jiajie, et al.
Published: (2025)
by: He, Jiajie, et al.
Published: (2025)
Grounded Cache Routing for Retrieval-Augmented Generation: When Is It Safe to Reuse an Answer?
by: Shah, Syed Huma
Published: (2026)
by: Shah, Syed Huma
Published: (2026)
User-Side Realization
by: Sato, Ryoma
Published: (2024)
by: Sato, Ryoma
Published: (2024)
EmojiPrompt: Generative Prompt Obfuscation for Privacy-Preserving Communication with Cloud-based LLMs
by: Lin, Sam, et al.
Published: (2024)
by: Lin, Sam, et al.
Published: (2024)
BadRAG: Identifying Vulnerabilities in Retrieval Augmented Generation of Large Language Models
by: Xue, Jiaqi, et al.
Published: (2024)
by: Xue, Jiaqi, et al.
Published: (2024)
Measuring Technological Convergence in Encryption Technologies with Proximity Indices: A Text Mining and Bibliometric Analysis using OpenAlex
by: Tavazzi, Alessandro, et al.
Published: (2024)
by: Tavazzi, Alessandro, et al.
Published: (2024)
Taxonomy Inference for Tabular Data Using Large Language Models
by: Wu, Zhenyu, et al.
Published: (2025)
by: Wu, Zhenyu, et al.
Published: (2025)
DRAMA: Unifying Data Retrieval and Analysis for Open-Domain Analytic Queries
by: Hu, Chuxuan, et al.
Published: (2025)
by: Hu, Chuxuan, et al.
Published: (2025)
NormTab: Improving Symbolic Reasoning in LLMs Through Tabular Data Normalization
by: Nahid, Md Mahadi Hasan, et al.
Published: (2024)
by: Nahid, Md Mahadi Hasan, et al.
Published: (2024)
MemPot: Defending Against Memory Extraction Attack with Optimized Honeypots
by: Wang, Yuhao, et al.
Published: (2026)
by: Wang, Yuhao, et al.
Published: (2026)
Toward Autonomous SOC Operations: End-to-End LLM Framework for Threat Detection, Query Generation, and Resolution in Security Operations
by: Saju, Md Hasan, et al.
Published: (2026)
by: Saju, Md Hasan, et al.
Published: (2026)
Personalized w-Event Privacy for Infinite Stream Estimation
by: Du, Leilei, et al.
Published: (2026)
by: Du, Leilei, et al.
Published: (2026)
SQaLe: A Large Text-to-SQL Corpus Grounded in Real Schemas
by: Wolff, Cornelius, et al.
Published: (2025)
by: Wolff, Cornelius, et al.
Published: (2025)
Exposing Citation Vulnerabilities in Generative Engines
by: Mochizuki, Riku, et al.
Published: (2025)
by: Mochizuki, Riku, et al.
Published: (2025)
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models
by: Gong, Yuyang, et al.
Published: (2025)
by: Gong, Yuyang, et al.
Published: (2025)
Retrieval-Augmented Review Generation for Poisoning Recommender Systems
by: Yang, Shiyi, et al.
Published: (2025)
by: Yang, Shiyi, et al.
Published: (2025)
Tuning for TraceTarnish: Techniques, Trends, and Testing Tangible Traits
by: Dilworth, Robert
Published: (2025)
by: Dilworth, Robert
Published: (2025)
Unveiling Unicode's Unseen Underpinnings in Undermining Authorship Attribution
by: Dilworth, Robert
Published: (2025)
by: Dilworth, Robert
Published: (2025)
A Decentralized Retrieval Augmented Generation System with Source Reliabilities Secured on Blockchain
by: Lu, Yining, et al.
Published: (2025)
by: Lu, Yining, et al.
Published: (2025)
StegoStylo: Squelching Stylometric Scrutiny through Steganographic Stitching
by: Dilworth, Robert
Published: (2026)
by: Dilworth, Robert
Published: (2026)
A Wolf in Sheep's Clothing: Targeted Routing Hijacking in Federated RAG
by: Mu, Junjie, et al.
Published: (2026)
by: Mu, Junjie, et al.
Published: (2026)
Hijacking Text Heritage: Hiding the Human Signature through Homoglyphic Substitution
by: Dilworth, Robert
Published: (2026)
by: Dilworth, Robert
Published: (2026)
Similar Items
-
TARGET: Benchmarking Table Retrieval for Generative Tasks
by: Ji, Xingyu, et al.
Published: (2025) -
Fine-Grained Table Retrieval Through the Lens of Complex Queries
by: Kosiuk, Wojciech, et al.
Published: (2026) -
Token-wise Influential Training Data Retrieval for Large Language Models
by: Lin, Huawei, et al.
Published: (2024) -
PII-Bench: Evaluating Query-Aware Privacy Protection Systems
by: Shen, Hao, et al.
Published: (2025) -
TITAN: Graph-Executable Reasoning for Cyber Threat Intelligence
by: Simoni, Marco, et al.
Published: (2025)