Local Obfuscation by GLINER for Impartial Context Aware Lineage: Development and evaluation of PII Removal system
Fuente:
arXiv
Saved in:
| Main Authors: | Shivaprakash, Prakrithi, Shukla, Lekhansh, Mukherjee, Animesh, Chand, Prabhat, Murthy, Pratima |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lost without translation -- Can transformer (language models) understand mood states?
by: Shivaprakash, Prakrithi, et al.
Published: (2025)
by: Shivaprakash, Prakrithi, et al.
Published: (2025)
Benchmarking Motivational Interviewing Competence of Large Language Models
by: Jha, Aishwariya, et al.
Published: (2026)
by: Jha, Aishwariya, et al.
Published: (2026)
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
by: Haritha Gireesh, et al.
Published: (2025)
by: Haritha Gireesh, et al.
Published: (2025)
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
by: Kumar, Subham, et al.
Published: (2025)
by: Kumar, Subham, et al.
Published: (2025)
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
by: Kumar, Subham, et al.
Published: (2025)
by: Kumar, Subham, et al.
Published: (2025)
CAPID: Context-Aware PII Detection for Question-Answering Systems
by: Ponomarenko, Mariia, et al.
Published: (2026)
by: Ponomarenko, Mariia, et al.
Published: (2026)
Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages
by: Banerjee, Somnath, et al.
Published: (2026)
by: Banerjee, Somnath, et al.
Published: (2026)
RA-MTR: A Retrieval Augmented Multi-Task Reader based Approach for Inspirational Quote Extraction from Long Documents
by: Adak, Sayantan, et al.
Published: (2025)
by: Adak, Sayantan, et al.
Published: (2025)
Comparing Feature-based and Context-aware Approaches to PII Generalization Level Prediction
by: Zhang, Kailin, et al.
Published: (2024)
by: Zhang, Kailin, et al.
Published: (2024)
Context Matters: Pushing the Boundaries of Open-Ended Answer Generation with Graph-Structured Knowledge Context
by: Banerjee, Somnath, et al.
Published: (2024)
by: Banerjee, Somnath, et al.
Published: (2024)
SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models
by: Banerjee, Somnath, et al.
Published: (2024)
by: Banerjee, Somnath, et al.
Published: (2024)
PII-Scope: A Comprehensive Study on Training Data PII Extraction Attacks in LLMs
by: Nakka, Krishna Kanth, et al.
Published: (2024)
by: Nakka, Krishna Kanth, et al.
Published: (2024)
PII-Bench: Evaluating Query-Aware Privacy Protection Systems
by: Shen, Hao, et al.
Published: (2025)
by: Shen, Hao, et al.
Published: (2025)
PATCH: Mitigating PII Leakage in Language Models with Privacy-Aware Targeted Circuit PatcHing
by: Hughes, Anthony, et al.
Published: (2025)
by: Hughes, Anthony, et al.
Published: (2025)
TEXT2AFFORD: Probing Object Affordance Prediction abilities of Language Models solely from Text
by: Adak, Sayantan, et al.
Published: (2024)
by: Adak, Sayantan, et al.
Published: (2024)
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs
by: Nag, Arijit, et al.
Published: (2024)
by: Nag, Arijit, et al.
Published: (2024)
CrowdCounter: A benchmark type-specific multi-target counterspeech dataset
by: Saha, Punyajoy, et al.
Published: (2024)
by: Saha, Punyajoy, et al.
Published: (2024)
Lost in Interpretation: The Plausibility-Faithfulness Trade-off in Cross-Lingual Explanations
by: Banerjee, Somnath, et al.
Published: (2026)
by: Banerjee, Somnath, et al.
Published: (2026)
MemeSense: An Adaptive In-Context Framework for Social Commonsense Driven Meme Moderation
by: Adak, Sayantan, et al.
Published: (2025)
by: Adak, Sayantan, et al.
Published: (2025)
Scalable multilingual PII annotation for responsible AI in LLMs
by: Meena, Bharti, et al.
Published: (2025)
by: Meena, Bharti, et al.
Published: (2025)
Locale-Conditioned Few-Shot Prompting Mitigates Demonstration Regurgitation in On-Device PII Substitution with Small Language Models
by: Sadani, Anuj, et al.
Published: (2026)
by: Sadani, Anuj, et al.
Published: (2026)
PII-Compass: Guiding LLM training data extraction prompts towards the target PII via grounding
by: Nakka, Krishna Kanth, et al.
Published: (2024)
by: Nakka, Krishna Kanth, et al.
Published: (2024)
Unmasking the Reality of PII Masking Models: Performance Gaps and the Call for Accountability
by: Singh, Devansh, et al.
Published: (2025)
by: Singh, Devansh, et al.
Published: (2025)
Efficient Continual Pre-training of LLMs for Low-resource Languages
by: Nag, Arijit, et al.
Published: (2024)
by: Nag, Arijit, et al.
Published: (2024)
How (un)ethical are instruction-centric responses of LLMs? Unveiling the vulnerabilities of safety guardrails to harmful queries
by: Banerjee, Somnath, et al.
Published: (2024)
by: Banerjee, Somnath, et al.
Published: (2024)
PIIGuard: Mitigating PII Harvesting under Adversarial Sanitization
by: Liu, Mingshuo, et al.
Published: (2026)
by: Liu, Mingshuo, et al.
Published: (2026)
All that is English may be Hindi: Enhancing language identification through automatic ranking of likeliness of word borrowing in social media
by: Patro, Jasabanta, et al.
Published: (2017)
by: Patro, Jasabanta, et al.
Published: (2017)
On Zero-Shot Counterspeech Generation by LLMs
by: Saha, Punyajoy, et al.
Published: (2024)
by: Saha, Punyajoy, et al.
Published: (2024)
InfFeed: Influence Functions as a Feedback to Improve the Performance of Subjective Tasks
by: Banerjee, Somnath, et al.
Published: (2024)
by: Banerjee, Somnath, et al.
Published: (2024)
DistALANER: Distantly Supervised Active Learning Augmented Named Entity Recognition in the Open Source Software Ecosystem
by: Banerjee, Somnath, et al.
Published: (2024)
by: Banerjee, Somnath, et al.
Published: (2024)
Proactive Privacy Amnesia for Large Language Models: Safeguarding PII with Negligible Impact on Model Utility
by: Kuo, Martin, et al.
Published: (2025)
by: Kuo, Martin, et al.
Published: (2025)
REVERSUM: A Multi-staged Retrieval-Augmented Generation Method to Enhance Wikipedia Tail Biographies through Personal Narratives
by: Adak, Sayantan, et al.
Published: (2025)
by: Adak, Sayantan, et al.
Published: (2025)
A Comparative Study of Light-weight Language Models for PII Masking and their Deployment for Real Conversational Texts
by: Acharya, Prabigya, et al.
Published: (2025)
by: Acharya, Prabigya, et al.
Published: (2025)
Privacy-Preserving Language Model Inference with Instance Obfuscation
by: Yao, Yixiang, et al.
Published: (2024)
by: Yao, Yixiang, et al.
Published: (2024)
Authorship Obfuscation in Multilingual Machine-Generated Text Detection
by: Macko, Dominik, et al.
Published: (2024)
by: Macko, Dominik, et al.
Published: (2024)
Mitigating Self-Preference by Authorship Obfuscation
by: Mahbub, Taslim, et al.
Published: (2025)
by: Mahbub, Taslim, et al.
Published: (2025)
ASR-FAIRBENCH: Measuring and Benchmarking Equity Across Speech Recognition Systems
by: Rai, Anand, et al.
Published: (2025)
by: Rai, Anand, et al.
Published: (2025)
On the effective transfer of knowledge from English to Hindi Wikipedia
by: Das, Paramita, et al.
Published: (2024)
by: Das, Paramita, et al.
Published: (2024)
PANORAMA: A synthetic PII-laced dataset for studying sensitive data memorization in LLMs
by: Selvam, Sriram, et al.
Published: (2025)
by: Selvam, Sriram, et al.
Published: (2025)
GLiNER2-PII: A Multilingual Model for Personally Identifiable Information Extraction
by: Zaratiana, Urchade, et al.
Published: (2026)
by: Zaratiana, Urchade, et al.
Published: (2026)
Similar Items
-
Lost without translation -- Can transformer (language models) understand mood states?
by: Shivaprakash, Prakrithi, et al.
Published: (2025) -
Benchmarking Motivational Interviewing Competence of Large Language Models
by: Jha, Aishwariya, et al.
Published: (2026) -
Language Models for Standardising Clinical Notes and Information Extraction in Addiction Psychiatry—An Empirical Study
by: Haritha Gireesh, et al.
Published: (2025) -
ASR Under the Stethoscope: Evaluating Biases in Clinical Speech Recognition across Indian Languages
by: Kumar, Subham, et al.
Published: (2025) -
GRACE: Graph Neural Networks for Locus-of-Care Prediction under Extreme Class Imbalance
by: Kumar, Subham, et al.
Published: (2025)