Current State in Privacy-Preserving Text Preprocessing for Domain-Agnostic NLP
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sinha, Abhirup, Saha, Pritilata, Saha, Tithi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Study of Privacy-preserving Language Modeling Approaches
von: Saha, Pritilata, et al.
Veröffentlicht: (2025)
von: Saha, Pritilata, et al.
Veröffentlicht: (2025)
BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources
von: Kumar, Raghvendra, et al.
Veröffentlicht: (2026)
von: Kumar, Raghvendra, et al.
Veröffentlicht: (2026)
Detecting Statements in Text: A Domain-Agnostic Few-Shot Solution
von: Chausson, Sandrine, et al.
Veröffentlicht: (2024)
von: Chausson, Sandrine, et al.
Veröffentlicht: (2024)
Understanding Cross-Domain Adaptation in Low-Resource Topic Modeling
von: Akash, Pritom Saha, et al.
Veröffentlicht: (2025)
von: Akash, Pritom Saha, et al.
Veröffentlicht: (2025)
Text Categorization Can Enhance Domain-Agnostic Stopword Extraction
von: Turki, Houcemeddine, et al.
Veröffentlicht: (2024)
von: Turki, Houcemeddine, et al.
Veröffentlicht: (2024)
SciNLP: A Domain-Specific Benchmark for Full-Text Scientific Entity and Relation Extraction in NLP
von: Duan, Decheng, et al.
Veröffentlicht: (2025)
von: Duan, Decheng, et al.
Veröffentlicht: (2025)
NLP Workbench: Efficient and Extensible Integration of State-of-the-art Text Mining Tools
von: Yao, Peiran, et al.
Veröffentlicht: (2023)
von: Yao, Peiran, et al.
Veröffentlicht: (2023)
Privacy Evaluation Benchmarks for NLP Models
von: Huang, Wei, et al.
Veröffentlicht: (2024)
von: Huang, Wei, et al.
Veröffentlicht: (2024)
TACIT: A Target-Agnostic Feature Disentanglement Framework for Cross-Domain Text Classification
von: Song, Rui, et al.
Veröffentlicht: (2023)
von: Song, Rui, et al.
Veröffentlicht: (2023)
Two eyes, Two views, and finally, One summary! Towards Multi-modal Multi-tasking Knowledge-Infused Medical Dialogue Summarization
von: Saha, Anisha, et al.
Veröffentlicht: (2024)
von: Saha, Anisha, et al.
Veröffentlicht: (2024)
MASE: Interpretable NLP Models via Model-Agnostic Saliency Estimation
von: Yang, Zhou, et al.
Veröffentlicht: (2025)
von: Yang, Zhou, et al.
Veröffentlicht: (2025)
Empirical Analysis of the Effect of Context in the Task of Automated Essay Scoring in Transformer-Based Models
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
von: Chakravarty, Abhirup
Veröffentlicht: (2025)
Comparison of Large Language Models for Deployment Requirements
von: Yaman, Alper, et al.
Veröffentlicht: (2025)
von: Yaman, Alper, et al.
Veröffentlicht: (2025)
Knowledge-Aware Self-Correction in Language Models via Structured Memory Graphs
von: Saha, Swayamjit
Veröffentlicht: (2025)
von: Saha, Swayamjit
Veröffentlicht: (2025)
Measuring the Robustness of NLP Models to Domain Shifts
von: Calderon, Nitay, et al.
Veröffentlicht: (2023)
von: Calderon, Nitay, et al.
Veröffentlicht: (2023)
Explainability of Text Processing and Retrieval Methods: A Survey
von: Saha, Sourav, et al.
Veröffentlicht: (2022)
von: Saha, Sourav, et al.
Veröffentlicht: (2022)
Investigating Large Language Models' Linguistic Abilities for Text Preprocessing
von: Braga, Marco, et al.
Veröffentlicht: (2025)
von: Braga, Marco, et al.
Veröffentlicht: (2025)
Evolutionary Feature-wise Thresholding for Binary Representation of NLP Embeddings
von: Sinha, Soumen, et al.
Veröffentlicht: (2025)
von: Sinha, Soumen, et al.
Veröffentlicht: (2025)
Enhancing Short-Text Topic Modeling with LLM-Driven Context Expansion and Prefix-Tuned VAEs
von: Akash, Pritom Saha, et al.
Veröffentlicht: (2024)
von: Akash, Pritom Saha, et al.
Veröffentlicht: (2024)
When does MAML Work the Best? An Empirical Study on Model-Agnostic Meta-Learning in NLP Applications
von: Liu, Zequn, et al.
Veröffentlicht: (2020)
von: Liu, Zequn, et al.
Veröffentlicht: (2020)
Transparent NLP: Using RAG and LLM Alignment for Privacy Q&A
von: Leschanowsky, Anna, et al.
Veröffentlicht: (2025)
von: Leschanowsky, Anna, et al.
Veröffentlicht: (2025)
NLP Privacy Risk Identification in Social Media (NLP-PRISM): A Survey
von: Goswami, Dhiman, et al.
Veröffentlicht: (2026)
von: Goswami, Dhiman, et al.
Veröffentlicht: (2026)
How Does A Text Preprocessing Pipeline Affect Ontology Matching?
von: Qiang, Zhangcheng, et al.
Veröffentlicht: (2024)
von: Qiang, Zhangcheng, et al.
Veröffentlicht: (2024)
WavRx: a Disease-Agnostic, Generalizable, and Privacy-Preserving Speech Health Diagnostic Model
von: Zhu, Yi, et al.
Veröffentlicht: (2024)
von: Zhu, Yi, et al.
Veröffentlicht: (2024)
Prompting Away Stereotypes? Evaluating Bias in Text-to-Image Models for Occupations
von: Raza, Shaina, et al.
Veröffentlicht: (2025)
von: Raza, Shaina, et al.
Veröffentlicht: (2025)
Adversarial Paraphrasing: A Universal Attack for Humanizing AI-Generated Text
von: Cheng, Yize, et al.
Veröffentlicht: (2025)
von: Cheng, Yize, et al.
Veröffentlicht: (2025)
Synthesizing Privacy-Preserving Text Data via Finetuning without Finetuning Billion-Scale LLMs
von: Tan, Bowen, et al.
Veröffentlicht: (2025)
von: Tan, Bowen, et al.
Veröffentlicht: (2025)
NAP^2: A Benchmark for Naturalness and Privacy-Preserving Text Rewriting by Learning from Human
von: Huang, Shuo, et al.
Veröffentlicht: (2024)
von: Huang, Shuo, et al.
Veröffentlicht: (2024)
Anonymous-by-Construction: An LLM-Driven Framework for Privacy-Preserving Text
von: Albanese, Federico, et al.
Veröffentlicht: (2026)
von: Albanese, Federico, et al.
Veröffentlicht: (2026)
Transforming Sensitive Documents into Quantitative Data: An AI-Based Preprocessing Toolchain for Structured and Privacy-Conscious Analysis
von: Ledberg, Anders, et al.
Veröffentlicht: (2025)
von: Ledberg, Anders, et al.
Veröffentlicht: (2025)
QuIM-RAG: Advancing Retrieval-Augmented Generation with Inverted Question Matching for Enhanced QA Performance
von: Saha, Binita, et al.
Veröffentlicht: (2025)
von: Saha, Binita, et al.
Veröffentlicht: (2025)
ParsiPy: NLP Toolkit for Historical Persian Texts in Python
von: Farsi, Farhan, et al.
Veröffentlicht: (2025)
von: Farsi, Farhan, et al.
Veröffentlicht: (2025)
State-of-the-art generalisation research in NLP: A taxonomy and review
von: Hupkes, Dieuwke, et al.
Veröffentlicht: (2022)
von: Hupkes, Dieuwke, et al.
Veröffentlicht: (2022)
Evaluation Metrics for Text Data Augmentation in NLP
von: Amadeus, Marcellus, et al.
Veröffentlicht: (2024)
von: Amadeus, Marcellus, et al.
Veröffentlicht: (2024)
1-Diffractor: Efficient and Utility-Preserving Text Obfuscation Leveraging Word-Level Metric Differential Privacy
von: Meisenbacher, Stephen, et al.
Veröffentlicht: (2024)
von: Meisenbacher, Stephen, et al.
Veröffentlicht: (2024)
An EcoSage Assistant: Towards Building A Multimodal Plant Care Dialogue Assistant
von: Tomar, Mohit, et al.
Veröffentlicht: (2024)
von: Tomar, Mohit, et al.
Veröffentlicht: (2024)
One Word is Enough: Minimal Adversarial Perturbations for Neural Text Ranking
von: Karmakar, Tanmay, et al.
Veröffentlicht: (2026)
von: Karmakar, Tanmay, et al.
Veröffentlicht: (2026)
ATEB: Evaluating and Improving Advanced NLP Tasks for Text Embedding Models
von: Han, Simeng, et al.
Veröffentlicht: (2025)
von: Han, Simeng, et al.
Veröffentlicht: (2025)
State of NLP in Kenya: A Survey
von: Amol, Cynthia Jayne, et al.
Veröffentlicht: (2024)
von: Amol, Cynthia Jayne, et al.
Veröffentlicht: (2024)
User Behavior Prediction as a Generic, Robust, Scalable, and Low-Cost Evaluation Strategy for Estimating Generalization in LLMs
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
von: Saha, Sougata, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Study of Privacy-preserving Language Modeling Approaches
von: Saha, Pritilata, et al.
Veröffentlicht: (2025) -
BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources
von: Kumar, Raghvendra, et al.
Veröffentlicht: (2026) -
Detecting Statements in Text: A Domain-Agnostic Few-Shot Solution
von: Chausson, Sandrine, et al.
Veröffentlicht: (2024) -
Understanding Cross-Domain Adaptation in Low-Resource Topic Modeling
von: Akash, Pritom Saha, et al.
Veröffentlicht: (2025) -
Text Categorization Can Enhance Domain-Agnostic Stopword Extraction
von: Turki, Houcemeddine, et al.
Veröffentlicht: (2024)