Towards Fair and Efficient De-identification: Quantifying the Efficiency and Generalizability of De-identification Approaches
Fuente:
arXiv
Guardado en:
| Autores principales: | Zambare, Noopur, Aghakasiri, Kiana, Lin, Carissa, Ye, Carrie, Mitchell, J. Ross, Abdalla, Mohamed |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Not What the Doctor Ordered: Surveying LLM-based De-identification and Quantifying Clinical Information Loss
por: Aghakasiri, Kiana, et al.
Publicado: (2025)
por: Aghakasiri, Kiana, et al.
Publicado: (2025)
AIOptimizer - Software performance optimisation prototype for cost minimisation
por: Zambare, Noopur
Publicado: (2023)
por: Zambare, Noopur
Publicado: (2023)
In the Name of Fairness: Assessing the Bias in Clinical Record De-identification
por: Xiao, Yuxin, et al.
Publicado: (2023)
por: Xiao, Yuxin, et al.
Publicado: (2023)
Thunder-DeID: Accurate and Efficient De-identification Framework for Korean Court Judgments
por: Hahm, Sungeun, et al.
Publicado: (2025)
por: Hahm, Sungeun, et al.
Publicado: (2025)
Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMs
por: Tavakoli, Mohammad, et al.
Publicado: (2025)
por: Tavakoli, Mohammad, et al.
Publicado: (2025)
DeIDClinic: A Risk-Aware Pseudonymization Framework for Clinical Text De-identification and Re-identification Risk Assessment
por: Paul, Angel, et al.
Publicado: (2024)
por: Paul, Angel, et al.
Publicado: (2024)
Enhancing Clinical Models with Pseudo Data for De-identification
por: Landes, Paul, et al.
Publicado: (2025)
por: Landes, Paul, et al.
Publicado: (2025)
Enhancing the De-identification of Personally Identifiable Information in Educational Data
por: Ji, Zilyu, et al.
Publicado: (2025)
por: Ji, Zilyu, et al.
Publicado: (2025)
Re-identification of De-identified Documents with Autoregressive Infilling
por: Charpentier, Lucas Georges Gabriel, et al.
Publicado: (2025)
por: Charpentier, Lucas Georges Gabriel, et al.
Publicado: (2025)
Comparing Multiclass Classification Algorithms for Financial Distress Prediction
por: Zambare, Noopur, et al.
Publicado: (2023)
por: Zambare, Noopur, et al.
Publicado: (2023)
Paradox of De-identification: A Critique of HIPAA Safe Harbour in the Age of LLMs
por: Jiang, Lavender Y., et al.
Publicado: (2026)
por: Jiang, Lavender Y., et al.
Publicado: (2026)
Collaboration or Corporate Capture? Quantifying NLP's Reliance on Industry Artifacts and Contributions
por: Aitken, Will, et al.
Publicado: (2023)
por: Aitken, Will, et al.
Publicado: (2023)
Anonpsy: A Graph-Based Framework for Structure-Preserving De-identification of Psychiatric Narratives
por: Lim, Kyung Ho, et al.
Publicado: (2026)
por: Lim, Kyung Ho, et al.
Publicado: (2026)
De-identification is not enough: a comparison between de-identified and synthetic clinical notes
por: Sarkar, Atiquer Rahman, et al.
Publicado: (2024)
por: Sarkar, Atiquer Rahman, et al.
Publicado: (2024)
Differentially Private De-identification of Dutch Clinical Notes: A Comparative Evaluation
por: Miranda, Michele, et al.
Publicado: (2026)
por: Miranda, Michele, et al.
Publicado: (2026)
Improving the Performance of Radiology Report De-identification with Large-Scale Training and Benchmarking Against Cloud Vendor Methods
por: Prakash, Eva, et al.
Publicado: (2025)
por: Prakash, Eva, et al.
Publicado: (2025)
Automated Identification of Discourse Markers Using the NLP Approach: The Case of "Okay"
por: Sanosi, Abdulaziz, et al.
Publicado: (2021)
por: Sanosi, Abdulaziz, et al.
Publicado: (2021)
A Trip Towards Fairness: Bias and De-Biasing in Large Language Models
por: Ranaldi, Leonardo, et al.
Publicado: (2023)
por: Ranaldi, Leonardo, et al.
Publicado: (2023)
Cross-lingual paraphrase identification
por: Fedorova, Inessa, et al.
Publicado: (2024)
por: Fedorova, Inessa, et al.
Publicado: (2024)
SHIELD: A Diverse Clinical Note Dataset and Distilled Small Language Models for Enterprise-Scale De-identification
por: Posada, Jose D., et al.
Publicado: (2026)
por: Posada, Jose D., et al.
Publicado: (2026)
GiusBERTo: A Legal Language Model for Personal Data De-identification in Italian Court of Auditors Decisions
por: Salierno, Giulio, et al.
Publicado: (2024)
por: Salierno, Giulio, et al.
Publicado: (2024)
CodeSCM: Causal Analysis for Multi-Modal Code Generation
por: Gupta, Mukur, et al.
Publicado: (2025)
por: Gupta, Mukur, et al.
Publicado: (2025)
LLMs-in-the-Loop Part 2: Expert Small AI Models for Anonymization and De-identification of PHI Across Multiple Languages
por: Gunay, Murat, et al.
Publicado: (2024)
por: Gunay, Murat, et al.
Publicado: (2024)
Synthetic Data for Veterinary EHR De-identification: Benefits, Limits, and Safety Trade-offs Under Fixed Compute
por: Brundage, David
Publicado: (2026)
por: Brundage, David
Publicado: (2026)
Efficiently Quantifying and Mitigating Ripple Effects in Model Editing
por: Wang, Jianchen, et al.
Publicado: (2024)
por: Wang, Jianchen, et al.
Publicado: (2024)
Stronger Re-identification Attacks through Reasoning and Aggregation
por: Charpentier, Lucas Georges Gabriel, et al.
Publicado: (2025)
por: Charpentier, Lucas Georges Gabriel, et al.
Publicado: (2025)
Hate Speech Detection with Generalizable Target-aware Fairness
por: Chen, Tong, et al.
Publicado: (2024)
por: Chen, Tong, et al.
Publicado: (2024)
De-identification of clinical free text using natural language processing: A systematic review of current approaches
por: Kovačević, Aleksandar, et al.
Publicado: (2023)
por: Kovačević, Aleksandar, et al.
Publicado: (2023)
Feature engineering vs. deep learning for paper section identification: Toward applications in Chinese medical literature
por: Zhou, Sijia, et al.
Publicado: (2024)
por: Zhou, Sijia, et al.
Publicado: (2024)
CTC-DID: CTC-Based Arabic dialect identification for streaming applications
por: Farooq, Muhammad Umar, et al.
Publicado: (2026)
por: Farooq, Muhammad Umar, et al.
Publicado: (2026)
Comparing representations of long clinical texts for the task of patient note-identification
por: Alsaidi, Safa, et al.
Publicado: (2025)
por: Alsaidi, Safa, et al.
Publicado: (2025)
Automatic register identification for the open web using multilingual deep learning
por: Henriksson, Erik, et al.
Publicado: (2024)
por: Henriksson, Erik, et al.
Publicado: (2024)
AROhI: An Interactive Tool for Estimating ROI of Data Analytics
por: Zambare, Noopur, et al.
Publicado: (2024)
por: Zambare, Noopur, et al.
Publicado: (2024)
AIDBench: A benchmark for evaluating the authorship identification capability of large language models
por: Wen, Zichen, et al.
Publicado: (2024)
por: Wen, Zichen, et al.
Publicado: (2024)
Classifier identification in Ancient Egyptian as a low-resource sequence-labelling task
por: Nikolaev, Dmitry, et al.
Publicado: (2024)
por: Nikolaev, Dmitry, et al.
Publicado: (2024)
Unveiling the Tapestry of Automated Essay Scoring: A Comprehensive Investigation of Accuracy, Fairness, and Generalizability
por: Yang, Kaixun, et al.
Publicado: (2024)
por: Yang, Kaixun, et al.
Publicado: (2024)
Brain-tuning Improves Generalizability and Efficiency of Brain Alignment in Speech Models
por: Moussa, Omer, et al.
Publicado: (2025)
por: Moussa, Omer, et al.
Publicado: (2025)
AIDetx: a compression-based method for identification of machine-learning generated text
por: Almeida, Leonardo, et al.
Publicado: (2024)
por: Almeida, Leonardo, et al.
Publicado: (2024)
AI-based approach to burnout identification from textual data
por: Zavertiaeva, Marina, et al.
Publicado: (2026)
por: Zavertiaeva, Marina, et al.
Publicado: (2026)
Towards Generalizable Implicit In-Context Learning with Attention Routing
por: Li, Jiaqian, et al.
Publicado: (2025)
por: Li, Jiaqian, et al.
Publicado: (2025)
Ejemplares similares
-
Not What the Doctor Ordered: Surveying LLM-based De-identification and Quantifying Clinical Information Loss
por: Aghakasiri, Kiana, et al.
Publicado: (2025) -
AIOptimizer - Software performance optimisation prototype for cost minimisation
por: Zambare, Noopur
Publicado: (2023) -
In the Name of Fairness: Assessing the Bias in Clinical Record De-identification
por: Xiao, Yuxin, et al.
Publicado: (2023) -
Thunder-DeID: Accurate and Efficient De-identification Framework for Korean Court Judgments
por: Hahm, Sungeun, et al.
Publicado: (2025) -
Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMs
por: Tavakoli, Mohammad, et al.
Publicado: (2025)