Beyond De-Identification: A Structured Approach for Defining and Detecting Indirect Identifiers in Medical Texts
Fuente:
arXiv
Saved in:
| Main Authors: | Baroud, Ibrahim, Raithel, Lisa, Möller, Sebastian, Roller, Roland |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MultiGraSCCo: A Multilingual Anonymization Benchmark with Annotations of Personal Identifiers
by: Baroud, Ibrahim, et al.
Published: (2026)
by: Baroud, Ibrahim, et al.
Published: (2026)
Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes
by: Frei, Johann, et al.
Published: (2025)
by: Frei, Johann, et al.
Published: (2025)
Evaluating the Robustness of Adverse Drug Event Classification Models Using Templates
by: MacPhail, Dorothea, et al.
Published: (2024)
by: MacPhail, Dorothea, et al.
Published: (2024)
Integrating Text and Time-Series into (Large) Language Models to Predict Medical Outcomes
by: Larbi, Iyadh Ben Cheikh, et al.
Published: (2025)
by: Larbi, Iyadh Ben Cheikh, et al.
Published: (2025)
DFKI-NLP at SemEval-2024 Task 2: Towards Robust LLMs Using Data Perturbations and MinMax Training
by: Verma, Bhuvanesh, et al.
Published: (2024)
by: Verma, Bhuvanesh, et al.
Published: (2024)
A Dataset for Pharmacovigilance in German, French, and Japanese: Annotating Adverse Drug Reactions across Languages
by: Raithel, Lisa, et al.
Published: (2024)
by: Raithel, Lisa, et al.
Published: (2024)
xMEN: A Modular Toolkit for Cross-Lingual Medical Entity Normalization
by: Borchert, Florian, et al.
Published: (2023)
by: Borchert, Florian, et al.
Published: (2023)
DeID-GPT: Zero-shot Medical Text De-Identification by GPT-4
by: Liu, Zhengliang, et al.
Published: (2023)
by: Liu, Zhengliang, et al.
Published: (2023)
Beyond Turing: A Comparative Analysis of Approaches for Detecting Machine-Generated Text
by: Adilazuarda, Muhammad Farid
Published: (2023)
by: Adilazuarda, Muhammad Farid
Published: (2023)
Interpretable Text Embeddings and Text Similarity Explanation: A Survey
by: Opitz, Juri, et al.
Published: (2025)
by: Opitz, Juri, et al.
Published: (2025)
Who Wrote This? Identifying Machine vs Human-Generated Text in Hausa
by: Sani, Babangida, et al.
Published: (2025)
by: Sani, Babangida, et al.
Published: (2025)
Beyond Easy Wins: A Text Hardness-Aware Benchmark for LLM-generated Text Detection
by: Ayoobi, Navid, et al.
Published: (2025)
by: Ayoobi, Navid, et al.
Published: (2025)
Identifying Speakers in Dialogue Transcripts: A Text-based Approach Using Pretrained Language Models
by: Nguyen, Minh, et al.
Published: (2024)
by: Nguyen, Minh, et al.
Published: (2024)
Factuality Beyond Coherence: Evaluating LLM Watermarking Methods for Medical Texts
by: Hastuti, Rochana Prih, et al.
Published: (2025)
by: Hastuti, Rochana Prih, et al.
Published: (2025)
Detecting Fallacies in Climate Misinformation: A Technocognitive Approach to Identifying Misleading Argumentation
by: Zanartu, Francisco, et al.
Published: (2024)
by: Zanartu, Francisco, et al.
Published: (2024)
OptBA: Optimizing Hyperparameters with the Bees Algorithm for Improved Medical Text Classification
by: Shaaban, Mai A., et al.
Published: (2023)
by: Shaaban, Mai A., et al.
Published: (2023)
Beyond LLMs: A Linguistic Approach to Causal Graph Generation from Narrative Texts
by: Li, Zehan, et al.
Published: (2025)
by: Li, Zehan, et al.
Published: (2025)
STRICTA: Structured Reasoning in Critical Text Assessment for Peer Review and Beyond
by: Dycke, Nils, et al.
Published: (2024)
by: Dycke, Nils, et al.
Published: (2024)
From Indirect Object Identification to Syllogisms: Exploring Binary Mechanisms in Transformer Circuits
by: Saraipour, Karim, et al.
Published: (2025)
by: Saraipour, Karim, et al.
Published: (2025)
Identifying Bias in Machine-generated Text Detection
by: Stowe, Kevin, et al.
Published: (2025)
by: Stowe, Kevin, et al.
Published: (2025)
Text2MDT: Extracting Medical Decision Trees from Medical Texts
by: Zhu, Wei, et al.
Published: (2024)
by: Zhu, Wei, et al.
Published: (2024)
Beyond Raw Detection Scores: Markov-Informed Calibration for Boosting Machine-Generated Text Detection
by: Wu, Chenwang, et al.
Published: (2026)
by: Wu, Chenwang, et al.
Published: (2026)
A Computational Framework to Identify Self-Aspects in Text
by: Caporusso, Jaya, et al.
Published: (2025)
by: Caporusso, Jaya, et al.
Published: (2025)
Multimodal Event Detection: Current Approaches and Defining the New Playground through LLMs and VLMs
by: Dey, Abhishek, et al.
Published: (2025)
by: Dey, Abhishek, et al.
Published: (2025)
Exons-Detect: Identifying and Amplifying Exonic Tokens via Hidden-State Discrepancy for Robust AI-Generated Text Detection
by: Zhu, Xiaowei, et al.
Published: (2026)
by: Zhu, Xiaowei, et al.
Published: (2026)
Beyond Perplexity: Character Distribution Signatures and the MDTA Benchmark for AI Text Detection
by: Narayanasamy, Priyadarshan, et al.
Published: (2026)
by: Narayanasamy, Priyadarshan, et al.
Published: (2026)
Identifying while Learning for Document Event Causality Identification
by: Liu, Cheng, et al.
Published: (2024)
by: Liu, Cheng, et al.
Published: (2024)
Emergence of Minimal Circuits for Indirect Object Identification in Attention-Only Transformers
by: Adhikari, Rabin
Published: (2025)
by: Adhikari, Rabin
Published: (2025)
Agnostic Language Identification and Generation
by: Høgsgaard, Mikael Møller, et al.
Published: (2026)
by: Høgsgaard, Mikael Møller, et al.
Published: (2026)
Uncovering Visual-Semantic Psycholinguistic Properties from the Distributional Structure of Text Embedding Space
by: Wu, Si, et al.
Published: (2025)
by: Wu, Si, et al.
Published: (2025)
HausaNLP at SemEval-2025 Task 11: Hausa Text Emotion Detection
by: Sani, Sani Abdullahi, et al.
Published: (2025)
by: Sani, Sani Abdullahi, et al.
Published: (2025)
Hybrid Annotation for Propaganda Detection: Integrating LLM Pre-Annotations with Human Intelligence
by: Sahitaj, Ariana, et al.
Published: (2025)
by: Sahitaj, Ariana, et al.
Published: (2025)
Defining, Understanding, and Detecting Online Toxicity: Challenges and Machine Learning Approaches
by: Shahi, Gautam Kishore, et al.
Published: (2025)
by: Shahi, Gautam Kishore, et al.
Published: (2025)
Simulating User Diversity in Task-Oriented Dialogue Systems using Large Language Models
by: Ahmad, Adnan, et al.
Published: (2025)
by: Ahmad, Adnan, et al.
Published: (2025)
Enhancing Text Authenticity: A Novel Hybrid Approach for AI-Generated Text Detection
by: Zhang, Ye, et al.
Published: (2024)
by: Zhang, Ye, et al.
Published: (2024)
Identifying Reliable Evaluation Metrics for Scientific Text Revision
by: Jourdan, Léane, et al.
Published: (2025)
by: Jourdan, Léane, et al.
Published: (2025)
Leveraging Implicit Feedback from Deployment Data in Dialogue
by: Pang, Richard Yuanzhe, et al.
Published: (2023)
by: Pang, Richard Yuanzhe, et al.
Published: (2023)
Efficiently Identifying Watermarked Segments in Mixed-Source Texts
by: Zhao, Xuandong, et al.
Published: (2024)
by: Zhao, Xuandong, et al.
Published: (2024)
A Multi-Strategy Approach for AI-Generated Text Detection
by: Zain, Ali, et al.
Published: (2025)
by: Zain, Ali, et al.
Published: (2025)
Detecting Toxic Language: Ontology and BERT-based Approaches for Bulgarian Text
by: Berbatova, Melania, et al.
Published: (2026)
by: Berbatova, Melania, et al.
Published: (2026)
Similar Items
-
MultiGraSCCo: A Multilingual Anonymization Benchmark with Annotations of Personal Identifiers
by: Baroud, Ibrahim, et al.
Published: (2026) -
Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes
by: Frei, Johann, et al.
Published: (2025) -
Evaluating the Robustness of Adverse Drug Event Classification Models Using Templates
by: MacPhail, Dorothea, et al.
Published: (2024) -
Integrating Text and Time-Series into (Large) Language Models to Predict Medical Outcomes
by: Larbi, Iyadh Ben Cheikh, et al.
Published: (2025) -
DFKI-NLP at SemEval-2024 Task 2: Towards Robust LLMs Using Data Perturbations and MinMax Training
by: Verma, Bhuvanesh, et al.
Published: (2024)