Re-identification of De-identified Documents with Autoregressive Infilling
Fuente:
arXiv
Saved in:
| Main Authors: | Charpentier, Lucas Georges Gabriel, Lison, Pierre |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stronger Re-identification Attacks through Reasoning and Aggregation
by: Charpentier, Lucas Georges Gabriel, et al.
Published: (2025)
by: Charpentier, Lucas Georges Gabriel, et al.
Published: (2025)
Protecting De-identified Documents from Search-based Linkage Attacks
by: Lison, Pierre, et al.
Published: (2025)
by: Lison, Pierre, et al.
Published: (2025)
GPT or BERT: why not both?
by: Charpentier, Lucas Georges Gabriel, et al.
Published: (2024)
by: Charpentier, Lucas Georges Gabriel, et al.
Published: (2024)
Dual-objective Language Models: Training Efficiency Without Overfitting
by: Samuel, David, et al.
Published: (2025)
by: Samuel, David, et al.
Published: (2025)
Systematic Generalization in Language Models Scales with Information Entropy
by: Wold, Sondre, et al.
Published: (2025)
by: Wold, Sondre, et al.
Published: (2025)
More Room for Language: Investigating the Effect of Retrieval on Language Models
by: Samuel, David, et al.
Published: (2024)
by: Samuel, David, et al.
Published: (2024)
Self-Infilling Code Generation
by: Zheng, Lin, et al.
Published: (2023)
by: Zheng, Lin, et al.
Published: (2023)
Compositional Generalization with Grounded Language Models
by: Wold, Sondre, et al.
Published: (2024)
by: Wold, Sondre, et al.
Published: (2024)
Enhancing Naturalness in LLM-Generated Utterances through Disfluency Insertion
by: Hassan, Syed Zohaib, et al.
Published: (2024)
by: Hassan, Syed Zohaib, et al.
Published: (2024)
DeIDClinic: A Risk-Aware Pseudonymization Framework for Clinical Text De-identification and Re-identification Risk Assessment
by: Paul, Angel, et al.
Published: (2024)
by: Paul, Angel, et al.
Published: (2024)
Prior Lessons of Incremental Dialogue and Robot Action Management for the Age of Language Models
by: Kennington, Casey, et al.
Published: (2025)
by: Kennington, Casey, et al.
Published: (2025)
Following Route Instructions using Large Vision-Language Models: A Comparison between Low-level and Panoramic Action Spaces
by: Kåsene, Vebjørn Haug, et al.
Published: (2025)
by: Kåsene, Vebjørn Haug, et al.
Published: (2025)
Small Languages, Big Models: A Study of Continual Training on Languages of Norway
by: Samuel, David, et al.
Published: (2024)
by: Samuel, David, et al.
Published: (2024)
De-identification is not enough: a comparison between de-identified and synthetic clinical notes
by: Sarkar, Atiquer Rahman, et al.
Published: (2024)
by: Sarkar, Atiquer Rahman, et al.
Published: (2024)
Conversational Feedback in Scripted versus Spontaneous Dialogues: A Comparative Analysis
by: Pilán, Ildikó, et al.
Published: (2023)
by: Pilán, Ildikó, et al.
Published: (2023)
Evaluating LLMs on Generating Age-Appropriate Child-Like Conversations
by: Hassan, Syed Zohaib, et al.
Published: (2025)
by: Hassan, Syed Zohaib, et al.
Published: (2025)
Truthful Text Sanitization Guided by Inference Attacks
by: Pilán, Ildikó, et al.
Published: (2024)
by: Pilán, Ildikó, et al.
Published: (2024)
Unlocking Prompt Infilling Capability for Diffusion Language Models
by: Fujinuma, Yoshinari, et al.
Published: (2026)
by: Fujinuma, Yoshinari, et al.
Published: (2026)
A Systematic Approach to Predict the Impact of Cybersecurity Vulnerabilities Using LLMs
by: Høst, Anders Mølmen, et al.
Published: (2025)
by: Høst, Anders Mølmen, et al.
Published: (2025)
Towards Fair and Efficient De-identification: Quantifying the Efficiency and Generalizability of De-identification Approaches
by: Zambare, Noopur, et al.
Published: (2026)
by: Zambare, Noopur, et al.
Published: (2026)
EFIM: Efficient Serving of LLMs for Infilling Tasks with Improved KV Cache Reuse
by: Guo, Tianyu, et al.
Published: (2025)
by: Guo, Tianyu, et al.
Published: (2025)
DreamOn: Diffusion Language Models For Code Infilling Beyond Fixed-size Canvas
by: Wu, Zirui, et al.
Published: (2026)
by: Wu, Zirui, et al.
Published: (2026)
Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings
by: Rose, Daniel, et al.
Published: (2023)
by: Rose, Daniel, et al.
Published: (2023)
MAGNET: Augmenting Generative Decoders with Representation Learning and Infilling Capabilities
by: Khosla, Savya, et al.
Published: (2025)
by: Khosla, Savya, et al.
Published: (2025)
Unlocking the Potential of Diffusion Language Models through Template Infilling
by: Lee, Junhoo, et al.
Published: (2025)
by: Lee, Junhoo, et al.
Published: (2025)
Empowering Character-level Text Infilling by Eliminating Sub-Tokens
by: Ren, Houxing, et al.
Published: (2024)
by: Ren, Houxing, et al.
Published: (2024)
Diffusion LMs Can Approximate Optimal Infilling Lengths Implicitly
by: Liu, Hengchang, et al.
Published: (2026)
by: Liu, Hengchang, et al.
Published: (2026)
[De|Re]constructing VLMs' Reasoning in Counting
by: Alghisi, Simone, et al.
Published: (2025)
by: Alghisi, Simone, et al.
Published: (2025)
Flexible-length Text Infilling for Discrete Diffusion Models
by: Zhang, Andrew, et al.
Published: (2025)
by: Zhang, Andrew, et al.
Published: (2025)
Planning-Aware Code Infilling via Horizon-Length Prediction
by: Ding, Yifeng, et al.
Published: (2024)
by: Ding, Yifeng, et al.
Published: (2024)
Advancing Semi-Supervised Learning for Automatic Post-Editing: Data-Synthesis by Mask-Infilling with Erroneous Terms
by: Lee, Wonkee, et al.
Published: (2022)
by: Lee, Wonkee, et al.
Published: (2022)
Improving the Performance of Radiology Report De-identification with Large-Scale Training and Benchmarking Against Cloud Vendor Methods
by: Prakash, Eva, et al.
Published: (2025)
by: Prakash, Eva, et al.
Published: (2025)
Enhancing Clinical Models with Pseudo Data for De-identification
by: Landes, Paul, et al.
Published: (2025)
by: Landes, Paul, et al.
Published: (2025)
Thunder-DeID: Accurate and Efficient De-identification Framework for Korean Court Judgments
by: Hahm, Sungeun, et al.
Published: (2025)
by: Hahm, Sungeun, et al.
Published: (2025)
Detecting Referring Expressions in Visually Grounded Dialogue with Autoregressive Language Models
by: Willemsen, Bram, et al.
Published: (2025)
by: Willemsen, Bram, et al.
Published: (2025)
From Completion to Editing: Unlocking Context-Aware Code Infilling via Search-and-Replace Instruction Tuning
by: Zhang, Jiajun, et al.
Published: (2026)
by: Zhang, Jiajun, et al.
Published: (2026)
Extracting Training Data from Diffusion Language Models via Infilling
by: Wang, Yihan, et al.
Published: (2026)
by: Wang, Yihan, et al.
Published: (2026)
ReFusion: A Diffusion Large Language Model with Parallel Autoregressive Decoding
by: Li, Jia-Nan, et al.
Published: (2025)
by: Li, Jia-Nan, et al.
Published: (2025)
Paradox of De-identification: A Critique of HIPAA Safe Harbour in the Age of LLMs
by: Jiang, Lavender Y., et al.
Published: (2026)
by: Jiang, Lavender Y., et al.
Published: (2026)
SaulLM-54B & SaulLM-141B: Scaling Up Domain Adaptation for the Legal Domain
by: Colombo, Pierre, et al.
Published: (2024)
by: Colombo, Pierre, et al.
Published: (2024)
Similar Items
-
Stronger Re-identification Attacks through Reasoning and Aggregation
by: Charpentier, Lucas Georges Gabriel, et al.
Published: (2025) -
Protecting De-identified Documents from Search-based Linkage Attacks
by: Lison, Pierre, et al.
Published: (2025) -
GPT or BERT: why not both?
by: Charpentier, Lucas Georges Gabriel, et al.
Published: (2024) -
Dual-objective Language Models: Training Efficiency Without Overfitting
by: Samuel, David, et al.
Published: (2025) -
Systematic Generalization in Language Models Scales with Information Entropy
by: Wold, Sondre, et al.
Published: (2025)