Comparing representations of long clinical texts for the task of patient note-identification
Fuente:
arXiv
Saved in:
| Main Authors: | Alsaidi, Safa, Vincent, Marc, Boyer, Olivia, Garcelon, Nicolas, Couceiro, Miguel, Coulet, Adrien |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Shapley Regression for Rare Disease Diagnosis Support: a case study on APDS
by: Alsaidi, Safa, et al.
Published: (2026)
by: Alsaidi, Safa, et al.
Published: (2026)
Facilitating phenotyping from clinical texts: the medkit library
by: Neuraz, Antoine, et al.
Published: (2024)
by: Neuraz, Antoine, et al.
Published: (2024)
Efficient extraction of medication information from clinical notes: an evaluation in two languages
by: Fabacher, Thibaut, et al.
Published: (2025)
by: Fabacher, Thibaut, et al.
Published: (2025)
Prompting Large Language Models for Supporting the Differential Diagnosis of Anemia
by: Castagnari, Elisa, et al.
Published: (2024)
by: Castagnari, Elisa, et al.
Published: (2024)
Analysing Lightweight Large Language Models for Biomedical Named Entity Recognition on Diverse Ouput Formats
by: Epron, Pierre, et al.
Published: (2026)
by: Epron, Pierre, et al.
Published: (2026)
De-identification is not enough: a comparison between de-identified and synthetic clinical notes
by: Sarkar, Atiquer Rahman, et al.
Published: (2024)
by: Sarkar, Atiquer Rahman, et al.
Published: (2024)
Domain-specific long text classification from sparse relevant information
by: D'Cruz, Célia, et al.
Published: (2024)
by: D'Cruz, Célia, et al.
Published: (2024)
Solving morphological analogies: from retrieval to generation
by: Marquer, Esteban, et al.
Published: (2023)
by: Marquer, Esteban, et al.
Published: (2023)
Serialized EHR make for good text representations
by: Chou, Zhirong, et al.
Published: (2025)
by: Chou, Zhirong, et al.
Published: (2025)
EntmaxKV: Support-Aware Decoding for Entmax Attention
by: Duarte, Gonçalo, et al.
Published: (2026)
by: Duarte, Gonçalo, et al.
Published: (2026)
Frankentext: Stitching random text fragments into long-form narratives
by: Pham, Chau Minh, et al.
Published: (2025)
by: Pham, Chau Minh, et al.
Published: (2025)
Can human clinical rationales improve the performance and explainability of clinical text classification models?
by: Metzner, Christoph, et al.
Published: (2025)
by: Metzner, Christoph, et al.
Published: (2025)
Classifier identification in Ancient Egyptian as a low-resource sequence-labelling task
by: Nikolaev, Dmitry, et al.
Published: (2024)
by: Nikolaev, Dmitry, et al.
Published: (2024)
Tgea: An error-annotated dataset and benchmark tasks for text generation from pretrained language models
by: He, Jie, et al.
Published: (2025)
by: He, Jie, et al.
Published: (2025)
The Balancing Act: Unmasking and Alleviating ASR Biases in Portuguese
by: Kulkarni, Ajinkya, et al.
Published: (2024)
by: Kulkarni, Ajinkya, et al.
Published: (2024)
VERISCORE: Evaluating the factuality of verifiable claims in long-form text generation
by: Song, Yixiao, et al.
Published: (2024)
by: Song, Yixiao, et al.
Published: (2024)
Just-in-time and distributed task representations in language models
by: Li, Yuxuan, et al.
Published: (2025)
by: Li, Yuxuan, et al.
Published: (2025)
Comparing energy consumption and accuracy in text classification inference
by: Zschache, Johannes, et al.
Published: (2025)
by: Zschache, Johannes, et al.
Published: (2025)
A quantitative analysis of semantic information in deep representations of text and images
by: Acevedo, Santiago, et al.
Published: (2025)
by: Acevedo, Santiago, et al.
Published: (2025)
AIDetx: a compression-based method for identification of machine-learning generated text
by: Almeida, Leonardo, et al.
Published: (2024)
by: Almeida, Leonardo, et al.
Published: (2024)
FrameNet Semantic Role Classification by Analogy
by: Ngo, Van-Duy, et al.
Published: (2026)
by: Ngo, Van-Duy, et al.
Published: (2026)
ARC-Encoder: learning compressed text representations for large language models
by: Pilchen, Hippolyte, et al.
Published: (2025)
by: Pilchen, Hippolyte, et al.
Published: (2025)
How reparametrization trick broke differentially-private text representation learning
by: Habernal, Ivan
Published: (2022)
by: Habernal, Ivan
Published: (2022)
EzSQL: An SQL intermediate representation for improving SQL-to-text Generation
by: Bhardwaj, Meher, et al.
Published: (2024)
by: Bhardwaj, Meher, et al.
Published: (2024)
ProText: A benchmark dataset for measuring (mis)gendering in long-form texts
by: Kotek, Hadas, et al.
Published: (2026)
by: Kotek, Hadas, et al.
Published: (2026)
Does quantization affect models' performance on long-context tasks?
by: Mekala, Anmol, et al.
Published: (2025)
by: Mekala, Anmol, et al.
Published: (2025)
De-identification of clinical free text using natural language processing: A systematic review of current approaches
by: Kovačević, Aleksandar, et al.
Published: (2023)
by: Kovačević, Aleksandar, et al.
Published: (2023)
Unveiling Biases while Embracing Sustainability: Assessing the Dual Challenges of Automatic Speech Recognition Systems
by: Kulkarni, Ajinkya, et al.
Published: (2025)
by: Kulkarni, Ajinkya, et al.
Published: (2025)
Impact of enriched meaning representations for language generation in dialogue tasks: A comprehensive exploration of the relevance of tasks, corpora and metrics
by: Vázquez, Alain, et al.
Published: (2026)
by: Vázquez, Alain, et al.
Published: (2026)
Comparing Labeled Markov Chains: A Cantor-Kantorovich Approach
by: Banse, Adrien, et al.
Published: (2025)
by: Banse, Adrien, et al.
Published: (2025)
Using reasoning LLMs to extract SDOH events from clinical notes
by: Dogan, Ertan, et al.
Published: (2026)
by: Dogan, Ertan, et al.
Published: (2026)
Evaluating the fairness of task-adaptive pretraining on unlabeled test data before few-shot text classification
by: Dubey, Kush
Published: (2024)
by: Dubey, Kush
Published: (2024)
How much do contextualized representations encode long-range context?
by: Sun, Simeng, et al.
Published: (2024)
by: Sun, Simeng, et al.
Published: (2024)
Evaluation of the phi-3-mini SLM for identification of texts related to medicine, health, and sports injuries
by: Brogly, Chris, et al.
Published: (2025)
by: Brogly, Chris, et al.
Published: (2025)
Nationality encoding in language model hidden states: Probing culturally differentiated representations in persona-conditioned academic text
by: Jackson, Paul, et al.
Published: (2026)
by: Jackson, Paul, et al.
Published: (2026)
SNOBERT: A Benchmark for clinical notes entity linking in the SNOMED CT clinical terminology
by: Kulyabin, Mikhail, et al.
Published: (2024)
by: Kulyabin, Mikhail, et al.
Published: (2024)
Are LLMs reliable? An exploration of the reliability of large language models in clinical note generation
by: Carandang, Kristine Ann M., et al.
Published: (2025)
by: Carandang, Kristine Ann M., et al.
Published: (2025)
Comparing LLM-generated and human-authored news text using formal syntactic theory
by: Zamaraeva, Olga, et al.
Published: (2025)
by: Zamaraeva, Olga, et al.
Published: (2025)
People who frequently use ChatGPT for writing tasks are accurate and robust detectors of AI-generated text
by: Russell, Jenna, et al.
Published: (2025)
by: Russell, Jenna, et al.
Published: (2025)
Synthetically generated text for supervised text analysis
by: Halterman, Andrew
Published: (2023)
by: Halterman, Andrew
Published: (2023)
Similar Items
-
Shapley Regression for Rare Disease Diagnosis Support: a case study on APDS
by: Alsaidi, Safa, et al.
Published: (2026) -
Facilitating phenotyping from clinical texts: the medkit library
by: Neuraz, Antoine, et al.
Published: (2024) -
Efficient extraction of medication information from clinical notes: an evaluation in two languages
by: Fabacher, Thibaut, et al.
Published: (2025) -
Prompting Large Language Models for Supporting the Differential Diagnosis of Anemia
by: Castagnari, Elisa, et al.
Published: (2024) -
Analysing Lightweight Large Language Models for Biomedical Named Entity Recognition on Diverse Ouput Formats
by: Epron, Pierre, et al.
Published: (2026)