Evaluating LLM Abilities to Understand Tabular Electronic Health Records: A Comprehensive Study of Patient Data Extraction and Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Lovon, Jesus, Mouysset, Martin, Oleiwan, Jo, Moreno, Jose G., Damase-Michel, Christine, Tamine, Lynda |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ReToP: Learning to Rewrite Electronic Health Records for Clinical Prediction
by: Lovon-Melgarejo, Jesus, et al.
Published: (2026)
by: Lovon-Melgarejo, Jesus, et al.
Published: (2026)
PatientDx: Merging Large Language Models for Protecting Data-Privacy in Healthcare
by: Moreno, Jose G., et al.
Published: (2025)
by: Moreno, Jose G., et al.
Published: (2025)
Revisiting the MIMIC-IV Benchmark: Experiments Using Language Models for Electronic Health Records
by: Lovon, Jesus, et al.
Published: (2025)
by: Lovon, Jesus, et al.
Published: (2025)
Jointly Generating and Attributing Answers using Logits of Document-Identifier Tokens
by: Albarede, Lucas, et al.
Published: (2025)
by: Albarede, Lucas, et al.
Published: (2025)
An Evaluation Framework for Attributed Information Retrieval using Large Language Models
by: Djeddal, Hanane, et al.
Published: (2024)
by: Djeddal, Hanane, et al.
Published: (2024)
NEXT-EVAL: Next Evaluation of Traditional and LLM Web Data Record Extraction
by: Kim, Soyeon, et al.
Published: (2025)
by: Kim, Soyeon, et al.
Published: (2025)
Towards System Modelling to Support Diseases Data Extraction from the Electronic Health Records for Physicians Research Activities
by: Alsaqer, Bushra F., et al.
Published: (2024)
by: Alsaqer, Bushra F., et al.
Published: (2024)
DR.EHR: Dense Retrieval for Electronic Health Record with Knowledge Injection and Synthetic Data
by: Zhao, Zhengyun, et al.
Published: (2025)
by: Zhao, Zhengyun, et al.
Published: (2025)
HyST: LLM-Powered Hybrid Retrieval over Semi-Structured Tabular Data
by: Myung, Jiyoon, et al.
Published: (2025)
by: Myung, Jiyoon, et al.
Published: (2025)
Experience Retrieval-Augmentation with Electronic Health Records Enables Accurate Discharge QA
by: Ou, Justice, et al.
Published: (2025)
by: Ou, Justice, et al.
Published: (2025)
SampleLLM: Optimizing Tabular Data Synthesis in Recommendations
by: Gao, Jingtong, et al.
Published: (2025)
by: Gao, Jingtong, et al.
Published: (2025)
CliniQ: A Multi-faceted Benchmark for Electronic Health Record Retrieval with Semantic Match Assessment
by: Zhao, Zhengyun, et al.
Published: (2025)
by: Zhao, Zhengyun, et al.
Published: (2025)
Lessons Learned on Information Retrieval in Electronic Health Records: A Comparison of Embedding Models and Pooling Strategies
by: Myers, Skatje, et al.
Published: (2024)
by: Myers, Skatje, et al.
Published: (2024)
EHR-RAGp: Retrieval-Augmented Prototype-Guided Foundation Model for Electronic Health Records
by: Shurrab, Saeed, et al.
Published: (2026)
by: Shurrab, Saeed, et al.
Published: (2026)
Structure-Aware Chunking for Tabular Data in Retrieval-Augmented Generation
by: Guttal, Pooja, et al.
Published: (2026)
by: Guttal, Pooja, et al.
Published: (2026)
Synergizing Data Imputation and Electronic Health Records for Advancing Prostate Cancer Research: Challenges, and Practical Applications
by: Batouche, Abderrahim Oussama, et al.
Published: (2023)
by: Batouche, Abderrahim Oussama, et al.
Published: (2023)
Generating Querying Code from Text for Multi-Modal Electronic Health Record
by: ZHang, Mengliang
Published: (2025)
by: ZHang, Mengliang
Published: (2025)
RAM-EHR: Retrieval Augmentation Meets Clinical Predictions on Electronic Health Records
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
Pneuma: Leveraging LLMs for Tabular Data Representation and Retrieval in an End-to-End System
by: Balaka, Muhammad Imam Luthfi, et al.
Published: (2025)
by: Balaka, Muhammad Imam Luthfi, et al.
Published: (2025)
Retrieval Augmented Generation Evaluation for Health Documents
by: Ceresa, Mario, et al.
Published: (2025)
by: Ceresa, Mario, et al.
Published: (2025)
Retrieval Effectiveness of Enhanced Bibliographic Records.
by: Dillon, Martin, et al.
Published: (1990)
by: Dillon, Martin, et al.
Published: (1990)
Evaluating the Effectiveness and Scalability of LLM-Based Data Augmentation for Retrieval
by: Chitale, Pranjal A., et al.
Published: (2025)
by: Chitale, Pranjal A., et al.
Published: (2025)
CURE: A Dataset for Clinical Understanding & Retrieval Evaluation
by: Sheikh, Nadia Athar, et al.
Published: (2024)
by: Sheikh, Nadia Athar, et al.
Published: (2024)
Enhancing AI Accessibility in Veterinary Medicine: Linking Classifiers and Electronic Health Records
by: Kong, Chun Yin, et al.
Published: (2024)
by: Kong, Chun Yin, et al.
Published: (2024)
Unlocking Electronic Health Records: A Hybrid Graph RAG Approach to Safe Clinical AI for Patient QA
by: Thio, Samuel, et al.
Published: (2025)
by: Thio, Samuel, et al.
Published: (2025)
VeriFact: Verifying Facts in LLM-Generated Clinical Text with Electronic Health Records
by: Chung, Philip, et al.
Published: (2025)
by: Chung, Philip, et al.
Published: (2025)
RAC: Retrieval-Augmented Clarification for Faithful Conversational Search
by: Kebir, Ahmed Rayane, et al.
Published: (2026)
by: Kebir, Ahmed Rayane, et al.
Published: (2026)
CRAFT: Training-Free Cascaded Retrieval for Tabular QA
by: Singh, Adarsh, et al.
Published: (2025)
by: Singh, Adarsh, et al.
Published: (2025)
An Interpretable Deep-Learning Framework for Predicting Hospital Readmissions From Electronic Health Records
by: Azzalini, Fabio, et al.
Published: (2023)
by: Azzalini, Fabio, et al.
Published: (2023)
Tabular PDF Information Extraction with Local LLMs and Layout-Aware Parsing: A Reliability Evaluation
by: Hilmi, Muhammad Anis Al, et al.
Published: (2026)
by: Hilmi, Muhammad Anis Al, et al.
Published: (2026)
Uncovering the Bigger Picture: Comprehensive Event Understanding Via Diverse News Retrieval
by: Tang, Yixuan, et al.
Published: (2025)
by: Tang, Yixuan, et al.
Published: (2025)
OpenExtract: Automated Data Extraction for Systematic Reviews in Health
by: Achterberg, Jim, et al.
Published: (2026)
by: Achterberg, Jim, et al.
Published: (2026)
Beyond Chunk-Then-Embed: A Comprehensive Taxonomy and Evaluation of Document Chunking Strategies for Information Retrieval
by: Zhou, Yongjie, et al.
Published: (2026)
by: Zhou, Yongjie, et al.
Published: (2026)
Data, Not Model: Explaining Bias toward LLM Texts in Neural Retrievers
by: Huang, Wei, et al.
Published: (2026)
by: Huang, Wei, et al.
Published: (2026)
TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understanding
by: Qiang, Minjie, et al.
Published: (2026)
by: Qiang, Minjie, et al.
Published: (2026)
Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integration
by: Dai, Sunhao, et al.
Published: (2024)
by: Dai, Sunhao, et al.
Published: (2024)
Revisiting Prompt Engineering: A Comprehensive Evaluation for LLM-based Personalized Recommendation
by: Kusano, Genki, et al.
Published: (2025)
by: Kusano, Genki, et al.
Published: (2025)
Information Extraction from Historical Well Records Using A Large Language Model
by: Ma, Zhiwei, et al.
Published: (2024)
by: Ma, Zhiwei, et al.
Published: (2024)
Automatic Cardiac Risk Management Classification using large-context Electronic Patients Health Records
by: Vitale, Jacopo, et al.
Published: (2026)
by: Vitale, Jacopo, et al.
Published: (2026)
Comprehensive Evaluation of Matrix Factorization Models for Collaborative Filtering Recommender Systems
by: Bobadilla, Jesús, et al.
Published: (2024)
by: Bobadilla, Jesús, et al.
Published: (2024)
Similar Items
-
ReToP: Learning to Rewrite Electronic Health Records for Clinical Prediction
by: Lovon-Melgarejo, Jesus, et al.
Published: (2026) -
PatientDx: Merging Large Language Models for Protecting Data-Privacy in Healthcare
by: Moreno, Jose G., et al.
Published: (2025) -
Revisiting the MIMIC-IV Benchmark: Experiments Using Language Models for Electronic Health Records
by: Lovon, Jesus, et al.
Published: (2025) -
Jointly Generating and Attributing Answers using Logits of Document-Identifier Tokens
by: Albarede, Lucas, et al.
Published: (2025) -
An Evaluation Framework for Attributed Information Retrieval using Large Language Models
by: Djeddal, Hanane, et al.
Published: (2024)