Lessons Learned on Information Retrieval in Electronic Health Records: A Comparison of Embedding Models and Pooling Strategies
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Myers, Skatje, Miller, Timothy A., Gao, Yanjun, Churpek, Matthew M., Mayampurath, Anoop, Dligach, Dmitriy, Afshar, Majid |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating Retrieval-Augmented Generation vs. Long-Context Input for Clinical Reasoning over EHRs
von: Myers, Skatje, et al.
Veröffentlicht: (2025)
von: Myers, Skatje, et al.
Veröffentlicht: (2025)
Position Paper On Diagnostic Uncertainty Estimation from Large Language Models: Next-Word Probability Is Not Pre-test Probability
von: Gao, Yanjun, et al.
Veröffentlicht: (2024)
von: Gao, Yanjun, et al.
Veröffentlicht: (2024)
When Raw Data Prevails: Are Large Language Model Embeddings Effective in Numerical Data Representation for Medical Machine Learning Applications?
von: Gao, Yanjun, et al.
Veröffentlicht: (2024)
von: Gao, Yanjun, et al.
Veröffentlicht: (2024)
Leveraging Medical Knowledge Graphs Into Large Language Models for Diagnosis Prediction: Design and Application Study
von: Gao, Yanjun, et al.
Veröffentlicht: (2023)
von: Gao, Yanjun, et al.
Veröffentlicht: (2023)
Simple Yet Effective: An Information-Theoretic Approach to Multi-LLM Uncertainty Quantification
von: Kruse, Maya, et al.
Veröffentlicht: (2025)
von: Kruse, Maya, et al.
Veröffentlicht: (2025)
Brittleness and Promise: Knowledge Graph Based Reward Modeling for Diagnostic Reasoning
von: Khatwani, Saksham, et al.
Veröffentlicht: (2025)
von: Khatwani, Saksham, et al.
Veröffentlicht: (2025)
LogosKG: Hardware-Optimized Scalable and Interpretable Knowledge Graph Retrieval
von: Cheng, He, et al.
Veröffentlicht: (2026)
von: Cheng, He, et al.
Veröffentlicht: (2026)
Pooling and Semantic Shift: The Fundamental Challenges in Long Text Embedding and Retrieval
von: Gao, Hang, et al.
Veröffentlicht: (2026)
von: Gao, Hang, et al.
Veröffentlicht: (2026)
Improving Clinical NLP Performance through Language Model-Generated Synthetic Clinical Data
von: Chen, Shan, et al.
Veröffentlicht: (2024)
von: Chen, Shan, et al.
Veröffentlicht: (2024)
ReinPool: Reinforcement Learning Pooling Multi-Vector Embeddings for Retrieval System
von: Cha, Sungguk, et al.
Veröffentlicht: (2026)
von: Cha, Sungguk, et al.
Veröffentlicht: (2026)
CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation
von: Yoon, WonJin, et al.
Veröffentlicht: (2026)
von: Yoon, WonJin, et al.
Veröffentlicht: (2026)
Experience Retrieval-Augmentation with Electronic Health Records Enables Accurate Discharge QA
von: Ou, Justice, et al.
Veröffentlicht: (2025)
von: Ou, Justice, et al.
Veröffentlicht: (2025)
CliniQ: A Multi-faceted Benchmark for Electronic Health Record Retrieval with Semantic Match Assessment
von: Zhao, Zhengyun, et al.
Veröffentlicht: (2025)
von: Zhao, Zhengyun, et al.
Veröffentlicht: (2025)
EHR-RAGp: Retrieval-Augmented Prototype-Guided Foundation Model for Electronic Health Records
von: Shurrab, Saeed, et al.
Veröffentlicht: (2026)
von: Shurrab, Saeed, et al.
Veröffentlicht: (2026)
DR.EHR: Dense Retrieval for Electronic Health Record with Knowledge Injection and Synthetic Data
von: Zhao, Zhengyun, et al.
Veröffentlicht: (2025)
von: Zhao, Zhengyun, et al.
Veröffentlicht: (2025)
Beyond Chunk-Then-Embed: A Comprehensive Taxonomy and Evaluation of Document Chunking Strategies for Information Retrieval
von: Zhou, Yongjie, et al.
Veröffentlicht: (2026)
von: Zhou, Yongjie, et al.
Veröffentlicht: (2026)
Entity Retrieval for Answering Entity-Centric Questions
von: Shavarani, Hassan S., et al.
Veröffentlicht: (2024)
von: Shavarani, Hassan S., et al.
Veröffentlicht: (2024)
Generating Querying Code from Text for Multi-Modal Electronic Health Record
von: ZHang, Mengliang
Veröffentlicht: (2025)
von: ZHang, Mengliang
Veröffentlicht: (2025)
Evaluating LLM Abilities to Understand Tabular Electronic Health Records: A Comprehensive Study of Patient Data Extraction and Retrieval
von: Lovon, Jesus, et al.
Veröffentlicht: (2025)
von: Lovon, Jesus, et al.
Veröffentlicht: (2025)
RAM-EHR: Retrieval Augmentation Meets Clinical Predictions on Electronic Health Records
von: Xu, Ran, et al.
Veröffentlicht: (2024)
von: Xu, Ran, et al.
Veröffentlicht: (2024)
LMK > CLS: Landmark Pooling for Dense Embeddings
von: Doshi, Meet, et al.
Veröffentlicht: (2026)
von: Doshi, Meet, et al.
Veröffentlicht: (2026)
MoMA: A Mixture-of-Multimodal-Agents Architecture for Enhancing Clinical Prediction Modelling
von: Gao, Jifan, et al.
Veröffentlicht: (2025)
von: Gao, Jifan, et al.
Veröffentlicht: (2025)
Retrieval Effectiveness of Enhanced Bibliographic Records.
von: Dillon, Martin, et al.
Veröffentlicht: (1990)
von: Dillon, Martin, et al.
Veröffentlicht: (1990)
TurkEmbed4Retrieval: Turkish Embedding Model for Retrieval Task
von: Ezerceli, Özay, et al.
Veröffentlicht: (2025)
von: Ezerceli, Özay, et al.
Veröffentlicht: (2025)
Towards Reliable Testing for Multiple Information Retrieval System Comparisons
von: Otero, David, et al.
Veröffentlicht: (2025)
von: Otero, David, et al.
Veröffentlicht: (2025)
Enhancing AI Accessibility in Veterinary Medicine: Linking Classifiers and Electronic Health Records
von: Kong, Chun Yin, et al.
Veröffentlicht: (2024)
von: Kong, Chun Yin, et al.
Veröffentlicht: (2024)
Scaling Laws for Embedding Dimension in Information Retrieval
von: Killingback, Julian, et al.
Veröffentlicht: (2026)
von: Killingback, Julian, et al.
Veröffentlicht: (2026)
Unlocking Electronic Health Records: A Hybrid Graph RAG Approach to Safe Clinical AI for Patient QA
von: Thio, Samuel, et al.
Veröffentlicht: (2025)
von: Thio, Samuel, et al.
Veröffentlicht: (2025)
An Interpretable Deep-Learning Framework for Predicting Hospital Readmissions From Electronic Health Records
von: Azzalini, Fabio, et al.
Veröffentlicht: (2023)
von: Azzalini, Fabio, et al.
Veröffentlicht: (2023)
Pooling And Attention: What Are Effective Designs For LLM-Based Embedding Models?
von: Tang, Yixuan, et al.
Veröffentlicht: (2024)
von: Tang, Yixuan, et al.
Veröffentlicht: (2024)
Synergizing Data Imputation and Electronic Health Records for Advancing Prostate Cancer Research: Challenges, and Practical Applications
von: Batouche, Abderrahim Oussama, et al.
Veröffentlicht: (2023)
von: Batouche, Abderrahim Oussama, et al.
Veröffentlicht: (2023)
Comparison of Information Retrieval Techniques Applied to IT Support Tickets
von: Pereira, Leonardo Santiago Benitez, et al.
Veröffentlicht: (2025)
von: Pereira, Leonardo Santiago Benitez, et al.
Veröffentlicht: (2025)
Applying Embedding-Based Retrieval to Airbnb Search
von: Abdool, Mustafa, et al.
Veröffentlicht: (2026)
von: Abdool, Mustafa, et al.
Veröffentlicht: (2026)
Why Large Language Models can Secretly Outperform Embedding Similarity in Information Retrieval
von: Benescu, Matei, et al.
Veröffentlicht: (2026)
von: Benescu, Matei, et al.
Veröffentlicht: (2026)
Graph-Embedding Empowered Entity Retrieval
von: Gerritse, Emma J., et al.
Veröffentlicht: (2025)
von: Gerritse, Emma J., et al.
Veröffentlicht: (2025)
Automated Evaluation of Retrieval-Augmented Language Models with Task-Specific Exam Generation
von: Guinet, Gauthier, et al.
Veröffentlicht: (2024)
von: Guinet, Gauthier, et al.
Veröffentlicht: (2024)
Enhancing Health Information Retrieval with RAG by Prioritizing Topical Relevance and Factual Accuracy
von: Uapadhyay, Rishabh, et al.
Veröffentlicht: (2025)
von: Uapadhyay, Rishabh, et al.
Veröffentlicht: (2025)
A Retrieval Comparison of Six Published Indexes in the Field of Library and Information Science
von: Keen, E. Michael
Veröffentlicht: (1976)
von: Keen, E. Michael
Veröffentlicht: (1976)
Towards System Modelling to Support Diseases Data Extraction from the Electronic Health Records for Physicians Research Activities
von: Alsaqer, Bushra F., et al.
Veröffentlicht: (2024)
von: Alsaqer, Bushra F., et al.
Veröffentlicht: (2024)
Almanac Copilot: Towards Autonomous Electronic Health Record Navigation
von: Zakka, Cyril, et al.
Veröffentlicht: (2024)
von: Zakka, Cyril, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Evaluating Retrieval-Augmented Generation vs. Long-Context Input for Clinical Reasoning over EHRs
von: Myers, Skatje, et al.
Veröffentlicht: (2025) -
Position Paper On Diagnostic Uncertainty Estimation from Large Language Models: Next-Word Probability Is Not Pre-test Probability
von: Gao, Yanjun, et al.
Veröffentlicht: (2024) -
When Raw Data Prevails: Are Large Language Model Embeddings Effective in Numerical Data Representation for Medical Machine Learning Applications?
von: Gao, Yanjun, et al.
Veröffentlicht: (2024) -
Leveraging Medical Knowledge Graphs Into Large Language Models for Diagnosis Prediction: Design and Application Study
von: Gao, Yanjun, et al.
Veröffentlicht: (2023) -
Simple Yet Effective: An Information-Theoretic Approach to Multi-LLM Uncertainty Quantification
von: Kruse, Maya, et al.
Veröffentlicht: (2025)