EHR-MCP: Real-world Evaluation of Clinical Information Retrieval by Large Language Models via Model Context Protocol
Fuente:
arXiv
Saved in:
| Main Authors: | Masayoshi, Kanato, Hashimoto, Masahiro, Yokoyama, Ryoichi, Toda, Naoki, Uwamino, Yoshifumi, Fukuda, Shogo, Namkoong, Ho, Jinzaki, Masahiro |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers
by: Luo, Ziyang, et al.
Published: (2025)
by: Luo, Ziyang, et al.
Published: (2025)
Enhancing Model Context Protocol (MCP) with Context-Aware Server Collaboration
by: Jayanti, Meenakshi Amulya, et al.
Published: (2026)
by: Jayanti, Meenakshi Amulya, et al.
Published: (2026)
ScaleMCP: Dynamic and Auto-Synchronizing Model Context Protocol Tools for LLM Agents
by: Lumer, Elias, et al.
Published: (2025)
by: Lumer, Elias, et al.
Published: (2025)
FinMCP-Bench: Benchmarking LLM Agents for Real-World Financial Tool Use under the Model Context Protocol
by: Zhu, Jie, et al.
Published: (2026)
by: Zhu, Jie, et al.
Published: (2026)
IoT-MCP: Bridging LLMs and IoT Systems Through Model Context Protocol
by: Yang, Ningyuan, et al.
Published: (2025)
by: Yang, Ningyuan, et al.
Published: (2025)
Agent-Based Simulation of a Financial Market with Large Language Models
by: Hashimoto, Ryuji, et al.
Published: (2025)
by: Hashimoto, Ryuji, et al.
Published: (2025)
Balanced Multi-Factor In-Context Learning for Multilingual Large Language Models
by: Kaneko, Masahiro, et al.
Published: (2025)
by: Kaneko, Masahiro, et al.
Published: (2025)
Digital Business Model Analysis Using a Large Language Model
by: Watanabe, Masahiro, et al.
Published: (2024)
by: Watanabe, Masahiro, et al.
Published: (2024)
Validation of musculoskeletal segmentation model with uncertainty estimation for bone and muscle assessment in hip-to-knee clinical CT images
by: Soufi, Mazen, et al.
Published: (2024)
by: Soufi, Mazen, et al.
Published: (2024)
MCIP: Protecting MCP Safety via Model Contextual Integrity Protocol
by: Jing, Huihao, et al.
Published: (2025)
by: Jing, Huihao, et al.
Published: (2025)
RAM-EHR: Retrieval Augmentation Meets Clinical Predictions on Electronic Health Records
by: Xu, Ran, et al.
Published: (2024)
by: Xu, Ran, et al.
Published: (2024)
Enhancing Clinical Decision Support and EHR Insights through LLMs and the Model Context Protocol: An Open-Source MCP-FHIR Framework
by: Ehtesham, Abul, et al.
Published: (2025)
by: Ehtesham, Abul, et al.
Published: (2025)
Parameterizations for Gradient-based Markov Chain Monte Carlo on the Stiefel Manifold: A Comparative Study
by: Tanaka, Masahiro
Published: (2024)
by: Tanaka, Masahiro
Published: (2024)
Delayed Acceptance Markov Chain Monte Carlo for Robust Bayesian Analysis
by: Tanaka, Masahiro
Published: (2025)
by: Tanaka, Masahiro
Published: (2025)
Generalized Posterior Calibration via Sequential Monte Carlo Sampler
by: Tanaka, Masahiro
Published: (2024)
by: Tanaka, Masahiro
Published: (2024)
A Little Leak Will Sink a Great Ship: Survey of Transparency for Large Language Models from Start to Finish
by: Kaneko, Masahiro, et al.
Published: (2024)
by: Kaneko, Masahiro, et al.
Published: (2024)
Technical Report: Small Language Model for Japanese Clinical and Medicine
by: Watanabe, Shogo
Published: (2024)
by: Watanabe, Shogo
Published: (2024)
Eagle: Ethical Dataset Given from Real Interactions
by: Kaneko, Masahiro, et al.
Published: (2024)
by: Kaneko, Masahiro, et al.
Published: (2024)
Modified Delayed Acceptance MCMC for Quasi-Bayesian Inference with Linear Moment Conditions
by: Tanaka, Masahiro
Published: (2025)
by: Tanaka, Masahiro
Published: (2025)
Weighted Particle-Based Optimization for Efficient Generalized Posterior Calibration
by: Tanaka, Masahiro
Published: (2024)
by: Tanaka, Masahiro
Published: (2024)
OUTFOX: LLM-Generated Essay Detection Through In-Context Learning with Adversarially Generated Examples
by: Koike, Ryuto, et al.
Published: (2023)
by: Koike, Ryuto, et al.
Published: (2023)
Real-world validation of a multimodal LLM-powered pipeline for High-Accuracy Clinical Trial Patient Matching leveraging EHR data
by: Callies, Anatole, et al.
Published: (2025)
by: Callies, Anatole, et al.
Published: (2025)
Onco-Retriever: Generative Classifier for Retrieval of EHR Records in Oncology
by: Gupta, Shashi Kant, et al.
Published: (2024)
by: Gupta, Shashi Kant, et al.
Published: (2024)
MCP-SafetyBench: A Benchmark for Safety Evaluation of Large Language Models with Real-World MCP Servers
by: Zong, Xuanjun, et al.
Published: (2025)
by: Zong, Xuanjun, et al.
Published: (2025)
In-Contextual Gender Bias Suppression for Large Language Models
by: Oba, Daisuke, et al.
Published: (2023)
by: Oba, Daisuke, et al.
Published: (2023)
On the Alignment of Large Language Models with Global Human Opinion
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Temporal Entailment Pretraining for Clinical Language Models over EHR Data
by: Tanaka, Tatsunori, et al.
Published: (2025)
by: Tanaka, Tatsunori, et al.
Published: (2025)
AgriBench: A Hierarchical Agriculture Benchmark for Multimodal Large Language Models
by: Zhou, Yutong, et al.
Published: (2024)
by: Zhou, Yutong, et al.
Published: (2024)
Social Bias Evaluation for Large Language Models Requires Prompt Variations
by: Hida, Rem, et al.
Published: (2024)
by: Hida, Rem, et al.
Published: (2024)
Emergence from Emergence: Financial Market Simulation via Learning with Heterogeneous Preferences
by: Hashimoto, Ryuji, et al.
Published: (2025)
by: Hashimoto, Ryuji, et al.
Published: (2025)
Refined and Segmented Price Sentiment Indices from Survey Comments
by: Suzuki, Masahiro, et al.
Published: (2024)
by: Suzuki, Masahiro, et al.
Published: (2024)
From Generative Modeling to Clinical Classification: A GPT-Based Architecture for EHR Notes
by: Irany, Fariba Afrin, et al.
Published: (2026)
by: Irany, Fariba Afrin, et al.
Published: (2026)
Interactive DualChecker for Mitigating Hallucinations in Distilling Large Language Models
by: Wang, Meiyun, et al.
Published: (2024)
by: Wang, Meiyun, et al.
Published: (2024)
Double Glueing over Free Exponential: with Measure Theoretic Applications
by: Hamano, Masahiro
Published: (2021)
by: Hamano, Masahiro
Published: (2021)
Evaluating Gender Bias of Pre-trained Language Models in Natural Language Inference by Considering All Labels
by: Anantaprayoon, Panatchakorn, et al.
Published: (2023)
by: Anantaprayoon, Panatchakorn, et al.
Published: (2023)
Economy Watchers Survey Provides Datasets and Tasks for Japanese Financial Domain
by: Suzuki, Masahiro, et al.
Published: (2024)
by: Suzuki, Masahiro, et al.
Published: (2024)
Paraphrasing Adversarial Attack on LLM-as-a-Reviewer
by: Kaneko, Masahiro
Published: (2026)
by: Kaneko, Masahiro
Published: (2026)
CMT-LLM: Contextual Multi-Talker ASR Utilizing Large Language Models
by: He, Jiajun, et al.
Published: (2025)
by: He, Jiajun, et al.
Published: (2025)
From Heard to Lived Opinions: Simulating Opinion Dynamics with Grounded LLM Agents in Economic Environments
by: Hashimoto, Ryuji, et al.
Published: (2026)
by: Hashimoto, Ryuji, et al.
Published: (2026)
Intent-Aware Self-Correction for Mitigating Social Biases in Large Language Models
by: Anantaprayoon, Panatchakorn, et al.
Published: (2025)
by: Anantaprayoon, Panatchakorn, et al.
Published: (2025)
Similar Items
-
MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers
by: Luo, Ziyang, et al.
Published: (2025) -
Enhancing Model Context Protocol (MCP) with Context-Aware Server Collaboration
by: Jayanti, Meenakshi Amulya, et al.
Published: (2026) -
ScaleMCP: Dynamic and Auto-Synchronizing Model Context Protocol Tools for LLM Agents
by: Lumer, Elias, et al.
Published: (2025) -
FinMCP-Bench: Benchmarking LLM Agents for Real-World Financial Tool Use under the Model Context Protocol
by: Zhu, Jie, et al.
Published: (2026) -
IoT-MCP: Bridging LLMs and IoT Systems Through Model Context Protocol
by: Yang, Ningyuan, et al.
Published: (2025)