ECG-Reasoning-Benchmark: A Benchmark for Evaluating Clinical Reasoning Capabilities in ECG Interpretation
Fuente:
arXiv
Saved in:
| Main Authors: | Oh, Jungwoo, Chung, Hyunseung, Lee, Junhee, Kim, Min-Gyu, Yoon, Hangyul, Lee, Ki Seong, Lee, Youngchae, Yeo, Muhan, Choi, Edward |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ECG-Agent: On-Device Tool-Calling Agent for ECG Multi-Turn Dialogue
by: Chung, Hyunseung, et al.
Published: (2026)
by: Chung, Hyunseung, et al.
Published: (2026)
CXReasonBench: A Benchmark for Evaluating Structured Diagnostic Reasoning in Chest X-rays
by: Lee, Hyungyung, et al.
Published: (2025)
by: Lee, Hyungyung, et al.
Published: (2025)
CXReasonAgent: Evidence-Grounded Diagnostic Reasoning Agent for Chest X-rays
by: Lee, Hyungyung, et al.
Published: (2026)
by: Lee, Hyungyung, et al.
Published: (2026)
Modeling Clinical Uncertainty in Radiology Reports: from Explicit Uncertainty Markers to Implicit Reasoning Pathways
by: Rabaey, Paloma, et al.
Published: (2025)
by: Rabaey, Paloma, et al.
Published: (2025)
CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation
by: Khatibi, Elahe, et al.
Published: (2026)
by: Khatibi, Elahe, et al.
Published: (2026)
Lunguage: A Benchmark for Structured and Sequential Chest X-ray Interpretation
by: Moon, Jong Hak, et al.
Published: (2025)
by: Moon, Jong Hak, et al.
Published: (2025)
Towards Error-Free EHRs: Reasoning-Intensive Consistency Verification Between Clinical Notes and Structured Tables in Electronic Health Records
by: Kwon, Yeonsu, et al.
Published: (2026)
by: Kwon, Yeonsu, et al.
Published: (2026)
RAISE: Enhancing Scientific Reasoning in LLMs via Step-by-Step Retrieval
by: Oh, Minhae, et al.
Published: (2025)
by: Oh, Minhae, et al.
Published: (2025)
EHRNoteQA: An LLM Benchmark for Real-World Clinical Practice Using Discharge Summaries
by: Kweon, Sunjun, et al.
Published: (2024)
by: Kweon, Sunjun, et al.
Published: (2024)
Vib2ECG: A Paired Chest-Lead SCG-ECG Dataset and Benchmark for ECG Reconstruction
by: Lu, Guorui, et al.
Published: (2026)
by: Lu, Guorui, et al.
Published: (2026)
DPP‐DTT Nanowire Phototransistors for Optoelectronic Synapses in EMG and ECG Signal Classification
by: Wangmyung Choi, et al.
Published: (2025)
by: Wangmyung Choi, et al.
Published: (2025)
ECG-biometrics-bench: A Unified Framework for Reproducible Benchmarking of ECG Biometrics
by: Parvan, Milad
Published: (2026)
by: Parvan, Milad
Published: (2026)
ARCTraj: A Dataset and Benchmark of Human Reasoning Trajectories for Abstract Problem Solving
by: Kim, Sejin, et al.
Published: (2025)
by: Kim, Sejin, et al.
Published: (2025)
OpenECG: Benchmarking ECG Foundation Models with Public 1.2 Million Records
by: Wan, Zhijiang, et al.
Published: (2025)
by: Wan, Zhijiang, et al.
Published: (2025)
Benchmarking ECG FMs: A Reality Check Across Clinical Tasks
by: Al-Masud, M A, et al.
Published: (2025)
by: Al-Masud, M A, et al.
Published: (2025)
Towards Efficient Visual-Language Alignment of the Q-Former for Visual Reasoning Tasks
by: Kim, Sungkyung, et al.
Published: (2024)
by: Kim, Sungkyung, et al.
Published: (2024)
KoBALT: Korean Benchmark For Advanced Linguistic Tasks
by: Shin, Hyopil, et al.
Published: (2025)
by: Shin, Hyopil, et al.
Published: (2025)
DPP‐DTT Nanowire Phototransistors for Optoelectronic Synapses in EMG and ECG Signal Classification (Small 48/2025)
by: Wangmyung Choi, et al.
Published: (2025)
by: Wangmyung Choi, et al.
Published: (2025)
From Static Benchmarks to Dynamic Protocol: Agent-Centric Text Anomaly Detection for Evaluating LLM Reasoning
by: Yoa, Seungdong, et al.
Published: (2026)
by: Yoa, Seungdong, et al.
Published: (2026)
K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology
by: Kim, Soyeon, et al.
Published: (2026)
by: Kim, Soyeon, et al.
Published: (2026)
Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History
by: Kim, Serin, et al.
Published: (2026)
by: Kim, Serin, et al.
Published: (2026)
Consciousness-ECG Transformer for Conscious State Estimation System with Real-Time Monitoring
by: Kweon, Young-Seok, et al.
Published: (2025)
by: Kweon, Young-Seok, et al.
Published: (2025)
CIRCUIT: A Benchmark for Circuit Interpretation and Reasoning Capabilities of LLMs
by: Skelic, Lejla, et al.
Published: (2025)
by: Skelic, Lejla, et al.
Published: (2025)
STAR: Detecting Inference-time Backdoors in LLM Reasoning via State-Transition Amplification Ratio
by: Park, Seong-Gyu, et al.
Published: (2026)
by: Park, Seong-Gyu, et al.
Published: (2026)
The Curious Case of Analogies: Investigating Analogical Reasoning in Large Language Models
by: Lee, Taewhoo, et al.
Published: (2025)
by: Lee, Taewhoo, et al.
Published: (2025)
SemiSegECG: A Multi-Dataset Benchmark for Semi-Supervised Semantic Segmentation in ECG Delineation
by: Park, Minje, et al.
Published: (2025)
by: Park, Minje, et al.
Published: (2025)
Korean Canonical Legal Benchmark: Toward Knowledge-Independent Evaluation of LLMs' Legal Reasoning Capabilities
by: Oh, Hongseok, et al.
Published: (2025)
by: Oh, Hongseok, et al.
Published: (2025)
PILOT-Bench: A Benchmark for Legal Reasoning in the Patent Domain with IRAC-Aligned Classification Tasks
by: Jang, Yehoon, et al.
Published: (2026)
by: Jang, Yehoon, et al.
Published: (2026)
Latent Self-Consistency for Reliable Majority-Set Selection in Short- and Long-Answer Reasoning
by: Oh, Jungsuk, et al.
Published: (2025)
by: Oh, Jungsuk, et al.
Published: (2025)
ALFRED: Ask a Large-language model For Reliable ECG Diagnosis
by: Yu, Jin, et al.
Published: (2025)
by: Yu, Jin, et al.
Published: (2025)
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
by: Kim, Kyuhee, et al.
Published: (2025)
by: Kim, Kyuhee, et al.
Published: (2025)
EVL-ECG: Efficient ECG Interpretation With Multi-Aspect Heterogeneous Knowledge Distillation
by: Hong, Dang Nguyen, et al.
Published: (2026)
by: Hong, Dang Nguyen, et al.
Published: (2026)
Pedagogy-R1: Pedagogically-Aligned Reasoning Model with Balanced Educational Benchmark
by: Lee, Unggi, et al.
Published: (2025)
by: Lee, Unggi, et al.
Published: (2025)
RFEval: Benchmarking Reasoning Faithfulness under Counterfactual Reasoning Intervention in Large Reasoning Models
by: Han, Yunseok, et al.
Published: (2026)
by: Han, Yunseok, et al.
Published: (2026)
PICon: A Multi-Turn Interrogation Framework for Evaluating Persona Agent Consistency
by: Kim, Minseo, et al.
Published: (2026)
by: Kim, Minseo, et al.
Published: (2026)
Meta-cognitive Multi-scale Hierarchical Reasoning for Motor Imagery Decoding
by: Kim, Si-Hyun, et al.
Published: (2025)
by: Kim, Si-Hyun, et al.
Published: (2025)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
by: Oh, Yeongtak, et al.
Published: (2026)
by: Oh, Yeongtak, et al.
Published: (2026)
VisDoT : Enhancing Visual Reasoning through Human-Like Interpretation Grounding and Decomposition of Thought
by: Lee, Eunsoo, et al.
Published: (2026)
by: Lee, Eunsoo, et al.
Published: (2026)
ECG-R1: Protocol-Guided and Modality-Agnostic MLLM for Reliable ECG Interpretation
by: Jin, Jiarui, et al.
Published: (2026)
by: Jin, Jiarui, et al.
Published: (2026)
Optimizing Neural Network Scale for ECG Classification
by: Lee, Byeong Tak, et al.
Published: (2023)
by: Lee, Byeong Tak, et al.
Published: (2023)
Similar Items
-
ECG-Agent: On-Device Tool-Calling Agent for ECG Multi-Turn Dialogue
by: Chung, Hyunseung, et al.
Published: (2026) -
CXReasonBench: A Benchmark for Evaluating Structured Diagnostic Reasoning in Chest X-rays
by: Lee, Hyungyung, et al.
Published: (2025) -
CXReasonAgent: Evidence-Grounded Diagnostic Reasoning Agent for Chest X-rays
by: Lee, Hyungyung, et al.
Published: (2026) -
Modeling Clinical Uncertainty in Radiology Reports: from Explicit Uncertainty Markers to Implicit Reasoning Pathways
by: Rabaey, Paloma, et al.
Published: (2025) -
CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation
by: Khatibi, Elahe, et al.
Published: (2026)