SCARE: A Benchmark for SQL Correction and Question Answerability Classification for Reliable EHR Question Answering
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Gyubok, Chay, Woosog, Choi, Edward |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TrustSQL: Benchmarking Text-to-SQL Reliability with Penalty-Based Scoring
by: Lee, Gyubok, et al.
Published: (2024)
by: Lee, Gyubok, et al.
Published: (2024)
EHR-SeqSQL : A Sequential Text-to-SQL Dataset For Interactively Exploring Electronic Health Records
by: Ryu, Jaehee, et al.
Published: (2024)
by: Ryu, Jaehee, et al.
Published: (2024)
FHIR-AgentBench: Benchmarking LLM Agents for Realistic Interoperable EHR Question Answering
by: Lee, Gyubok, et al.
Published: (2025)
by: Lee, Gyubok, et al.
Published: (2025)
TCM-Ladder: A Benchmark for Multimodal Question Answering on Traditional Chinese Medicine
by: Xie, Jiacheng, et al.
Published: (2025)
by: Xie, Jiacheng, et al.
Published: (2025)
Comprehensive Evaluation for a Large Scale Knowledge Graph Question Answering Service
by: Potdar, Saloni, et al.
Published: (2025)
by: Potdar, Saloni, et al.
Published: (2025)
MEBench: Benchmarking Large Language Models for Cross-Document Multi-Entity Question Answering
by: Lin, Teng, et al.
Published: (2025)
by: Lin, Teng, et al.
Published: (2025)
Accurate Table Question Answering with Accessible LLMs
by: Jiang, Yangfan, et al.
Published: (2026)
by: Jiang, Yangfan, et al.
Published: (2026)
Knowledge Graph-Guided Multi-Agent Distillation for Reliable Industrial Question Answering with Datasets
by: Pan, Jiqun, et al.
Published: (2025)
by: Pan, Jiqun, et al.
Published: (2025)
Reliable Answers for Recurring Questions: Boosting Text-to-SQL Accuracy with Template Constrained Decoding
by: Jivani, Smit, et al.
Published: (2026)
by: Jivani, Smit, et al.
Published: (2026)
Answerability in Retrieval-Augmented Open-Domain Question Answering
by: Abdumalikov, Rustam, et al.
Published: (2024)
by: Abdumalikov, Rustam, et al.
Published: (2024)
Overview of the EHRSQL 2024 Shared Task on Reliable Text-to-SQL Modeling on Electronic Health Records
by: Lee, Gyubok, et al.
Published: (2024)
by: Lee, Gyubok, et al.
Published: (2024)
Weakly Supervised Text-to-SQL Parsing through Question Decomposition
by: Wolfson, Tomer, et al.
Published: (2021)
by: Wolfson, Tomer, et al.
Published: (2021)
Domain Specific Question to SQL Conversion with Embedded Data Balancing Technique
by: Jyothi, et al.
Published: (2025)
by: Jyothi, et al.
Published: (2025)
STARQA: A Question Answering Dataset for Complex Analytical Reasoning over Structured Databases
by: Maddela, Mounica, et al.
Published: (2025)
by: Maddela, Mounica, et al.
Published: (2025)
Text to Query Plans for Question Answering on Large Tables
by: Zhang, Yipeng, et al.
Published: (2025)
by: Zhang, Yipeng, et al.
Published: (2025)
KU-DMIS at EHRSQL 2024:Generating SQL query via question templatization in EHR
by: Kim, Hajung, et al.
Published: (2024)
by: Kim, Hajung, et al.
Published: (2024)
Towards Question Answering over Large Semi-structured Tables
by: Wang, Yuxiang, et al.
Published: (2025)
by: Wang, Yuxiang, et al.
Published: (2025)
Memory-QA: Answering Recall Questions Based on Multimodal Memories
by: Jiang, Hongda, et al.
Published: (2025)
by: Jiang, Hongda, et al.
Published: (2025)
RetinaQA: A Robust Knowledge Base Question Answering Model for both Answerable and Unanswerable Questions
by: Faldu, Prayushi, et al.
Published: (2024)
by: Faldu, Prayushi, et al.
Published: (2024)
From Conversation to Query Execution: Benchmarking User and Tool Interactions for EHR Database Agents
by: Lee, Gyubok, et al.
Published: (2025)
by: Lee, Gyubok, et al.
Published: (2025)
Automatic Answerability Evaluation for Question Generation
by: Wang, Zifan, et al.
Published: (2023)
by: Wang, Zifan, et al.
Published: (2023)
Q-NL Verifier: Leveraging Synthetic Data for Robust Knowledge Graph Question Answering
by: Schwabe, Tim, et al.
Published: (2025)
by: Schwabe, Tim, et al.
Published: (2025)
Multi-hop Question Answering over Knowledge Graphs using Large Language Models
by: Chakraborty, Abir
Published: (2024)
by: Chakraborty, Abir
Published: (2024)
Leveraging LLM-GNN Integration for Open-World Question Answering over Knowledge Graphs
by: Abdallah, Hussein, et al.
Published: (2026)
by: Abdallah, Hussein, et al.
Published: (2026)
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search
by: Luo, Haoran, et al.
Published: (2025)
by: Luo, Haoran, et al.
Published: (2025)
RouterKGQA: Specialized--General Model Routing for Constraint-Aware Knowledge Graph Question Answering
by: Yuan, Bo, et al.
Published: (2026)
by: Yuan, Bo, et al.
Published: (2026)
Text2SQL-Flow: A Robust SQL-Aware Data Augmentation Framework for Text-to-SQL
by: Cai, Qifeng, et al.
Published: (2025)
by: Cai, Qifeng, et al.
Published: (2025)
BEAVER: An Enterprise Benchmark for Text-to-SQL
by: Chen, Peter Baile, et al.
Published: (2024)
by: Chen, Peter Baile, et al.
Published: (2024)
ReViSQL: Achieving Human-Level Text-to-SQL
by: Zhu, Yuxuan, et al.
Published: (2026)
by: Zhu, Yuxuan, et al.
Published: (2026)
ProPerSim: Developing Proactive and Personalized AI Assistants through User-Assistant Simulation
by: Kim, Jiho, et al.
Published: (2025)
by: Kim, Jiho, et al.
Published: (2025)
V-SQL: A View-based Two-stage Text-to-SQL Framework
by: You, Zeshun, et al.
Published: (2024)
by: You, Zeshun, et al.
Published: (2024)
AmbiSQL: Interactive Ambiguity Detection and Resolution for Text-to-SQL
by: Ding, Zhongjun, et al.
Published: (2025)
by: Ding, Zhongjun, et al.
Published: (2025)
LLM-SQL-Solver: Can LLMs Determine SQL Equivalence?
by: Zhao, Fuheng, et al.
Published: (2023)
by: Zhao, Fuheng, et al.
Published: (2023)
ErrorLLM: Modeling SQL Errors for Text-to-SQL Refinement
by: Hong, Zijin, et al.
Published: (2026)
by: Hong, Zijin, et al.
Published: (2026)
OmniSQL: Synthesizing High-quality Text-to-SQL Data at Scale
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
Towards Unbiased Evaluation of Detecting Unanswerable Questions in EHRSQL
by: Yang, Yongjin, et al.
Published: (2024)
by: Yang, Yongjin, et al.
Published: (2024)
TailorSQL: An NL2SQL System Tailored to Your Query Workload
by: Vaidya, Kapil, et al.
Published: (2025)
by: Vaidya, Kapil, et al.
Published: (2025)
Dubo-SQL: Diverse Retrieval-Augmented Generation and Fine Tuning for Text-to-SQL
by: Thorpe, Dayton G., et al.
Published: (2024)
by: Thorpe, Dayton G., et al.
Published: (2024)
Training Table Question Answering via SQL Query Decomposition
by: Mouravieff, Raphaël, et al.
Published: (2024)
by: Mouravieff, Raphaël, et al.
Published: (2024)
Agentar-Scale-SQL: Advancing Text-to-SQL through Orchestrated Test-Time Scaling
by: Wang, Pengfei, et al.
Published: (2025)
by: Wang, Pengfei, et al.
Published: (2025)
Similar Items
-
TrustSQL: Benchmarking Text-to-SQL Reliability with Penalty-Based Scoring
by: Lee, Gyubok, et al.
Published: (2024) -
EHR-SeqSQL : A Sequential Text-to-SQL Dataset For Interactively Exploring Electronic Health Records
by: Ryu, Jaehee, et al.
Published: (2024) -
FHIR-AgentBench: Benchmarking LLM Agents for Realistic Interoperable EHR Question Answering
by: Lee, Gyubok, et al.
Published: (2025) -
TCM-Ladder: A Benchmark for Multimodal Question Answering on Traditional Chinese Medicine
by: Xie, Jiacheng, et al.
Published: (2025) -
Comprehensive Evaluation for a Large Scale Knowledge Graph Question Answering Service
by: Potdar, Saloni, et al.
Published: (2025)