Beyond SELECT: A Comprehensive Taxonomy-Guided Benchmark for Real-World Text-to-SQL Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Hao, Song, Yuanfeng, Yin, Xiaoming, Chen, Xing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MultiTEND: A Multilingual Benchmark for Natural Language to NoSQL Query Translation
by: Qin, Zhiqian, et al.
Published: (2025)
by: Qin, Zhiqian, et al.
Published: (2025)
DataSage: Multi-agent Collaboration for Insight Discovery with External Knowledge Retrieval, Multi-role Debating, and Multi-path Reasoning
by: Liu, Xiaochuan, et al.
Published: (2025)
by: Liu, Xiaochuan, et al.
Published: (2025)
Falcon: A Comprehensive Chinese Text-to-SQL Benchmark for Enterprise-Grade Evaluation
by: Luo, Wenzhen, et al.
Published: (2025)
by: Luo, Wenzhen, et al.
Published: (2025)
Beyond Query-Level Comparison: Fine-Grained Reinforcement Learning for Text-to-SQL with Automated Interpretable Critiques
by: Wang, Guifeng, et al.
Published: (2025)
by: Wang, Guifeng, et al.
Published: (2025)
Beyond Perfect APIs: A Comprehensive Evaluation of LLM Agents Under Real-World API Complexity
by: Kim, Doyoung, et al.
Published: (2026)
by: Kim, Doyoung, et al.
Published: (2026)
Towards Robustness of Text-to-Visualization Translation against Lexical and Phrasal Variability
by: Lu, Jinwei, et al.
Published: (2024)
by: Lu, Jinwei, et al.
Published: (2024)
Beyond Static Pipelines: Learning Dynamic Workflows for Text-to-SQL
by: Wang, Yihan, et al.
Published: (2026)
by: Wang, Yihan, et al.
Published: (2026)
Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows
by: Lei, Fangyu, et al.
Published: (2024)
by: Lei, Fangyu, et al.
Published: (2024)
C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts
by: Qing, Chenxi, et al.
Published: (2026)
by: Qing, Chenxi, et al.
Published: (2026)
BEAVER: An Enterprise Benchmark for Text-to-SQL
by: Chen, Peter Baile, et al.
Published: (2024)
by: Chen, Peter Baile, et al.
Published: (2024)
EPI-SQL: Enhancing Text-to-SQL Translation with Error-Prevention Instructions
by: Liu, Xiping, et al.
Published: (2024)
by: Liu, Xiping, et al.
Published: (2024)
SQLfuse: Enhancing Text-to-SQL Performance through Comprehensive LLM Synergy
by: Zhang, Tingkai, et al.
Published: (2024)
by: Zhang, Tingkai, et al.
Published: (2024)
EHRSQL: A Practical Text-to-SQL Benchmark for Electronic Health Records
by: Lee, Gyubok, et al.
Published: (2023)
by: Lee, Gyubok, et al.
Published: (2023)
SQLBench: A Comprehensive Evaluation for Text-to-SQL Capabilities of Large Language Models
by: Zhang, Bin, et al.
Published: (2024)
by: Zhang, Bin, et al.
Published: (2024)
Arctic-Text2SQL-R1: Simple Rewards, Strong Reasoning in Text-to-SQL
by: Yao, Zhewei, et al.
Published: (2025)
by: Yao, Zhewei, et al.
Published: (2025)
PTD-SQL: Partitioning and Targeted Drilling with LLMs in Text-to-SQL
by: Luo, Ruilin, et al.
Published: (2024)
by: Luo, Ruilin, et al.
Published: (2024)
Open-SQL Framework: Enhancing Text-to-SQL on Open-source Large Language Models
by: Chen, Xiaojun, et al.
Published: (2024)
by: Chen, Xiaojun, et al.
Published: (2024)
CRED-SQL: Enhancing Real-world Large Scale Database Text-to-SQL Parsing through Cluster Retrieval and Execution Description
by: Duan, Shaoming, et al.
Published: (2025)
by: Duan, Shaoming, et al.
Published: (2025)
Pi-SQL: Enhancing Text-to-SQL with Fine-Grained Guidance from Pivot Programming Languages
by: chi, Yongdong, et al.
Published: (2025)
by: chi, Yongdong, et al.
Published: (2025)
RuleArena: A Benchmark for Rule-Guided Reasoning with LLMs in Real-World Scenarios
by: Zhou, Ruiwen, et al.
Published: (2024)
by: Zhou, Ruiwen, et al.
Published: (2024)
SQL-Trail: Multi-Turn Reinforcement Learning with Interleaved Feedback for Text-to-SQL
by: Hua, Harper, et al.
Published: (2026)
by: Hua, Harper, et al.
Published: (2026)
Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs
by: Singh, Gundeep, et al.
Published: (2026)
by: Singh, Gundeep, et al.
Published: (2026)
Structure Guided Large Language Model for SQL Generation
by: Zhang, Qinggang, et al.
Published: (2024)
by: Zhang, Qinggang, et al.
Published: (2024)
RepoTransBench: A Real-World Multilingual Benchmark for Repository-Level Code Translation
by: Wang, Yanli, et al.
Published: (2024)
by: Wang, Yanli, et al.
Published: (2024)
Ar-Spider: Text-to-SQL in Arabic
by: Almohaimeed, Saleh, et al.
Published: (2024)
by: Almohaimeed, Saleh, et al.
Published: (2024)
RSL-SQL: Robust Schema Linking in Text-to-SQL Generation
by: Cao, Zhenbiao, et al.
Published: (2024)
by: Cao, Zhenbiao, et al.
Published: (2024)
LLMSQL: Upgrading WikiSQL for the LLM Era of Text-to-SQL
by: Pihulski, Dzmitry, et al.
Published: (2025)
by: Pihulski, Dzmitry, et al.
Published: (2025)
SQL-PaLM: Improved Large Language Model Adaptation for Text-to-SQL (extended)
by: Sun, Ruoxi, et al.
Published: (2023)
by: Sun, Ruoxi, et al.
Published: (2023)
PExA: Parallel Exploration Agent for Complex Text-to-SQL
by: Parekh, Tanmay, et al.
Published: (2026)
by: Parekh, Tanmay, et al.
Published: (2026)
DetectRL: Benchmarking LLM-Generated Text Detection in Real-World Scenarios
by: Wu, Junchao, et al.
Published: (2024)
by: Wu, Junchao, et al.
Published: (2024)
IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages
by: Dawar, Aviral, et al.
Published: (2026)
by: Dawar, Aviral, et al.
Published: (2026)
PET-SQL: A Prompt-Enhanced Two-Round Refinement of Text-to-SQL with Cross-consistency
by: Li, Zhishuai, et al.
Published: (2024)
by: Li, Zhishuai, et al.
Published: (2024)
MultiVis-Agent: A Multi-Agent Framework with Logic Rules for Reliable and Comprehensive Cross-Modal Data Visualization
by: Lu, Jinwei, et al.
Published: (2026)
by: Lu, Jinwei, et al.
Published: (2026)
From Intentions to Techniques: A Comprehensive Taxonomy and Challenges in Text Watermarking for Large Language Models
by: Lalai, Harsh Nishant, et al.
Published: (2024)
by: Lalai, Harsh Nishant, et al.
Published: (2024)
TableCache: Primary Foreign Key Guided KV Cache Precomputation for Low Latency Text-to-SQL
by: Su, Jinbo, et al.
Published: (2026)
by: Su, Jinbo, et al.
Published: (2026)
Solid-SQL: Enhanced Schema-linking based In-context Learning for Robust Text-to-SQL
by: Liu, Geling, et al.
Published: (2024)
by: Liu, Geling, et al.
Published: (2024)
ROUTE: Robust Multitask Tuning and Collaboration for Text-to-SQL
by: Qin, Yang, et al.
Published: (2024)
by: Qin, Yang, et al.
Published: (2024)
EviLink: Multi-Path Schema Linking with Uncertainty-Guided Evidence Acquisition for Large-Scale Text-to-SQL
by: Zheng, Huawei, et al.
Published: (2026)
by: Zheng, Huawei, et al.
Published: (2026)
SQLCritic: Correcting Text-to-SQL Generation via Clause-wise Critic
by: Chen, Jikai, et al.
Published: (2025)
by: Chen, Jikai, et al.
Published: (2025)
RealMem: Benchmarking LLMs in Real-World Memory-Driven Interaction
by: Bian, Haonan, et al.
Published: (2026)
by: Bian, Haonan, et al.
Published: (2026)
Similar Items
-
MultiTEND: A Multilingual Benchmark for Natural Language to NoSQL Query Translation
by: Qin, Zhiqian, et al.
Published: (2025) -
DataSage: Multi-agent Collaboration for Insight Discovery with External Knowledge Retrieval, Multi-role Debating, and Multi-path Reasoning
by: Liu, Xiaochuan, et al.
Published: (2025) -
Falcon: A Comprehensive Chinese Text-to-SQL Benchmark for Enterprise-Grade Evaluation
by: Luo, Wenzhen, et al.
Published: (2025) -
Beyond Query-Level Comparison: Fine-Grained Reinforcement Learning for Text-to-SQL with Automated Interpretable Critiques
by: Wang, Guifeng, et al.
Published: (2025) -
Beyond Perfect APIs: A Comprehensive Evaluation of LLM Agents Under Real-World API Complexity
by: Kim, Doyoung, et al.
Published: (2026)