Falcon: A Comprehensive Chinese Text-to-SQL Benchmark for Enterprise-Grade Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Wenzhen, Guan, Wei, Yao, Yifan, Pan, Yimin, Wang, Feng, Yu, Zhipeng, Wen, Zhe, Chen, Liang, Zhuang, Yihong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BEAVER: An Enterprise Benchmark for Text-to-SQL
by: Chen, Peter Baile, et al.
Published: (2024)
by: Chen, Peter Baile, et al.
Published: (2024)
Text-to-SQL for Enterprise Data Analytics
by: Chen, Albert, et al.
Published: (2025)
by: Chen, Albert, et al.
Published: (2025)
Beyond Text-to-SQL: Can LLMs Really Debug Enterprise ETL SQL?
by: Ye, Jing, et al.
Published: (2026)
by: Ye, Jing, et al.
Published: (2026)
Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows
by: Lei, Fangyu, et al.
Published: (2024)
by: Lei, Fangyu, et al.
Published: (2024)
SIRIUS-SQL: Anchoring Multi-Candidate Text-to-SQL in Execution Feedback
by: Luo, Leo, et al.
Published: (2026)
by: Luo, Leo, et al.
Published: (2026)
VLRMBench: A Comprehensive and Challenging Benchmark for Vision-Language Reward Models
by: Ruan, Jiacheng, et al.
Published: (2025)
by: Ruan, Jiacheng, et al.
Published: (2025)
CSR-RAG: An Efficient Retrieval System for Text-to-SQL on the Enterprise Scale
by: Singh, Rajpreet, et al.
Published: (2026)
by: Singh, Rajpreet, et al.
Published: (2026)
TCMBench: A Comprehensive Benchmark for Evaluating Large Language Models in Traditional Chinese Medicine
by: Yue, Wenjing, et al.
Published: (2024)
by: Yue, Wenjing, et al.
Published: (2024)
CDTP: A Large-Scale Chinese Data-Text Pair Dataset for Comprehensive Evaluation of Chinese LLMs
by: Wu, Chengwei, et al.
Published: (2025)
by: Wu, Chengwei, et al.
Published: (2025)
Fùxì: A Benchmark for Evaluating Language Models on Ancient Chinese Text Understanding and Generation
by: Zhao, Shangqing, et al.
Published: (2025)
by: Zhao, Shangqing, et al.
Published: (2025)
V-SQL: A View-based Two-stage Text-to-SQL Framework
by: You, Zeshun, et al.
Published: (2024)
by: You, Zeshun, et al.
Published: (2024)
TrustSQL: Benchmarking Text-to-SQL Reliability with Penalty-Based Scoring
by: Lee, Gyubok, et al.
Published: (2024)
by: Lee, Gyubok, et al.
Published: (2024)
SQLBench: A Comprehensive Evaluation for Text-to-SQL Capabilities of Large Language Models
by: Zhang, Bin, et al.
Published: (2024)
by: Zhang, Bin, et al.
Published: (2024)
Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs
by: Singh, Gundeep, et al.
Published: (2026)
by: Singh, Gundeep, et al.
Published: (2026)
GenEdit: Compounding Operators and Continuous Improvement to Tackle Text-to-SQL in the Enterprise
by: Maamari, Karime, et al.
Published: (2025)
by: Maamari, Karime, et al.
Published: (2025)
LogicCat: A Chain-of-Thought Text-to-SQL Benchmark for Complex Reasoning
by: Liu, Tao, et al.
Published: (2025)
by: Liu, Tao, et al.
Published: (2025)
AgentArch: A Comprehensive Benchmark to Evaluate Agent Architectures in Enterprise
by: Bogavelli, Tara, et al.
Published: (2025)
by: Bogavelli, Tara, et al.
Published: (2025)
Beyond SELECT: A Comprehensive Taxonomy-Guided Benchmark for Real-World Text-to-SQL Translation
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
Text2SQL-Flow: A Robust SQL-Aware Data Augmentation Framework for Text-to-SQL
by: Cai, Qifeng, et al.
Published: (2025)
by: Cai, Qifeng, et al.
Published: (2025)
Evaluating LLMs for Text-to-SQL Generation With Complex SQL Workload
by: Ma, Limin, et al.
Published: (2024)
by: Ma, Limin, et al.
Published: (2024)
ErrorLLM: Modeling SQL Errors for Text-to-SQL Refinement
by: Hong, Zijin, et al.
Published: (2026)
by: Hong, Zijin, et al.
Published: (2026)
SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQL
by: Lyu, Shuai, et al.
Published: (2025)
by: Lyu, Shuai, et al.
Published: (2025)
Rose-SQL: Role-State Evolution Guided Structured Reasoning for Multi-Turn Text-to-SQL
by: Zhou, Le, et al.
Published: (2026)
by: Zhou, Le, et al.
Published: (2026)
FastAT Benchmark: A Comprehensive Framework for Fair Evaluation of Fast Adversarial Training Methods
by: Pan, Chao, et al.
Published: (2026)
by: Pan, Chao, et al.
Published: (2026)
R$^3$-SQL: Ranking Reward and Resampling for Text-to-SQL
by: Han, Hojae, et al.
Published: (2026)
by: Han, Hojae, et al.
Published: (2026)
PTD-SQL: Partitioning and Targeted Drilling with LLMs in Text-to-SQL
by: Luo, Ruilin, et al.
Published: (2024)
by: Luo, Ruilin, et al.
Published: (2024)
SQLfuse: Enhancing Text-to-SQL Performance through Comprehensive LLM Synergy
by: Zhang, Tingkai, et al.
Published: (2024)
by: Zhang, Tingkai, et al.
Published: (2024)
MAC-SQL: A Multi-Agent Collaborative Framework for Text-to-SQL
by: Wang, Bing, et al.
Published: (2023)
by: Wang, Bing, et al.
Published: (2023)
GradeSQL: Test-Time Inference with Outcome Reward Models for Text-to-SQL Generation from Large Language Models
by: Tritto, Mattia, et al.
Published: (2025)
by: Tritto, Mattia, et al.
Published: (2025)
Rethinking Text-to-SQL: Dynamic Multi-turn SQL Interaction for Real-world Database Exploration
by: Sun, Linzhuang, et al.
Published: (2025)
by: Sun, Linzhuang, et al.
Published: (2025)
Agent-Agnostic Evaluation of SQL Accuracy in Production Text-to-SQL Systems
by: Arif, Taslim Jamal, et al.
Published: (2026)
by: Arif, Taslim Jamal, et al.
Published: (2026)
Taming SQL Complexity: LLM-Based Equivalence Evaluation for Text-to-SQL
by: Zeng, Qingyun, et al.
Published: (2025)
by: Zeng, Qingyun, et al.
Published: (2025)
FloodSQL-Bench: A Retrieval-Augmented Benchmark for Geospatially-Grounded Text-to-SQL
by: Liu, Hanzhou, et al.
Published: (2025)
by: Liu, Hanzhou, et al.
Published: (2025)
LR-SQL: A Supervised Fine-Tuning Method for Text2SQL Tasks under Low-Resource Scenarios
by: Wuzhenghong, Wen, et al.
Published: (2024)
by: Wuzhenghong, Wen, et al.
Published: (2024)
Schema-R1: A reasoning training approach for schema linking in Text-to-SQL Task
by: Wen, Wuzhenghong, et al.
Published: (2025)
by: Wen, Wuzhenghong, et al.
Published: (2025)
Data Caching for Enterprise-Grade Petabyte-Scale OLAP
by: Tang, Chunxu, et al.
Published: (2024)
by: Tang, Chunxu, et al.
Published: (2024)
Memory-Efficient FastText: A Comprehensive Approach Using Double-Array Trie Structures and Mark-Compact Memory Management
by: Du, Yimin
Published: (2025)
by: Du, Yimin
Published: (2025)
ReEx-SQL: Reasoning with Execution-Aware Reinforcement Learning for Text-to-SQL
by: Dai, Yaxun, et al.
Published: (2025)
by: Dai, Yaxun, et al.
Published: (2025)
Prognosis of Acute HEV Infection in Patients With Liver Cirrhosis: A Retrospective Study of 628 Chinese Patients
by: Wen An, et al.
Published: (2024)
by: Wen An, et al.
Published: (2024)
EllieSQL: Cost-Efficient Text-to-SQL with Complexity-Aware Routing
by: Zhu, Yizhang, et al.
Published: (2025)
by: Zhu, Yizhang, et al.
Published: (2025)
Similar Items
-
BEAVER: An Enterprise Benchmark for Text-to-SQL
by: Chen, Peter Baile, et al.
Published: (2024) -
Text-to-SQL for Enterprise Data Analytics
by: Chen, Albert, et al.
Published: (2025) -
Beyond Text-to-SQL: Can LLMs Really Debug Enterprise ETL SQL?
by: Ye, Jing, et al.
Published: (2026) -
Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows
by: Lei, Fangyu, et al.
Published: (2024) -
SIRIUS-SQL: Anchoring Multi-Candidate Text-to-SQL in Execution Feedback
by: Luo, Leo, et al.
Published: (2026)