Gespeichert in:
| Hauptverfasser: | Caglayan, Bora, Wang, Mingxue, Kelleher, John D., Fei, Shen, Tong, Gui, Ding, Jiandong, Zhang, Puchao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2410.22925 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating NL2SQL via SQL2NL
von: Safarzadeh, Mohammadtaher, et al.
Veröffentlicht: (2025)
von: Safarzadeh, Mohammadtaher, et al.
Veröffentlicht: (2025)
ROSE: An Intent-Centered Evaluation Metric for NL2SQL
von: Pei, Wenqi, et al.
Veröffentlicht: (2026)
von: Pei, Wenqi, et al.
Veröffentlicht: (2026)
Blar-SQL: Faster, Stronger, Smaller NL2SQL
von: Domínguez, José Manuel, et al.
Veröffentlicht: (2024)
von: Domínguez, José Manuel, et al.
Veröffentlicht: (2024)
NL2SQLBench: A Modular Benchmarking Framework for LLM-Enabled NL2SQL Solutions
von: Hou, Shizheng, et al.
Veröffentlicht: (2026)
von: Hou, Shizheng, et al.
Veröffentlicht: (2026)
SPENCE: A Syntactic Probe for Detecting Contamination in NL2SQL Benchmarks
von: Safarzadeh, Mohammadtaher, et al.
Veröffentlicht: (2026)
von: Safarzadeh, Mohammadtaher, et al.
Veröffentlicht: (2026)
Agentic NL2SQL to Reduce Computational Costs
von: Jehle, Dominik, et al.
Veröffentlicht: (2025)
von: Jehle, Dominik, et al.
Veröffentlicht: (2025)
SHREC: a SRE Behaviour Knowledge Graph Model for Shell Command Recommendations
von: Tonon, Andrea, et al.
Veröffentlicht: (2024)
von: Tonon, Andrea, et al.
Veröffentlicht: (2024)
Memo-SQL: Structured Decomposition and Experience-Driven Self-Correction for Training-Free NL2SQL
von: Yang, Zerui, et al.
Veröffentlicht: (2026)
von: Yang, Zerui, et al.
Veröffentlicht: (2026)
GeoSQL-Eval: First Evaluation of LLMs on PostGIS-Based NL2GeoSQL Queries
von: Hou, Shuyang, et al.
Veröffentlicht: (2025)
von: Hou, Shuyang, et al.
Veröffentlicht: (2025)
VeriMinder: Mitigating Analytical Vulnerabilities in NL2SQL
von: Mohole, Shubham, et al.
Veröffentlicht: (2025)
von: Mohole, Shubham, et al.
Veröffentlicht: (2025)
OraPlan-SQL: A Planning-Centric Framework for Complex Bilingual NL2SQL Reasoning
von: Liu, Marianne Menglin, et al.
Veröffentlicht: (2025)
von: Liu, Marianne Menglin, et al.
Veröffentlicht: (2025)
SQLong: Enhanced NL2SQL for Longer Contexts with LLMs
von: Nguyen, Dai Quoc, et al.
Veröffentlicht: (2025)
von: Nguyen, Dai Quoc, et al.
Veröffentlicht: (2025)
Distill-C: Enhanced NL2SQL via Distilled Customization with LLMs
von: Hoang, Cong Duy Vu, et al.
Veröffentlicht: (2025)
von: Hoang, Cong Duy Vu, et al.
Veröffentlicht: (2025)
MTIR-SQL: Multi-turn Tool-Integrated Reasoning Reinforcement Learning for Text-to-SQL
von: Xu, Zekun, et al.
Veröffentlicht: (2025)
von: Xu, Zekun, et al.
Veröffentlicht: (2025)
Optimizing Small Language Models for NL2SQL via Chain-of-Thought Fine-Tuning
von: Solanki, Anshul, et al.
Veröffentlicht: (2026)
von: Solanki, Anshul, et al.
Veröffentlicht: (2026)
Feather-SQL: A Lightweight NL2SQL Framework with Dual-Model Collaboration Paradigm for Small Language Models
von: Pei, Wenqi, et al.
Veröffentlicht: (2025)
von: Pei, Wenqi, et al.
Veröffentlicht: (2025)
RubikSQL: Lifelong Learning Agentic Knowledge Base as an Industrial NL2SQL System
von: Chen, Zui, et al.
Veröffentlicht: (2025)
von: Chen, Zui, et al.
Veröffentlicht: (2025)
Is Long Context All You Need? Leveraging LLM's Extended Context for NL2SQL
von: Chung, Yeounoh, et al.
Veröffentlicht: (2025)
von: Chung, Yeounoh, et al.
Veröffentlicht: (2025)
Fact-Consistency Evaluation of Text-to-SQL Generation for Business Intelligence Using Exaone 3.5
von: Choi, Jeho
Veröffentlicht: (2025)
von: Choi, Jeho
Veröffentlicht: (2025)
Dial: A Knowledge-Grounded Dialect-Specific NL2SQL System
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
von: Zhang, Xiang, et al.
Veröffentlicht: (2026)
LLM NL2SQL Robustness: Surface Noise vs. Linguistic Variation in Traditional and Agentic Settings
von: Tu, Lifu, et al.
Veröffentlicht: (2026)
von: Tu, Lifu, et al.
Veröffentlicht: (2026)
LR-SQL: A Supervised Fine-Tuning Method for Text2SQL Tasks under Low-Resource Scenarios
von: Wuzhenghong, Wen, et al.
Veröffentlicht: (2024)
von: Wuzhenghong, Wen, et al.
Veröffentlicht: (2024)
RedParrot: Accelerating NL-to-DSL for Business Analytics via Query Semantic Caching
von: Wang, Tong, et al.
Veröffentlicht: (2026)
von: Wang, Tong, et al.
Veröffentlicht: (2026)
A framework for measuring the training efficiency of a neural architecture
von: Cueto-Mendoza, Eduardo, et al.
Veröffentlicht: (2024)
von: Cueto-Mendoza, Eduardo, et al.
Veröffentlicht: (2024)
Pre-Hoc Predictions in AutoML: Leveraging LLMs to Enhance Model Selection and Benchmarking for Tabular datasets
von: Belkhiter, Yannis, et al.
Veröffentlicht: (2025)
von: Belkhiter, Yannis, et al.
Veröffentlicht: (2025)
TrustSQL: Benchmarking Text-to-SQL Reliability with Penalty-Based Scoring
von: Lee, Gyubok, et al.
Veröffentlicht: (2024)
von: Lee, Gyubok, et al.
Veröffentlicht: (2024)
MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation
von: Shbita, Basel, et al.
Veröffentlicht: (2025)
von: Shbita, Basel, et al.
Veröffentlicht: (2025)
Evaluating Tokenizer Performance of Large Language Models Across Official Indian Languages
von: Tamang, S., et al.
Veröffentlicht: (2024)
von: Tamang, S., et al.
Veröffentlicht: (2024)
ScenicNL: Generating Probabilistic Scenario Programs from Natural Language
von: Elmaaroufi, Karim, et al.
Veröffentlicht: (2024)
von: Elmaaroufi, Karim, et al.
Veröffentlicht: (2024)
NL2SQL-BUGs: A Benchmark for Detecting Semantic Errors in NL2SQL Translation
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
Agent-Agnostic Evaluation of SQL Accuracy in Production Text-to-SQL Systems
von: Arif, Taslim Jamal, et al.
Veröffentlicht: (2026)
von: Arif, Taslim Jamal, et al.
Veröffentlicht: (2026)
Hybrid-NL2SVA: Integrating RAG and Finetuning for LLM-based NL2SVA
von: Xiao, Weihua, et al.
Veröffentlicht: (2025)
von: Xiao, Weihua, et al.
Veröffentlicht: (2025)
Evaluating LLMs for Text-to-SQL Generation With Complex SQL Workload
von: Ma, Limin, et al.
Veröffentlicht: (2024)
von: Ma, Limin, et al.
Veröffentlicht: (2024)
Scaling LLM Planning: NL2FLOW for Parametric Problem Generation and Rigorous Evaluation
von: Kang, Jungkoo
Veröffentlicht: (2025)
von: Kang, Jungkoo
Veröffentlicht: (2025)
SelECT-SQL: Self-correcting ensemble Chain-of-Thought for Text-to-SQL
von: Shen, Ke, et al.
Veröffentlicht: (2024)
von: Shen, Ke, et al.
Veröffentlicht: (2024)
Falcon: A Comprehensive Chinese Text-to-SQL Benchmark for Enterprise-Grade Evaluation
von: Luo, Wenzhen, et al.
Veröffentlicht: (2025)
von: Luo, Wenzhen, et al.
Veröffentlicht: (2025)
Safety2Drive: Safety-Critical Scenario Benchmark for the Evaluation of Autonomous Driving
von: Li, Jingzheng, et al.
Veröffentlicht: (2025)
von: Li, Jingzheng, et al.
Veröffentlicht: (2025)
LocalSearchBench: Benchmarking Agentic Search in Real-World Local Life Services
von: He, Hang, et al.
Veröffentlicht: (2025)
von: He, Hang, et al.
Veröffentlicht: (2025)
BEAVER: An Enterprise Benchmark for Text-to-SQL
von: Chen, Peter Baile, et al.
Veröffentlicht: (2024)
von: Chen, Peter Baile, et al.
Veröffentlicht: (2024)
RESTestBench: A Benchmark for Evaluating the Effectiveness of LLM-Generated REST API Test Cases from NL Requirements
von: Kogler, Leon, et al.
Veröffentlicht: (2026)
von: Kogler, Leon, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Evaluating NL2SQL via SQL2NL
von: Safarzadeh, Mohammadtaher, et al.
Veröffentlicht: (2025) -
ROSE: An Intent-Centered Evaluation Metric for NL2SQL
von: Pei, Wenqi, et al.
Veröffentlicht: (2026) -
Blar-SQL: Faster, Stronger, Smaller NL2SQL
von: Domínguez, José Manuel, et al.
Veröffentlicht: (2024) -
NL2SQLBench: A Modular Benchmarking Framework for LLM-Enabled NL2SQL Solutions
von: Hou, Shizheng, et al.
Veröffentlicht: (2026) -
SPENCE: A Syntactic Probe for Detecting Contamination in NL2SQL Benchmarks
von: Safarzadeh, Mohammadtaher, et al.
Veröffentlicht: (2026)