Both Ends Count! Just How Good are LLM Agents at "Text-to-Big SQL"?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Eizaguirre, Germán T., Tissen, Lars, Sánchez-Artigas, Marc |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DataFactory: Collaborative Multi-Agent Framework for Advanced Table Question Answering
von: Wang, Tong, et al.
Veröffentlicht: (2026)
von: Wang, Tong, et al.
Veröffentlicht: (2026)
Free Access to World News: Reconstructing Full-Text Articles from GDELT
von: Colladon, A. Fronzetti, et al.
Veröffentlicht: (2025)
von: Colladon, A. Fronzetti, et al.
Veröffentlicht: (2025)
EMR-AGENT: Automating Cohort and Feature Extraction from EMR Databases
von: Lee, Kwanhyung, et al.
Veröffentlicht: (2025)
von: Lee, Kwanhyung, et al.
Veröffentlicht: (2025)
From Search to Reasoning: A Five-Level RAG Capability Framework for Enterprise Data
von: Gill, Gurbinder, et al.
Veröffentlicht: (2025)
von: Gill, Gurbinder, et al.
Veröffentlicht: (2025)
The Science Data Lake: A Unified Open Infrastructure Integrating 293 Million Papers Across Eight Scholarly Sources with Embedding-Based Ontology Alignment
von: Wilinski, Jonas
Veröffentlicht: (2026)
von: Wilinski, Jonas
Veröffentlicht: (2026)
Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations
von: Mandarapu, Madhulatha, et al.
Veröffentlicht: (2026)
von: Mandarapu, Madhulatha, et al.
Veröffentlicht: (2026)
Retrieval and Augmentation of Domain Knowledge for Text-to-SQL Semantic Parsing
von: Patwardhan, Manasi, et al.
Veröffentlicht: (2025)
von: Patwardhan, Manasi, et al.
Veröffentlicht: (2025)
VulCPE: Context-Aware Cybersecurity Vulnerability Retrieval and Management
von: Jiang, Yuning, et al.
Veröffentlicht: (2025)
von: Jiang, Yuning, et al.
Veröffentlicht: (2025)
How Clued up are LLMs? Evaluating Multi-Step Deductive Reasoning in a Text-Based Game Environment
von: Ansell, Rebecca, et al.
Veröffentlicht: (2026)
von: Ansell, Rebecca, et al.
Veröffentlicht: (2026)
TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination
von: Li, Xu, et al.
Veröffentlicht: (2026)
von: Li, Xu, et al.
Veröffentlicht: (2026)
ChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level Generation
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2024)
HELEA: Hard-Negative Benchmark and LLM-based Reranking for Robust Entity Alignment
von: Jang, Yoonjin, et al.
Veröffentlicht: (2026)
von: Jang, Yoonjin, et al.
Veröffentlicht: (2026)
STEP: Stepwise Curriculum Learning for Context-Knowledge Fusion in Conversational Recommendation
von: Yang, Zhenye, et al.
Veröffentlicht: (2025)
von: Yang, Zhenye, et al.
Veröffentlicht: (2025)
Reviewing the Reviewer: Graph-Enhanced LLMs for E-commerce Appeal Adjudication
von: Du, Yuchen, et al.
Veröffentlicht: (2026)
von: Du, Yuchen, et al.
Veröffentlicht: (2026)
Comparison of Unsupervised Metrics for Evaluating Judicial Decision Extraction
von: Litvak, Ivan Leonidovich, et al.
Veröffentlicht: (2025)
von: Litvak, Ivan Leonidovich, et al.
Veröffentlicht: (2025)
BERTopic for Topic Modeling of Hindi Short Texts: A Comparative Study
von: Mutsaddi, Atharva, et al.
Veröffentlicht: (2025)
von: Mutsaddi, Atharva, et al.
Veröffentlicht: (2025)
Can AI Assist in Olympiad Coding
von: Ren, Samuel
Veröffentlicht: (2025)
von: Ren, Samuel
Veröffentlicht: (2025)
TRIZ Agents: A Multi-Agent LLM Approach for TRIZ-Based Innovation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
Open-TI: Open Traffic Intelligence with Augmented Language Model
von: Da, Longchao, et al.
Veröffentlicht: (2023)
von: Da, Longchao, et al.
Veröffentlicht: (2023)
FATHOMS-RAG: A Framework for the Assessment of Thinking and Observation in Multimodal Systems that use Retrieval Augmented Generation
von: Hildebrand, Samuel, et al.
Veröffentlicht: (2025)
von: Hildebrand, Samuel, et al.
Veröffentlicht: (2025)
Collaborative LLM Agents for C4 Software Architecture Design Automation
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
von: Szczepanik, Kamil, et al.
Veröffentlicht: (2025)
Method for Aggregating Unstructured Data Using Large Language Models
von: Lazebnyi, Vsevolod, et al.
Veröffentlicht: (2026)
von: Lazebnyi, Vsevolod, et al.
Veröffentlicht: (2026)
How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval
von: Ashrafi, Nazmus
Veröffentlicht: (2026)
von: Ashrafi, Nazmus
Veröffentlicht: (2026)
HySemRAG: A Hybrid Semantic Retrieval-Augmented Generation Framework for Automated Literature Synthesis and Methodological Gap Analysis
von: Godinez, Alejandro
Veröffentlicht: (2025)
von: Godinez, Alejandro
Veröffentlicht: (2025)
Critical Insights into Leading Conversational AI Models
von: Kohli, Urja, et al.
Veröffentlicht: (2025)
von: Kohli, Urja, et al.
Veröffentlicht: (2025)
When Words Change the Model: Sensitivity of LLMs for Constraint Programming Modelling
von: Pellegrino, Alessio, et al.
Veröffentlicht: (2025)
von: Pellegrino, Alessio, et al.
Veröffentlicht: (2025)
ChatGPT4PCG Competition: Character-like Level Generation for Science Birds
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2023)
von: Taveekitworachai, Pittawat, et al.
Veröffentlicht: (2023)
EMOS: Embodiment-aware Heterogeneous Multi-robot Operating System with LLM Agents
von: Chen, Junting, et al.
Veröffentlicht: (2024)
von: Chen, Junting, et al.
Veröffentlicht: (2024)
ConfProBench: A Confidence Evaluation Benchmark for MLLM-Based Process Judges
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
von: Zhou, Yue, et al.
Veröffentlicht: (2025)
WebArXiv: Evaluating Multimodal Agents on Time-Invariant arXiv Tasks
von: Sun, Zihao, et al.
Veröffentlicht: (2025)
von: Sun, Zihao, et al.
Veröffentlicht: (2025)
Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI
von: Bose, Joy
Veröffentlicht: (2026)
von: Bose, Joy
Veröffentlicht: (2026)
Interactive Text-to-SQL Generation via Editable Step-by-Step Explanations
von: Tian, Yuan, et al.
Veröffentlicht: (2023)
von: Tian, Yuan, et al.
Veröffentlicht: (2023)
HiFi-RAG: Hierarchical Content Filtering and Two-Pass Generation for Open-Domain RAG
von: Nuengsigkapian, Cattalyya
Veröffentlicht: (2025)
von: Nuengsigkapian, Cattalyya
Veröffentlicht: (2025)
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
von: Alshaikh, Moaath, et al.
Veröffentlicht: (2026)
von: Alshaikh, Moaath, et al.
Veröffentlicht: (2026)
MESS+: Dynamically Learned Inference-Time LLM Routing in Model Zoos with Service Level Guarantees
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2025)
von: Woisetschläger, Herbert, et al.
Veröffentlicht: (2025)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
von: Egami, Shusaku, et al.
Veröffentlicht: (2026)
PLUGH: A Benchmark for Spatial Understanding and Reasoning in Large Language Models
von: Tikhonov, Alexey
Veröffentlicht: (2024)
von: Tikhonov, Alexey
Veröffentlicht: (2024)
Tiny QA Benchmark++: Ultra-Lightweight, Synthetic Multilingual Dataset Generation & Smoke-Tests for Continuous LLM Evaluation
von: Koc, Vincent
Veröffentlicht: (2025)
von: Koc, Vincent
Veröffentlicht: (2025)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
von: Palacios, Diego Cabezas
Veröffentlicht: (2026)
Learning Natural Language Constraints for Safe Reinforcement Learning of Language Agents
von: Chua, Jaymari, et al.
Veröffentlicht: (2025)
von: Chua, Jaymari, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DataFactory: Collaborative Multi-Agent Framework for Advanced Table Question Answering
von: Wang, Tong, et al.
Veröffentlicht: (2026) -
Free Access to World News: Reconstructing Full-Text Articles from GDELT
von: Colladon, A. Fronzetti, et al.
Veröffentlicht: (2025) -
EMR-AGENT: Automating Cohort and Feature Extraction from EMR Databases
von: Lee, Kwanhyung, et al.
Veröffentlicht: (2025) -
From Search to Reasoning: A Five-Level RAG Capability Framework for Enterprise Data
von: Gill, Gurbinder, et al.
Veröffentlicht: (2025) -
The Science Data Lake: A Unified Open Infrastructure Integrating 293 Million Papers Across Eight Scholarly Sources with Embedding-Based Ontology Alignment
von: Wilinski, Jonas
Veröffentlicht: (2026)