A State-of-the-Art SQL Reasoning Model using RLVR
Fuente:
arXiv
Saved in:
| Main Authors: | Ali, Alnur, Baheti, Ashutosh, Chang, Jonathan, Chi, Ta-Chung, Cui, Brandon, Drozdov, Andrew, Frankle, Jonathan, Gupta, Abhay, Koppol, Pallavi, Kulinski, Sean, Li, Jonathan, Misra, Dipendra, Opsahl-Ong, Krista, Ortiz, Jose Javier Gonzalez, Zaharia, Matei, Zhang, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
KARL: Knowledge Agents via Reinforcement Learning
by: Chang, Jonathan D., et al.
Published: (2026)
by: Chang, Jonathan D., et al.
Published: (2026)
OfficeQA Pro: An Enterprise Benchmark for End-to-End Grounded Reasoning
by: Opsahl-Ong, Krista, et al.
Published: (2026)
by: Opsahl-Ong, Krista, et al.
Published: (2026)
Optimizing Instructions and Demonstrations for Multi-Stage Language Model Programs
by: Opsahl-Ong, Krista, et al.
Published: (2024)
by: Opsahl-Ong, Krista, et al.
Published: (2024)
Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines
by: Arabzadeh, Negar, et al.
Published: (2026)
by: Arabzadeh, Negar, et al.
Published: (2026)
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
by: Misra, Dipendra, et al.
Published: (2026)
by: Misra, Dipendra, et al.
Published: (2026)
Drowning in Documents: Consequences of Scaling Reranker Inference
by: Jacob, Mathew, et al.
Published: (2024)
by: Jacob, Mathew, et al.
Published: (2024)
Does your data spark joy? Performance gains from domain upsampling at the end of training
by: Blakeney, Cody, et al.
Published: (2024)
by: Blakeney, Cody, et al.
Published: (2024)
Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws
by: Sardana, Nikhil, et al.
Published: (2023)
by: Sardana, Nikhil, et al.
Published: (2023)
Non-Determinism and the Lawlessness of Machine Learning Code
by: Cooper, A. Feder, et al.
Published: (2022)
by: Cooper, A. Feder, et al.
Published: (2022)
Text2SQL is Not Enough: Unifying AI and Databases with TAG
by: Biswal, Asim, et al.
Published: (2024)
by: Biswal, Asim, et al.
Published: (2024)
A Fast Binary Splitting Approach for Non-Adaptive Learning of Erdős--Rényi Graphs
by: Ta, Hoang, et al.
Published: (2025)
by: Ta, Hoang, et al.
Published: (2025)
SQL for dBASE IV: a user's tutorial / Jonathan Sayles with Edwin F. Kerr ; edited by Meva Eringen
by: Sayles, Jonathan
Published: (1933)
by: Sayles, Jonathan
Published: (1933)
SIEVE: Sample-Efficient Parametric Learning from Natural Language
by: Asawa, Parth, et al.
Published: (2026)
by: Asawa, Parth, et al.
Published: (2026)
Automating the Enterprise with Foundation Models
by: Wornow, Michael, et al.
Published: (2024)
by: Wornow, Michael, et al.
Published: (2024)
Provable Interactive Learning with Hindsight Instruction Feedback
by: Misra, Dipendra, et al.
Published: (2024)
by: Misra, Dipendra, et al.
Published: (2024)
Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning
by: Foster, Dylan J., et al.
Published: (2024)
by: Foster, Dylan J., et al.
Published: (2024)
Dataset Reset Policy Optimization for RLHF
by: Chang, Jonathan D., et al.
Published: (2024)
by: Chang, Jonathan D., et al.
Published: (2024)
Robust Validation: Confident Predictions Even When Distributions Shift
by: Cauchois, Maxime, et al.
Published: (2020)
by: Cauchois, Maxime, et al.
Published: (2020)
Loss in French Romantic Art, Literature, and Politics
by: Ribner, Jonathan P.
Published: (2023)
by: Ribner, Jonathan P.
Published: (2023)
Dataset Difficulty and the Role of Inductive Bias
by: Kwok, Devin, et al.
Published: (2024)
by: Kwok, Devin, et al.
Published: (2024)
Dynamic Masking Rate Schedules for MLM Pretraining
by: Ankner, Zachary, et al.
Published: (2023)
by: Ankner, Zachary, et al.
Published: (2023)
Soup to go: mitigating forgetting during continual learning with model averaging
by: Kleiman, Anat, et al.
Published: (2025)
by: Kleiman, Anat, et al.
Published: (2025)
Creating Data Art: Authentic Learning and Visualisation Exhibition
by: Roberts, Jonathan C.
Published: (2024)
by: Roberts, Jonathan C.
Published: (2024)
Fostering Creative Visualisation Skills Through Data-Art Exhibitions
by: Roberts, Jonathan C.
Published: (2024)
by: Roberts, Jonathan C.
Published: (2024)
Engaging Data-Art: Conducting a Public Hands-On Workshop
by: Roberts, Jonathan C.
Published: (2024)
by: Roberts, Jonathan C.
Published: (2024)
ACORN: Performant and Predicate-Agnostic Search Over Vector Embeddings and Structured Data
by: Patel, Liana, et al.
Published: (2024)
by: Patel, Liana, et al.
Published: (2024)
World Model on Million-Length Video And Language With Blockwise RingAttention
by: Liu, Hao, et al.
Published: (2024)
by: Liu, Hao, et al.
Published: (2024)
RAG over Thinking Traces Can Improve Reasoning Tasks
by: Arabzadeh, Negar, et al.
Published: (2026)
by: Arabzadeh, Negar, et al.
Published: (2026)
Weakly Supervised Text-to-SQL Parsing through Question Decomposition
by: Wolfson, Tomer, et al.
Published: (2021)
by: Wolfson, Tomer, et al.
Published: (2021)
The Amate Tradition : Innovation and Dissent in Mexican Art / Jonathan D. Amith
by: D. Amith, Jonathan
by: D. Amith, Jonathan
Experimental Evidence for Longitudinal Scaling Exponent Saturation in Shear Turbulence
by: Gupta, Dipendra, et al.
Published: (2026)
by: Gupta, Dipendra, et al.
Published: (2026)
Fact or Fiction? Improving Fact Verification with Knowledge Graphs through Simplified Subgraph Retrievals
by: Opsahl, Tobias A.
Published: (2024)
by: Opsahl, Tobias A.
Published: (2024)
Thermodynamics of dynamical black holes beyond perturbation theory
by: Ashtekar, Abhay, et al.
Published: (2026)
by: Ashtekar, Abhay, et al.
Published: (2026)
Thermodynamics of Black Holes, far from Equilibrium
by: Ashtekar, Abhay, et al.
Published: (2025)
by: Ashtekar, Abhay, et al.
Published: (2025)
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
by: Saad-Falcon, Jon, et al.
Published: (2023)
by: Saad-Falcon, Jon, et al.
Published: (2023)
mAceReason-Math: A Dataset of High-Quality Multilingual Math Problems Ready For RLVR
by: Dobler, Konstantin, et al.
Published: (2026)
by: Dobler, Konstantin, et al.
Published: (2026)
When Can I Put down My Pen? If I Put down My Pen, Do I Get a Secret Decoder Ring?
by: Frankle, Raymond A.
Published: (1992)
by: Frankle, Raymond A.
Published: (1992)
Policy Improvement using Language Feedback Models
by: Zhong, Victor, et al.
Published: (2024)
by: Zhong, Victor, et al.
Published: (2024)
GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
by: Agrawal, Lakshya A, et al.
Published: (2025)
by: Agrawal, Lakshya A, et al.
Published: (2025)
Petrochemical Fantasies: The Art and Energy of American Comics. By DanielWorden. Columbus, OH : Ohio State University Press. 2024. 212 pp. $34.95 (pbk). ISBN : 9780‐81‐425918‐4
by: Jonathan Cummins
Published: (2026)
by: Jonathan Cummins
Published: (2026)
Similar Items
-
KARL: Knowledge Agents via Reinforcement Learning
by: Chang, Jonathan D., et al.
Published: (2026) -
OfficeQA Pro: An Enterprise Benchmark for End-to-End Grounded Reasoning
by: Opsahl-Ong, Krista, et al.
Published: (2026) -
Optimizing Instructions and Demonstrations for Multi-Stage Language Model Programs
by: Opsahl-Ong, Krista, et al.
Published: (2024) -
Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines
by: Arabzadeh, Negar, et al.
Published: (2026) -
Principled Fine-tuning of LLMs from User-Edits: A Medley of Preference, Supervision, and Reward
by: Misra, Dipendra, et al.
Published: (2026)