Agentic Jackal: Live Execution and Semantic Value Grounding for Text-to-JQL
Fuente:
arXiv
Saved in:
| Main Authors: | Murali, Vishnu, Gulati, Anmol, Lumer, Elias, Frank, Kevin, Campagna, Sindy, Subbiah, Vamse Kumar |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Jackal: A Real-World Execution-Based Benchmark Evaluating Large Language Models on Text-to-JQL Tasks
by: Frank, Kevin, et al.
Published: (2025)
by: Frank, Kevin, et al.
Published: (2025)
Chronos: Temporal-Aware Conversational Agents with Structured Event Retrieval for Long-Term Memory
by: Sen, Sahil, et al.
Published: (2026)
by: Sen, Sahil, et al.
Published: (2026)
Is Grep All You Need? How Agent Harnesses Reshape Agentic Search
by: Sen, Sahil, et al.
Published: (2026)
by: Sen, Sahil, et al.
Published: (2026)
Don't Break the Cache: An Evaluation of Prompt Caching for Long-Horizon Agentic Tasks
by: Lumer, Elias, et al.
Published: (2026)
by: Lumer, Elias, et al.
Published: (2026)
Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?
by: Gulati, Anmol, et al.
Published: (2026)
by: Gulati, Anmol, et al.
Published: (2026)
Agent-as-a-Graph: Knowledge Graph-Based Tool and Agent Retrieval for LLM Multi-Agent Systems
by: Nizar, Faheem, et al.
Published: (2025)
by: Nizar, Faheem, et al.
Published: (2025)
Tool-to-Agent Retrieval: Bridging Tools and Agents for Scalable LLM Multi-Agent Systems
by: Lumer, Elias, et al.
Published: (2025)
by: Lumer, Elias, et al.
Published: (2025)
ScaleMCP: Dynamic and Auto-Synchronizing Model Context Protocol Tools for LLM Agents
by: Lumer, Elias, et al.
Published: (2025)
by: Lumer, Elias, et al.
Published: (2025)
MemTool: Optimizing Short-Term Memory Management for Dynamic Tool Calling in LLM Agent Multi-Turn Conversations
by: Lumer, Elias, et al.
Published: (2025)
by: Lumer, Elias, et al.
Published: (2025)
Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents
by: Onweller, Hailey, et al.
Published: (2026)
by: Onweller, Hailey, et al.
Published: (2026)
Graph RAG-Tool Fusion
by: Lumer, Elias, et al.
Published: (2025)
by: Lumer, Elias, et al.
Published: (2025)
Toolshed: Scale Tool-Equipped Agents with Advanced RAG-Tool Fusion and Tool Knowledge Bases
by: Lumer, Elias, et al.
Published: (2024)
by: Lumer, Elias, et al.
Published: (2024)
Comparison of Text-Based and Image-Based Retrieval in Multimodal Retrieval Augmented Generation Large Language Model Systems
by: Lumer, Elias, et al.
Published: (2025)
by: Lumer, Elias, et al.
Published: (2025)
Rethinking Retrieval: From Traditional Retrieval Augmented Generation to Agentic and Non-Vector Reasoning Systems in the Financial Domain for Large Language Models
by: Lumer, Elias, et al.
Published: (2025)
by: Lumer, Elias, et al.
Published: (2025)
Beyond Rows to Reasoning: Agentic Retrieval for Multimodal Spreadsheet Understanding and Editing
by: Gulati, Anmol, et al.
Published: (2026)
by: Gulati, Anmol, et al.
Published: (2026)
Simulation Agent: A Framework for Integrating Simulation and Large Language Models for Enhanced Decision-Making
by: Kleiman, Jacob, et al.
Published: (2025)
by: Kleiman, Jacob, et al.
Published: (2025)
From Rows to Reasoning: A Retrieval-Augmented Multimodal Framework for Spreadsheet Understanding
by: Gulati, Anmol, et al.
Published: (2026)
by: Gulati, Anmol, et al.
Published: (2026)
EGREFINE: An Execution-Grounded Optimization Framework for Text-to-SQL Schema Refinement
by: Wang, Jiaqian, et al.
Published: (2026)
by: Wang, Jiaqian, et al.
Published: (2026)
ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution
by: Huang, Shouzheng, et al.
Published: (2026)
by: Huang, Shouzheng, et al.
Published: (2026)
Socratic Reasoning Improves Positive Text Rewriting
by: Goel, Anmol, et al.
Published: (2024)
by: Goel, Anmol, et al.
Published: (2024)
HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation
by: Deng, Zewei, et al.
Published: (2026)
by: Deng, Zewei, et al.
Published: (2026)
ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs
by: Wang, Zhipin, et al.
Published: (2026)
by: Wang, Zhipin, et al.
Published: (2026)
Fashion Recommendation: Outfit Compatibility using GNN
by: Gulati, Samaksh
Published: (2024)
by: Gulati, Samaksh
Published: (2024)
AdvSumm: Adversarial Training for Bias Mitigation in Text Summarization
by: Gupta, Mukur, et al.
Published: (2025)
by: Gupta, Mukur, et al.
Published: (2025)
Code Execution as Grounded Supervision for LLM Reasoning
by: Jung, Dongwon, et al.
Published: (2025)
by: Jung, Dongwon, et al.
Published: (2025)
Scaling Agentic Capabilities via Grounded Interaction Synthesis
by: Shi, Wenhang, et al.
Published: (2026)
by: Shi, Wenhang, et al.
Published: (2026)
Differentially Private Knowledge Distillation via Synthetic Text Generation
by: Flemings, James, et al.
Published: (2024)
by: Flemings, James, et al.
Published: (2024)
WebLists: Extracting Structured Information From Complex Interactive Websites Using Executable LLM Agents
by: Bohra, Arth, et al.
Published: (2025)
by: Bohra, Arth, et al.
Published: (2025)
From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning
by: Gado, Ahmed Y., et al.
Published: (2026)
by: Gado, Ahmed Y., et al.
Published: (2026)
Semantic Invariance in Agentic AI
by: de Zarzà, I., et al.
Published: (2026)
by: de Zarzà, I., et al.
Published: (2026)
PhysicsSolutionAgent: Towards Multimodal Explanations for Numerical Physics Problem Solving
by: Thole, Aditya, et al.
Published: (2026)
by: Thole, Aditya, et al.
Published: (2026)
DocDancer: Towards Agentic Document-Grounded Information Seeking
by: Zhang, Qintong, et al.
Published: (2026)
by: Zhang, Qintong, et al.
Published: (2026)
Towards Corpus-Grounded Agentic LLMs for Multilingual Grammatical Analysis
by: Klemen, Matej, et al.
Published: (2025)
by: Klemen, Matej, et al.
Published: (2025)
Verus-SpecGym: An Agentic Environment for Evaluating Specification Autoformalization
by: Agarwal, Anmol, et al.
Published: (2026)
by: Agarwal, Anmol, et al.
Published: (2026)
ToolGate: Contract-Grounded and Verified Tool Execution for LLMs
by: Liu, Yanming, et al.
Published: (2026)
by: Liu, Yanming, et al.
Published: (2026)
MemoBrain: Executive Memory as an Agentic Brain for Reasoning
by: Qian, Hongjin, et al.
Published: (2026)
by: Qian, Hongjin, et al.
Published: (2026)
Intrinsic Guardrails: How Semantic Geometry of Personality Interacts with Emergent Misalignment in LLMs
by: Aneja, Krishak, et al.
Published: (2026)
by: Aneja, Krishak, et al.
Published: (2026)
Towards Execution-Grounded Automated AI Research
by: Si, Chenglei, et al.
Published: (2026)
by: Si, Chenglei, et al.
Published: (2026)
RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
by: Gehring, Jonas, et al.
Published: (2024)
by: Gehring, Jonas, et al.
Published: (2024)
Nexus: Execution-Grounded Multi-Agent Test Oracle Synthesis
by: Huang, Dong, et al.
Published: (2025)
by: Huang, Dong, et al.
Published: (2025)
Similar Items
-
Jackal: A Real-World Execution-Based Benchmark Evaluating Large Language Models on Text-to-JQL Tasks
by: Frank, Kevin, et al.
Published: (2025) -
Chronos: Temporal-Aware Conversational Agents with Structured Event Retrieval for Long-Term Memory
by: Sen, Sahil, et al.
Published: (2026) -
Is Grep All You Need? How Agent Harnesses Reshape Agentic Search
by: Sen, Sahil, et al.
Published: (2026) -
Don't Break the Cache: An Evaluation of Prompt Caching for Long-Horizon Agentic Tasks
by: Lumer, Elias, et al.
Published: (2026) -
Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?
by: Gulati, Anmol, et al.
Published: (2026)