SEAR: Schema-Based Evaluation and Routing for LLM Gateways
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Zecheng, Zheng, Han, Xu, Yue |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
di: Wu, Zihao
Pubblicazione: (2025)
di: Wu, Zihao
Pubblicazione: (2025)
compar:IA: The French Government's LLM arena to collect French-language human prompts and preference data
di: Termignon, Lucie, et al.
Pubblicazione: (2026)
di: Termignon, Lucie, et al.
Pubblicazione: (2026)
Routing End User Queries to Enterprise Databases
di: Sudarshan, Saikrishna, et al.
Pubblicazione: (2026)
di: Sudarshan, Saikrishna, et al.
Pubblicazione: (2026)
Leveraging LLMs to Enable Natural Language Search on Go-to-market Platforms
di: Yao, Jesse, et al.
Pubblicazione: (2024)
di: Yao, Jesse, et al.
Pubblicazione: (2024)
TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments
di: Sakizli, Furkan
Pubblicazione: (2026)
di: Sakizli, Furkan
Pubblicazione: (2026)
AutoBench: Automating LLM Evaluation through Reciprocal Peer Assessment
di: Loi, Dario, et al.
Pubblicazione: (2025)
di: Loi, Dario, et al.
Pubblicazione: (2025)
Evaluating Tool-Using Language Agents: Judge Reliability, Propagation Cascades, and Runtime Mitigation in AgentProp-Bench
di: Gurram, Bhaskar
Pubblicazione: (2026)
di: Gurram, Bhaskar
Pubblicazione: (2026)
AgentPulse: A Continuous Multi-Signal Framework for Evaluating AI Agents in Deployment
di: Gao, Yuxuan, et al.
Pubblicazione: (2026)
di: Gao, Yuxuan, et al.
Pubblicazione: (2026)
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
di: Park, Sungho, et al.
Pubblicazione: (2026)
di: Park, Sungho, et al.
Pubblicazione: (2026)
RubikSQL: Lifelong Learning Agentic Knowledge Base as an Industrial NL2SQL System
di: Chen, Zui, et al.
Pubblicazione: (2025)
di: Chen, Zui, et al.
Pubblicazione: (2025)
HOME-KGQA: A Benchmark Dataset for Multimodal Knowledge Graph Question Answering on Household Daily Activities
di: Egami, Shusaku, et al.
Pubblicazione: (2026)
di: Egami, Shusaku, et al.
Pubblicazione: (2026)
Ada-MK: Adaptive MegaKernel Optimization via Automated DAG-based Search for LLM Inference
di: Dong, Wenxin, et al.
Pubblicazione: (2026)
di: Dong, Wenxin, et al.
Pubblicazione: (2026)
Datrics Text2SQL: A Framework for Natural Language to SQL Query Generation
di: Gladkykh, Tetiana, et al.
Pubblicazione: (2025)
di: Gladkykh, Tetiana, et al.
Pubblicazione: (2025)
Codebase-Memory: Tree-Sitter-Based Knowledge Graphs for LLM Code Exploration via MCP
di: Vogel, Martin, et al.
Pubblicazione: (2026)
di: Vogel, Martin, et al.
Pubblicazione: (2026)
Automating Pharmacovigilance Evidence Generation: Using Large Language Models to Produce Context-Aware SQL
di: Painter, Jeffery L., et al.
Pubblicazione: (2024)
di: Painter, Jeffery L., et al.
Pubblicazione: (2024)
StepCache: Step-Level Reuse with Lightweight Verification and Selective Patching for LLM Serving
di: Nouri, Azam
Pubblicazione: (2026)
di: Nouri, Azam
Pubblicazione: (2026)
Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations
di: Mandarapu, Madhulatha, et al.
Pubblicazione: (2026)
di: Mandarapu, Madhulatha, et al.
Pubblicazione: (2026)
A Grounded Memory System For Smart Personal Assistants
di: Ocker, Felix, et al.
Pubblicazione: (2025)
di: Ocker, Felix, et al.
Pubblicazione: (2025)
GISTBench: Evaluating LLM User Understanding via Evidence-Based Interest Verification
di: Fostiropoulos, Iordanis, et al.
Pubblicazione: (2026)
di: Fostiropoulos, Iordanis, et al.
Pubblicazione: (2026)
Diversification as Risk Minimization
di: Takehi, Rikiya, et al.
Pubblicazione: (2025)
di: Takehi, Rikiya, et al.
Pubblicazione: (2025)
SQL Query Engine: A Self-Healing LLM Pipeline for Natural Language to PostgreSQL Translation
di: Ijaz, Muhammad Adeel
Pubblicazione: (2026)
di: Ijaz, Muhammad Adeel
Pubblicazione: (2026)
SLA Management in Reconfigurable Multi-Agent RAG: A Systems Approach to Question Answering
di: Iannelli, Michael, et al.
Pubblicazione: (2024)
di: Iannelli, Michael, et al.
Pubblicazione: (2024)
Large Language Models for Simultaneous Named Entity Extraction and Spelling Correction
di: Whittaker, Edward, et al.
Pubblicazione: (2024)
di: Whittaker, Edward, et al.
Pubblicazione: (2024)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
di: Pawar, Tejas, et al.
Pubblicazione: (2025)
di: Pawar, Tejas, et al.
Pubblicazione: (2025)
MemForest: An Efficient Agent Memory System with Hierarchical Temporal Indexing
di: Chen, Han, et al.
Pubblicazione: (2026)
di: Chen, Han, et al.
Pubblicazione: (2026)
OpenGloss: A Synthetic Encyclopedic Dictionary and Semantic Knowledge Graph
di: Bommarito II, Michael J.
Pubblicazione: (2025)
di: Bommarito II, Michael J.
Pubblicazione: (2025)
Leveraging Large Language Models for Semantic Query Processing in a Scholarly Knowledge Graph
di: Jia, Runsong, et al.
Pubblicazione: (2024)
di: Jia, Runsong, et al.
Pubblicazione: (2024)
Robust LLM-based Column Type Annotation via Prompt Augmentation with LoRA Tuning
di: Meng, Hanze, et al.
Pubblicazione: (2025)
di: Meng, Hanze, et al.
Pubblicazione: (2025)
The Case for Intent-Based Query Rewriting
di: Nicolai, Gianna Lisa, et al.
Pubblicazione: (2025)
di: Nicolai, Gianna Lisa, et al.
Pubblicazione: (2025)
BatchBench: Toward a Workload-Aware Benchmark for Autoscaling Policies in Big Data Batch Processing -- A Proposed Framework
di: Budigi, Venkata Krishna Prasanth, et al.
Pubblicazione: (2026)
di: Budigi, Venkata Krishna Prasanth, et al.
Pubblicazione: (2026)
Understanding Multi-Agent LLM Frameworks: A Unified Benchmark and Experimental Analysis
di: Orogat, Abdelghny, et al.
Pubblicazione: (2026)
di: Orogat, Abdelghny, et al.
Pubblicazione: (2026)
A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web
di: Egami, Shusaku, et al.
Pubblicazione: (2026)
di: Egami, Shusaku, et al.
Pubblicazione: (2026)
Text-to-SQL based on Large Language Models and Database Keyword Search
di: Nascimento, Eduardo R., et al.
Pubblicazione: (2025)
di: Nascimento, Eduardo R., et al.
Pubblicazione: (2025)
VideoHEDGE: Entropy-Based Hallucination Detection for Video-VLMs via Semantic Clustering and Spatiotemporal Perturbations
di: Gautam, Sushant, et al.
Pubblicazione: (2026)
di: Gautam, Sushant, et al.
Pubblicazione: (2026)
MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition
di: Abtahi, Farhad, et al.
Pubblicazione: (2026)
di: Abtahi, Farhad, et al.
Pubblicazione: (2026)
Democratizing LLMs: An Exploration of Cost-Performance Trade-offs in Self-Refined Open-Source Models
di: Shashidhar, Sumuk, et al.
Pubblicazione: (2023)
di: Shashidhar, Sumuk, et al.
Pubblicazione: (2023)
UnWeaving the knots of GraphRAG -- turns out VectorRAG is almost enough
di: Tuora, Ryszard, et al.
Pubblicazione: (2026)
di: Tuora, Ryszard, et al.
Pubblicazione: (2026)
Neuromem: A Granular Decomposition of the Streaming Lifecycle in External Memory for LLMs
di: Zhang, Ruicheng, et al.
Pubblicazione: (2026)
di: Zhang, Ruicheng, et al.
Pubblicazione: (2026)
RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
When Does Data Augmentation Help? Evaluating LLM and Back-Translation Methods for Hausa and Fongbe NLP
di: Adjovi, Mahounan Pericles, et al.
Pubblicazione: (2026)
di: Adjovi, Mahounan Pericles, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Bhakti: A Lightweight Vector Database Management System for Endowing Large Language Models with Semantic Search Capabilities and Memory
di: Wu, Zihao
Pubblicazione: (2025) -
compar:IA: The French Government's LLM arena to collect French-language human prompts and preference data
di: Termignon, Lucie, et al.
Pubblicazione: (2026) -
Routing End User Queries to Enterprise Databases
di: Sudarshan, Saikrishna, et al.
Pubblicazione: (2026) -
Leveraging LLMs to Enable Natural Language Search on Go-to-market Platforms
di: Yao, Jesse, et al.
Pubblicazione: (2024) -
TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments
di: Sakizli, Furkan
Pubblicazione: (2026)