GLEAN: Grounded Lightweight Evaluation Anchors for Contamination-Aware Tabular Reasoning
Fuente:
arXiv
Guardado en:
| Autor principal: | Wang, Qizhi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Method for Aggregating Unstructured Data Using Large Language Models
por: Lazebnyi, Vsevolod, et al.
Publicado: (2026)
por: Lazebnyi, Vsevolod, et al.
Publicado: (2026)
SQuARE: Structured Query & Adaptive Retrieval Engine For Tabular Formats
por: Gondhalekar, Chinmay, et al.
Publicado: (2025)
por: Gondhalekar, Chinmay, et al.
Publicado: (2025)
From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction
por: Petrov, Alex, et al.
Publicado: (2026)
por: Petrov, Alex, et al.
Publicado: (2026)
ORPHEAS: A Cross-Lingual Greek-English Embedding Model for Retrieval-Augmented Generation
por: Livieris, Ioannis E., et al.
Publicado: (2026)
por: Livieris, Ioannis E., et al.
Publicado: (2026)
Democratizing GraphRAG: Linear, CPU-Only Graph Retrieval for Multi-Hop QA
por: Wang, Qizhi
Publicado: (2025)
por: Wang, Qizhi
Publicado: (2025)
Heterogeneity in Entity Matching: A Survey and Experimental Analysis
por: Moslemi, Mohammad Hossein, et al.
Publicado: (2025)
por: Moslemi, Mohammad Hossein, et al.
Publicado: (2025)
Enhancing Productivity in Database Management Through AI: A Three-Phase Approach for Database
por: Parashar, Kushagra, et al.
Publicado: (2025)
por: Parashar, Kushagra, et al.
Publicado: (2025)
Fact Grounded Attention: Eliminating Hallucination in Large Language Models Through Attention Level Knowledge Integration
por: Gupta, Aayush
Publicado: (2025)
por: Gupta, Aayush
Publicado: (2025)
HumanMCP: A Human-Like Query Dataset for Evaluating MCP Tool Retrieval Performance
por: Laddha, Shubh, et al.
Publicado: (2025)
por: Laddha, Shubh, et al.
Publicado: (2025)
CUBO: Self-Contained Retrieval-Augmented Generation on Consumer Laptops 10 GB Corpora, 16 GB RAM, Single-Device Deployment
por: Astrino, Paolo
Publicado: (2026)
por: Astrino, Paolo
Publicado: (2026)
EnterpriseRAG-Bench: A RAG Benchmark for Company Internal Knowledge
por: Sun, Yuhong, et al.
Publicado: (2026)
por: Sun, Yuhong, et al.
Publicado: (2026)
AI Agents-as-Judge: Automated Assessment of Accuracy, Consistency, Completeness and Clarity for Enterprise Documents
por: Dasgupta, Sudip, et al.
Publicado: (2025)
por: Dasgupta, Sudip, et al.
Publicado: (2025)
DCD: Domain-Oriented Design for Controlled Retrieval-Augmented Generation
por: Kovalskiy, Valeriy, et al.
Publicado: (2026)
por: Kovalskiy, Valeriy, et al.
Publicado: (2026)
LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics
por: Peyronnet, Antoine, et al.
Publicado: (2026)
por: Peyronnet, Antoine, et al.
Publicado: (2026)
DPDisc: From Factoid Questions to Data Product Requests for Open-World Data Product Discovery over Tables and Text
por: Zhang, Liangliang, et al.
Publicado: (2025)
por: Zhang, Liangliang, et al.
Publicado: (2025)
GE-Chat: A Graph Enhanced RAG Framework for Evidential Response Generation of LLMs
por: Da, Longchao, et al.
Publicado: (2025)
por: Da, Longchao, et al.
Publicado: (2025)
BMAM: Brain-inspired Multi-Agent Memory Framework
por: Li, Yang, et al.
Publicado: (2026)
por: Li, Yang, et al.
Publicado: (2026)
Towards Robust Retrieval-Augmented Generation Based on Knowledge Graph: A Comparative Analysis
por: Amamou, Hazem, et al.
Publicado: (2026)
por: Amamou, Hazem, et al.
Publicado: (2026)
SPARQL Generation with Entity Pre-trained GPT for KG Question Answering
por: Bustamante, Diego, et al.
Publicado: (2024)
por: Bustamante, Diego, et al.
Publicado: (2024)
AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning
por: Patel, Urjitkumar, et al.
Publicado: (2025)
por: Patel, Urjitkumar, et al.
Publicado: (2025)
A Reproducible, Scalable Pipeline for Synthesizing Autoregressive Model Literature
por: Alpay, Faruk, et al.
Publicado: (2025)
por: Alpay, Faruk, et al.
Publicado: (2025)
Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning
por: Pather, Kaviraj, et al.
Publicado: (2025)
por: Pather, Kaviraj, et al.
Publicado: (2025)
A Question Answering Dataset for Temporal-Sensitive Retrieval-Augmented Generation
por: Chen, Ziyang, et al.
Publicado: (2025)
por: Chen, Ziyang, et al.
Publicado: (2025)
Efficient fine-tuning methodology of text embedding models for information retrieval: contrastive learning penalty (clp)
por: Yu, Jeongsu
Publicado: (2024)
por: Yu, Jeongsu
Publicado: (2024)
CHORUS: An Agentic Framework for Generating Realistic Deliberation Data
por: Koursaris, A., et al.
Publicado: (2026)
por: Koursaris, A., et al.
Publicado: (2026)
Recent Advances in Data-Driven Business Process Management
por: Ackermann, Lars, et al.
Publicado: (2024)
por: Ackermann, Lars, et al.
Publicado: (2024)
MeVer at CheckThat! 2026: Cluster-Aware Hard-Negative Mining for Multilingual Scientific-Source Retrieval
por: Bakagianni, Juli, et al.
Publicado: (2026)
por: Bakagianni, Juli, et al.
Publicado: (2026)
TiCard: Deployable EXPLAIN-only Residual Learning for Cardinality Estimation
por: Wang, Qizhi
Publicado: (2025)
por: Wang, Qizhi
Publicado: (2025)
LinkedOut: Linking World Knowledge Representation Out of Video LLM for Next-Generation Video Recommendation
por: Zhang, Haichao, et al.
Publicado: (2025)
por: Zhang, Haichao, et al.
Publicado: (2025)
Improving Graph Embeddings in Machine Learning Using Knowledge Completion with Validation in a Case Study on COVID-19 Spread
por: Napoli, Rosario, et al.
Publicado: (2025)
por: Napoli, Rosario, et al.
Publicado: (2025)
Supporting software engineering tasks with agentic AI: Demonstration on document retrieval and test scenario generation
por: Kica, Marian, et al.
Publicado: (2026)
por: Kica, Marian, et al.
Publicado: (2026)
KNIGHT: Knowledge Graph-Driven Multiple-Choice Question Generation with Adaptive Hardness Calibration
por: Amanlou, Mohammad, et al.
Publicado: (2026)
por: Amanlou, Mohammad, et al.
Publicado: (2026)
A Case Study of Balanced Query Recommendation on Wikipedia
por: Mishra, Harshit, et al.
Publicado: (2025)
por: Mishra, Harshit, et al.
Publicado: (2025)
ByteRover: Agent-Native Memory Through LLM-Curated Hierarchical Context
por: Nguyen, Andy, et al.
Publicado: (2026)
por: Nguyen, Andy, et al.
Publicado: (2026)
IndoBERT-Sentiment: Context-Conditioned Sentiment Classification for Indonesian Text
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026)
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026)
IndoBERT-Relevancy: A Context-Conditioned Relevancy Classifier for Indonesian Text
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026)
por: Saputra, Muhammad Apriandito Arya, et al.
Publicado: (2026)
Structure and Destructure: Dual Forces in the Making of Knowledge Engines
por: Chen, Yihong
Publicado: (2025)
por: Chen, Yihong
Publicado: (2025)
Adaptive Multi-Stage Patent Claim Generation with Unified Quality Assessment
por: Liang, Chen-Wei, et al.
Publicado: (2026)
por: Liang, Chen-Wei, et al.
Publicado: (2026)
Combating data scarcity in recommendation services: Integrating cognitive types of VARK and neural network technologies (LLM)
por: Zmanovskii, Nikita
Publicado: (2026)
por: Zmanovskii, Nikita
Publicado: (2026)
AgenticAI-DialogGen: Topic-Guided Conversation Generation for Fine-Tuning and Evaluating Short- and Long-Term Memories of LLMs
por: Perera, Manoj Madushanka, et al.
Publicado: (2026)
por: Perera, Manoj Madushanka, et al.
Publicado: (2026)
Ejemplares similares
-
Method for Aggregating Unstructured Data Using Large Language Models
por: Lazebnyi, Vsevolod, et al.
Publicado: (2026) -
SQuARE: Structured Query & Adaptive Retrieval Engine For Tabular Formats
por: Gondhalekar, Chinmay, et al.
Publicado: (2025) -
From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction
por: Petrov, Alex, et al.
Publicado: (2026) -
ORPHEAS: A Cross-Lingual Greek-English Embedding Model for Retrieval-Augmented Generation
por: Livieris, Ioannis E., et al.
Publicado: (2026) -
Democratizing GraphRAG: Linear, CPU-Only Graph Retrieval for Multi-Hop QA
por: Wang, Qizhi
Publicado: (2025)