TransientTables: Evaluating LLMs' Reasoning on Temporally Evolving Semi-structured Tables
Fuente:
arXiv
Salvato in:
| Autori principali: | Shankarampeta, Abhilash, Mahajan, Harsh, Kataria, Tushar, Roth, Dan, Gupta, Vivek |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Evidence-Guided Schema Normalization for Temporal Tabular Reasoning
di: Thanga, Ashish, et al.
Pubblicazione: (2025)
di: Thanga, Ashish, et al.
Pubblicazione: (2025)
NTSEBENCH: Cognitive Reasoning Benchmark for Vision Language Models
di: Pandya, Pranshu, et al.
Pubblicazione: (2024)
di: Pandya, Pranshu, et al.
Pubblicazione: (2024)
Evaluating Concurrent Robustness of Language Models Across Diverse Challenge Sets
di: Gupta, Vatsal, et al.
Pubblicazione: (2023)
di: Gupta, Vatsal, et al.
Pubblicazione: (2023)
CORE-T: COherent REtrieval of Tables for Text-to-SQL
di: Soliman, Hassan, et al.
Pubblicazione: (2026)
di: Soliman, Hassan, et al.
Pubblicazione: (2026)
FlowVQA: Mapping Multimodal Logic in Visual Question Answering with Flowcharts
di: Singh, Shubhankar, et al.
Pubblicazione: (2024)
di: Singh, Shubhankar, et al.
Pubblicazione: (2024)
Is Table Retrieval a Solved Problem? Exploring Join-Aware Multi-Table Retrieval
di: Chen, Peter Baile, et al.
Pubblicazione: (2024)
di: Chen, Peter Baile, et al.
Pubblicazione: (2024)
Enhancing Temporal Understanding in LLMs for Semi-structured Tables
di: Deng, Irwin, et al.
Pubblicazione: (2024)
di: Deng, Irwin, et al.
Pubblicazione: (2024)
ADAM: A Diverse Archive of Mankind for Evaluating and Enhancing LLMs in Biographical Reasoning
di: Cekinmez, Jasin, et al.
Pubblicazione: (2025)
di: Cekinmez, Jasin, et al.
Pubblicazione: (2025)
TabSQLify: Enhancing Reasoning Capabilities of LLMs Through Table Decomposition
di: Nahid, Md Mahadi Hasan, et al.
Pubblicazione: (2024)
di: Nahid, Md Mahadi Hasan, et al.
Pubblicazione: (2024)
CRAFT: Training-Free Cascaded Retrieval for Tabular QA
di: Singh, Adarsh, et al.
Pubblicazione: (2025)
di: Singh, Adarsh, et al.
Pubblicazione: (2025)
Knowledge-Aware Reasoning over Multimodal Semi-structured Tables
di: Mathur, Suyash Vardhan, et al.
Pubblicazione: (2024)
di: Mathur, Suyash Vardhan, et al.
Pubblicazione: (2024)
Leveraging LLM For Synchronizing Information Across Multilingual Tables
di: Khincha, Siddharth, et al.
Pubblicazione: (2025)
di: Khincha, Siddharth, et al.
Pubblicazione: (2025)
Federated Retrieval-Augmented Generation: A Systematic Mapping Study
di: Chakraborty, Abhijit, et al.
Pubblicazione: (2025)
di: Chakraborty, Abhijit, et al.
Pubblicazione: (2025)
EnrichIndex: Using LLMs to Enrich Retrieval Indices Offline
di: Chen, Peter Baile, et al.
Pubblicazione: (2025)
di: Chen, Peter Baile, et al.
Pubblicazione: (2025)
State Space Models are Strong Text Rerankers
di: Xu, Zhichao, et al.
Pubblicazione: (2024)
di: Xu, Zhichao, et al.
Pubblicazione: (2024)
Knowledge in Triples for LLMs: Enhancing Table QA Accuracy with Semantic Extraction
di: Sholehrasa, Hossein, et al.
Pubblicazione: (2024)
di: Sholehrasa, Hossein, et al.
Pubblicazione: (2024)
Weaver: Interweaving SQL and LLM for Table Reasoning
di: Khoja, Rohit, et al.
Pubblicazione: (2025)
di: Khoja, Rohit, et al.
Pubblicazione: (2025)
From Facts to Conclusions : Integrating Deductive Reasoning in Retrieval-Augmented LLMs
di: Mishra, Shubham, et al.
Pubblicazione: (2025)
di: Mishra, Shubham, et al.
Pubblicazione: (2025)
TableRAG: A Retrieval Augmented Generation Framework for Heterogeneous Document Reasoning
di: Yu, Xiaohan, et al.
Pubblicazione: (2025)
di: Yu, Xiaohan, et al.
Pubblicazione: (2025)
RUST-BENCH: Benchmarking LLM Reasoning on Unstructured Text within Structured Tables
di: Abhyankar, Nikhil, et al.
Pubblicazione: (2025)
di: Abhyankar, Nikhil, et al.
Pubblicazione: (2025)
Combining Language and Graph Models for Semi-structured Information Extraction on the Web
di: Hong, Zhi, et al.
Pubblicazione: (2024)
di: Hong, Zhi, et al.
Pubblicazione: (2024)
Causal-Counterfactual RAG: The Integration of Causal-Counterfactual Reasoning into RAG
di: Khadilkar, Harshad, et al.
Pubblicazione: (2025)
di: Khadilkar, Harshad, et al.
Pubblicazione: (2025)
APEX-MEM: Agentic Semi-Structured Memory with Temporal Reasoning for Long-Term Conversational AI
di: Banerjee, Pratyay, et al.
Pubblicazione: (2026)
di: Banerjee, Pratyay, et al.
Pubblicazione: (2026)
Is Architectural Complexity Overrated? Competitive and Interpretable Knowledge Graph Completion with RelatE
di: Chakraborty, Abhijit, et al.
Pubblicazione: (2025)
di: Chakraborty, Abhijit, et al.
Pubblicazione: (2025)
A Pointer Network-based Approach for Joint Extraction and Detection of Multi-Label Multi-Class Intents
di: Mullick, Ankan, et al.
Pubblicazione: (2024)
di: Mullick, Ankan, et al.
Pubblicazione: (2024)
Beyond One-Size-Fits-All: Multi-Domain, Multi-Task Framework for Embedding Model Selection
di: Khetan, Vivek
Pubblicazione: (2024)
di: Khetan, Vivek
Pubblicazione: (2024)
Rank-K: Test-Time Reasoning for Listwise Reranking
di: Yang, Eugene, et al.
Pubblicazione: (2025)
di: Yang, Eugene, et al.
Pubblicazione: (2025)
Autofocus Retrieval: An Effective Pipeline for Multi-Hop Question Answering With Semi-Structured Knowledge
di: Boer, Derian, et al.
Pubblicazione: (2025)
di: Boer, Derian, et al.
Pubblicazione: (2025)
Redefining Retrieval Evaluation in the Era of LLMs
di: Trappolini, Giovanni, et al.
Pubblicazione: (2025)
di: Trappolini, Giovanni, et al.
Pubblicazione: (2025)
DistRAG: Towards Distance-Based Spatial Reasoning in LLMs
di: Schneider, Nicole R, et al.
Pubblicazione: (2025)
di: Schneider, Nicole R, et al.
Pubblicazione: (2025)
Temporal Fact Conflicts in LLMs: Reproducibility Insights from Unifying DYNAMICQA and MULAN
di: Dey, Ritajit, et al.
Pubblicazione: (2026)
di: Dey, Ritajit, et al.
Pubblicazione: (2026)
CiteFix: Enhancing RAG Accuracy Through Post-Processing Citation Correction
di: Maheshwari, Harsh, et al.
Pubblicazione: (2025)
di: Maheshwari, Harsh, et al.
Pubblicazione: (2025)
TQA-Bench: Evaluating LLMs for Multi-Table Question Answering with Scalable Context and Symbolic Extension
di: Qiu, Zipeng, et al.
Pubblicazione: (2024)
di: Qiu, Zipeng, et al.
Pubblicazione: (2024)
FollowTable: A Benchmark for Instruction-Following Table Retrieval
di: Jin, Rihui, et al.
Pubblicazione: (2026)
di: Jin, Rihui, et al.
Pubblicazione: (2026)
Improving Robustness of Tabular Retrieval via Representational Stability
di: Bhandari, Kushal Raj, et al.
Pubblicazione: (2026)
di: Bhandari, Kushal Raj, et al.
Pubblicazione: (2026)
Evaluating LLMs for Gender Disparities in Notable Persons
di: Rhue, Lauren, et al.
Pubblicazione: (2024)
di: Rhue, Lauren, et al.
Pubblicazione: (2024)
Evolving Text Data Stream Mining
di: Kumar, Jay
Pubblicazione: (2024)
di: Kumar, Jay
Pubblicazione: (2024)
TTQA-RS- A break-down prompting approach for Multi-hop Table-Text Question Answering with Reasoning and Summarization
di: Bardhan, Jayetri, et al.
Pubblicazione: (2024)
di: Bardhan, Jayetri, et al.
Pubblicazione: (2024)
Evaluating the Performance of LLMs on Technical Language Processing tasks
di: Kernycky, Andrew, et al.
Pubblicazione: (2024)
di: Kernycky, Andrew, et al.
Pubblicazione: (2024)
A Comprehensive Evaluation of Large Language Models on Temporal Event Forecasting
di: Chang, He, et al.
Pubblicazione: (2024)
di: Chang, He, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Evidence-Guided Schema Normalization for Temporal Tabular Reasoning
di: Thanga, Ashish, et al.
Pubblicazione: (2025) -
NTSEBENCH: Cognitive Reasoning Benchmark for Vision Language Models
di: Pandya, Pranshu, et al.
Pubblicazione: (2024) -
Evaluating Concurrent Robustness of Language Models Across Diverse Challenge Sets
di: Gupta, Vatsal, et al.
Pubblicazione: (2023) -
CORE-T: COherent REtrieval of Tables for Text-to-SQL
di: Soliman, Hassan, et al.
Pubblicazione: (2026) -
FlowVQA: Mapping Multimodal Logic in Visual Question Answering with Flowcharts
di: Singh, Shubhankar, et al.
Pubblicazione: (2024)