TabXEval: Why this is a Bad Table? An eXhaustive Rubric for Table Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Pancholi, Vihang, Bafna, Jainit, Anvekar, Tejas, Shrivastava, Manish, Gupta, Vivek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TabReX : Tabular Referenceless eXplainable Evaluation
by: Anvekar, Tejas, et al.
Published: (2025)
by: Anvekar, Tejas, et al.
Published: (2025)
Knowledge-Aware Reasoning over Multimodal Semi-structured Tables
by: Mathur, Suyash Vardhan, et al.
Published: (2024)
by: Mathur, Suyash Vardhan, et al.
Published: (2024)
Mast Kalandar at SemEval-2024 Task 8: On the Trail of Textual Origins: RoBERTa-BiLSTM Approach to Detect AI-Generated Text
by: Bafna, Jainit Sushil, et al.
Published: (2024)
by: Bafna, Jainit Sushil, et al.
Published: (2024)
ViTaB-A: Evaluating Multimodal Large Language Models on Visual Table Attribution
by: Alqurnawi, Yahia, et al.
Published: (2026)
by: Alqurnawi, Yahia, et al.
Published: (2026)
TraceBack: Multi-Agent Decomposition for Fine-Grained Table Attribution
by: Anvekar, Tejas, et al.
Published: (2026)
by: Anvekar, Tejas, et al.
Published: (2026)
Rethinking Information Synthesis in Multimodal Question Answering A Multi-Agent Perspective
by: Rajput, Krishna Singh, et al.
Published: (2025)
by: Rajput, Krishna Singh, et al.
Published: (2025)
Is Architectural Complexity Overrated? Competitive and Interpretable Knowledge Graph Completion with RelatE
by: Chakraborty, Abhijit, et al.
Published: (2025)
by: Chakraborty, Abhijit, et al.
Published: (2025)
DoPE: Decoy Oriented Perturbation Encapsulation Human-Readable, AI-Hostile Documents for Academic Integrity
by: Shekhar, Ashish Raj, et al.
Published: (2026)
by: Shekhar, Ashish Raj, et al.
Published: (2026)
Integrity Shield A System for Ethical AI Use & Authorship Transparency in Assessments
by: Shekhar, Ashish Raj, et al.
Published: (2026)
by: Shekhar, Ashish Raj, et al.
Published: (2026)
Can LLMs Compute with Reasons?
by: Sandilya, Harshit, et al.
Published: (2024)
by: Sandilya, Harshit, et al.
Published: (2024)
Is This a Bad Table? A Closer Look at the Evaluation of Table Generation from Text
by: Ramu, Pritika, et al.
Published: (2024)
by: Ramu, Pritika, et al.
Published: (2024)
ChiKhaPo: A Large-Scale Multilingual Benchmark for Evaluating Lexical Comprehension and Generation in Large Language Models
by: Chang, Emily, et al.
Published: (2025)
by: Chang, Emily, et al.
Published: (2025)
Segmentation Beyond Defaults: Asymmetrical Byte Pair Encoding for Optimal Machine Translation Performance
by: Yadav, Saumitra, et al.
Published: (2025)
by: Yadav, Saumitra, et al.
Published: (2025)
Can Constructions "SCAN" Compositionality ?
by: Katrapati, Ganesh, et al.
Published: (2025)
by: Katrapati, Ganesh, et al.
Published: (2025)
Get away with less: Need of source side data curation to build parallel corpus for low resource Machine Translation
by: Yadav, Saumitra, et al.
Published: (2026)
by: Yadav, Saumitra, et al.
Published: (2026)
Clash of the models: Comparing performance of BERT-based variants for generic news frame detection
by: Jumle, Vihang
Published: (2026)
by: Jumle, Vihang
Published: (2026)
SCOPE:Planning for Hybrid Querying over Clinical Trial Data
by: Chowdhury, Suparno Roy, et al.
Published: (2026)
by: Chowdhury, Suparno Roy, et al.
Published: (2026)
FD-NL2SQL: Feedback-Driven Clinical NL2SQL that Improves with Use
by: Chowdhury, Suparno Roy, et al.
Published: (2026)
by: Chowdhury, Suparno Roy, et al.
Published: (2026)
TransientTables: Evaluating LLMs' Reasoning on Temporally Evolving Semi-structured Tables
by: Shankarampeta, Abhilash, et al.
Published: (2025)
by: Shankarampeta, Abhilash, et al.
Published: (2025)
The Perceptual Observatory Characterizing Robustness and Grounding in MLLMs
by: Anvekar, Tejas, et al.
Published: (2025)
by: Anvekar, Tejas, et al.
Published: (2025)
Evaluating Large Language Models along Dimensions of Language Variation: A Systematik Invesdigatiom uv Cross-lingual Generalization
by: Bafna, Niyati, et al.
Published: (2024)
by: Bafna, Niyati, et al.
Published: (2024)
FreshTab: Sourcing Fresh Data for Table-to-Text Generation Evaluation
by: Onderková, Kristýna, et al.
Published: (2025)
by: Onderková, Kristýna, et al.
Published: (2025)
MARCUS: An Event-Centric NLP Pipeline that generates Character Arcs from Narratives
by: Bhyravajjula, Sriharsh, et al.
Published: (2025)
by: Bhyravajjula, Sriharsh, et al.
Published: (2025)
Automatic Normalization of Word Variations in Code-Mixed Social Media Text
by: Singh, Rajat, et al.
Published: (2018)
by: Singh, Rajat, et al.
Published: (2018)
Map&Make: Schema Guided Text to Table Generation
by: Ahuja, Naman, et al.
Published: (2025)
by: Ahuja, Naman, et al.
Published: (2025)
Mahalanobis k-NN: A Statistical Lens for Robust Point-Cloud Registrations
by: Anvekar, Tejas, et al.
Published: (2024)
by: Anvekar, Tejas, et al.
Published: (2024)
Leveraging LLM For Synchronizing Information Across Multilingual Tables
by: Khincha, Siddharth, et al.
Published: (2025)
by: Khincha, Siddharth, et al.
Published: (2025)
UNJOIN: Enhancing Multi-Table Text-to-SQL Generation via Schema Simplification
by: Ganesan, Poojah, et al.
Published: (2025)
by: Ganesan, Poojah, et al.
Published: (2025)
Preference-Aware Rubric Learning for Personalized Evaluation
by: Qiu, Yilun, et al.
Published: (2026)
by: Qiu, Yilun, et al.
Published: (2026)
Sentiment Analysis of Code-Mixed Languages leveraging Resource Rich Languages
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Zero-Shot Multi-task Hallucination Detection
by: Bhamidipati, Patanjali, et al.
Published: (2024)
by: Bhamidipati, Patanjali, et al.
Published: (2024)
TeClass: A Human-Annotated Relevance-based Headline Classification and Generation Dataset for Telugu
by: Kanumolu, Gopichand, et al.
Published: (2024)
by: Kanumolu, Gopichand, et al.
Published: (2024)
Emotions are Universal: Learning Sentiment Based Representations of Resource-Poor Languages using Siamese Networks
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Neural Network Architecture for Credibility Assessment of Textual Claims
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
Contrastive Learning of Emoji-based Representations for Resource-Poor Languages
by: Choudhary, Nurendra, et al.
Published: (2018)
by: Choudhary, Nurendra, et al.
Published: (2018)
AdaRubric: Task-Adaptive Rubrics for Reliable LLM Agent Evaluation and Reward Learning
by: Ding, Liang
Published: (2026)
by: Ding, Liang
Published: (2026)
Pointer-Generator Networks for Low-Resource Machine Translation: Don't Copy That!
by: Bafna, Niyati, et al.
Published: (2024)
by: Bafna, Niyati, et al.
Published: (2024)
The Good, The Bad, and Why: Unveiling Emotions in Generative AI
by: Li, Cheng, et al.
Published: (2023)
by: Li, Cheng, et al.
Published: (2023)
CORE-T: COherent REtrieval of Tables for Text-to-SQL
by: Soliman, Hassan, et al.
Published: (2026)
by: Soliman, Hassan, et al.
Published: (2026)
LID Models are Actually Accent Classifiers: Implications and Solutions for LID on Accented Speech
by: Bafna, Niyati, et al.
Published: (2025)
by: Bafna, Niyati, et al.
Published: (2025)
Similar Items
-
TabReX : Tabular Referenceless eXplainable Evaluation
by: Anvekar, Tejas, et al.
Published: (2025) -
Knowledge-Aware Reasoning over Multimodal Semi-structured Tables
by: Mathur, Suyash Vardhan, et al.
Published: (2024) -
Mast Kalandar at SemEval-2024 Task 8: On the Trail of Textual Origins: RoBERTa-BiLSTM Approach to Detect AI-Generated Text
by: Bafna, Jainit Sushil, et al.
Published: (2024) -
ViTaB-A: Evaluating Multimodal Large Language Models on Visual Table Attribution
by: Alqurnawi, Yahia, et al.
Published: (2026) -
TraceBack: Multi-Agent Decomposition for Fine-Grained Table Attribution
by: Anvekar, Tejas, et al.
Published: (2026)