An Empirical Investigation of Robustness in Large Language Models under Tabular Distortions
Fuente:
arXiv
Guardado en:
| Autores principales: | Dutta, Avik, Nigam, Harshit, Hasanbeig, Hosein, Radhakrishna, Arjun, Gulwani, Sumit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
ConDABench: Interactive Evaluation of Language Models for Data Analysis
por: Dutta, Avik, et al.
Publicado: (2025)
por: Dutta, Avik, et al.
Publicado: (2025)
Collaboration and Conflict between Humans and Language Models through the Lens of Game Theory
por: Singh, Mukul, et al.
Publicado: (2025)
por: Singh, Mukul, et al.
Publicado: (2025)
Do Code Models Suffer from the Dunning-Kruger Effect?
por: Singh, Mukul, et al.
Publicado: (2025)
por: Singh, Mukul, et al.
Publicado: (2025)
TEN: Table Explicitization, Neurosymbolically
por: Mehrotra, Nikita, et al.
Publicado: (2025)
por: Mehrotra, Nikita, et al.
Publicado: (2025)
MetaReflection: Learning Instructions for Language Agents using Past Reflections
por: Gupta, Priyanshu, et al.
Publicado: (2024)
por: Gupta, Priyanshu, et al.
Publicado: (2024)
Enhancing Creativity in Large Language Models through Associative Thinking Strategies
por: Mehrotra, Pronita, et al.
Publicado: (2024)
por: Mehrotra, Pronita, et al.
Publicado: (2024)
Exploring Interaction Patterns for Debugging: Enhancing Conversational Capabilities of AI-assistants
por: Chopra, Bhavya, et al.
Publicado: (2024)
por: Chopra, Bhavya, et al.
Publicado: (2024)
Scaling Competence, Shrinking Reasoning: Cognitive Signatures in Language Model Learning
por: Singh, Mukul, et al.
Publicado: (2025)
por: Singh, Mukul, et al.
Publicado: (2025)
Soft-Label Training Preserves Epistemic Uncertainty
por: Singh, Agamdeep, et al.
Publicado: (2025)
por: Singh, Agamdeep, et al.
Publicado: (2025)
TableTalk: Scaffolding Spreadsheet Development with a Language Agent
por: Liang, Jenny T., et al.
Publicado: (2025)
por: Liang, Jenny T., et al.
Publicado: (2025)
Training Emergent Joint Associations: A Reinforcement Learning Approach to Creative Thinking in Language Models
por: Singh, Mukul, et al.
Publicado: (2025)
por: Singh, Mukul, et al.
Publicado: (2025)
Mission-driven Exploration for Accelerated Deep Reinforcement Learning with Temporal Logic Task Specifications
por: Wang, Jun, et al.
Publicado: (2023)
por: Wang, Jun, et al.
Publicado: (2023)
STACKFEED: Structured Textual Actor-Critic Knowledge Base Editing with FeedBack
por: Kirtania, Shashank, et al.
Publicado: (2024)
por: Kirtania, Shashank, et al.
Publicado: (2024)
An Empirical Study of Validating Synthetic Data for Formula Generation
por: Singh, Usneek, et al.
Publicado: (2024)
por: Singh, Usneek, et al.
Publicado: (2024)
Large Language Models as Universal Predictors? An Empirical Study on Small Tabular Datasets
por: Pavlidis, Nikolaos, et al.
Publicado: (2025)
por: Pavlidis, Nikolaos, et al.
Publicado: (2025)
Redefining Developer Assistance: Through Large Language Models in Software Ecosystem
por: Banerjee, Somnath, et al.
Publicado: (2023)
por: Banerjee, Somnath, et al.
Publicado: (2023)
Improving Language Agents through BREW
por: Kirtania, Shashank, et al.
Publicado: (2025)
por: Kirtania, Shashank, et al.
Publicado: (2025)
IndiMathBench: Autoformalizing Mathematical Reasoning Problems with a Human Touch
por: Biyani, Param, et al.
Publicado: (2025)
por: Biyani, Param, et al.
Publicado: (2025)
Acceleron: A Tool to Accelerate Research Ideation
por: Nigam, Harshit, et al.
Publicado: (2024)
por: Nigam, Harshit, et al.
Publicado: (2024)
Diffusion is a code repair operator and generator
por: Singh, Mukul, et al.
Publicado: (2025)
por: Singh, Mukul, et al.
Publicado: (2025)
An Empirical Study on Large Language Models in Accuracy and Robustness under Chinese Industrial Scenarios
por: Li, Zongjie, et al.
Publicado: (2024)
por: Li, Zongjie, et al.
Publicado: (2024)
Goal Kernel Planning: Linearly-Solvable Non-Markovian Policies for Logical Tasks with Goal-Conditioned Options
por: Ringstrom, Thomas J., et al.
Publicado: (2020)
por: Ringstrom, Thomas J., et al.
Publicado: (2020)
Investigating Imperceptibility of Adversarial Attacks on Tabular Data: An Empirical Analysis
por: He, Zhipeng, et al.
Publicado: (2024)
por: He, Zhipeng, et al.
Publicado: (2024)
Multi-level Diagnosis and Evaluation for Robust Tabular Feature Engineering with Large Language Models
por: Lim, Yebin, et al.
Publicado: (2025)
por: Lim, Yebin, et al.
Publicado: (2025)
Tabularis Formatus: Predictive Formatting for Tables
por: Singh, Mukul, et al.
Publicado: (2025)
por: Singh, Mukul, et al.
Publicado: (2025)
Towards Understanding the Robustness of LLM-based Evaluations under Perturbations
por: Chaudhary, Manav, et al.
Publicado: (2024)
por: Chaudhary, Manav, et al.
Publicado: (2024)
Investigating the Robustness of Deductive Reasoning with Large Language Models
por: Hoppe, Fabian, et al.
Publicado: (2025)
por: Hoppe, Fabian, et al.
Publicado: (2025)
Investigating Persuasion Techniques in Arabic: An Empirical Study Leveraging Large Language Models
por: Alzahrani, Abdurahmman, et al.
Publicado: (2024)
por: Alzahrani, Abdurahmman, et al.
Publicado: (2024)
Robust Tabular Foundation Models
por: Peroni, Matthew, et al.
Publicado: (2025)
por: Peroni, Matthew, et al.
Publicado: (2025)
Automating Human Tutor-Style Programming Feedback: Leveraging GPT-4 Tutor Model for Hint Generation and GPT-3.5 Student Model for Hint Validation
por: Phung, Tung, et al.
Publicado: (2023)
por: Phung, Tung, et al.
Publicado: (2023)
Generating Realistic Tabular Data with Large Language Models
por: Nguyen, Dang, et al.
Publicado: (2024)
por: Nguyen, Dang, et al.
Publicado: (2024)
Analysis of Error Sources in LLM-based Hypothesis Search for Few-Shot Rule Induction
por: Parab, Aishni, et al.
Publicado: (2025)
por: Parab, Aishni, et al.
Publicado: (2025)
Harnessing the Power of Large Language Models for Empathetic Response Generation: Empirical Investigations and Improvements
por: Qian, Yushan, et al.
Publicado: (2023)
por: Qian, Yushan, et al.
Publicado: (2023)
Conceptual Schema Inference for Tabular Datasets using Large Language Models
por: Wu, Zhenyu, et al.
Publicado: (2025)
por: Wu, Zhenyu, et al.
Publicado: (2025)
Noise Immunity in In-Context Tabular Learning: An Empirical Robustness Analysis of TabPFN's Attention Mechanisms
por: Hu, James, et al.
Publicado: (2026)
por: Hu, James, et al.
Publicado: (2026)
Exploration vs. Fixation: Scaffolding Divergent and Convergent Thinking for Human-AI Co-Creation with Generative Models
por: Wen, Chao, et al.
Publicado: (2025)
por: Wen, Chao, et al.
Publicado: (2025)
Exploring the Robustness of Language Models for Tabular Question Answering via Attention Analysis
por: Bhandari, Kushal Raj, et al.
Publicado: (2024)
por: Bhandari, Kushal Raj, et al.
Publicado: (2024)
Rethinking Distribution Shifts: Empirical Analysis and Inductive Modeling for Tabular Data
por: Wang, Tianyu, et al.
Publicado: (2023)
por: Wang, Tianyu, et al.
Publicado: (2023)
Robust Detection of Synthetic Tabular Data under Schema Variability
por: Kindji, G. Charbel N., et al.
Publicado: (2025)
por: Kindji, G. Charbel N., et al.
Publicado: (2025)
SWEnergy: An Empirical Study on Energy Efficiency in Agentic Issue Resolution Frameworks with SLMs
por: Tripathy, Arihant, et al.
Publicado: (2025)
por: Tripathy, Arihant, et al.
Publicado: (2025)
Ejemplares similares
-
ConDABench: Interactive Evaluation of Language Models for Data Analysis
por: Dutta, Avik, et al.
Publicado: (2025) -
Collaboration and Conflict between Humans and Language Models through the Lens of Game Theory
por: Singh, Mukul, et al.
Publicado: (2025) -
Do Code Models Suffer from the Dunning-Kruger Effect?
por: Singh, Mukul, et al.
Publicado: (2025) -
TEN: Table Explicitization, Neurosymbolically
por: Mehrotra, Nikita, et al.
Publicado: (2025) -
MetaReflection: Learning Instructions for Language Agents using Past Reflections
por: Gupta, Priyanshu, et al.
Publicado: (2024)