Investigating the Impact of Data Contamination of Large Language Models in Text-to-SQL Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Ranaldi, Federico, Ruzzetti, Elena Sofia, Onorati, Dario, Ranaldi, Leonardo, Giannone, Cristina, Favalli, Andrea, Romagnoli, Raniero, Zanzotto, Fabio Massimo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MeMo: Towards Language Models with Associative Memory Mechanisms
by: Zanzotto, Fabio Massimo, et al.
Published: (2025)
by: Zanzotto, Fabio Massimo, et al.
Published: (2025)
A Trip Towards Fairness: Bias and De-Biasing in Large Language Models
by: Ranaldi, Leonardo, et al.
Published: (2023)
by: Ranaldi, Leonardo, et al.
Published: (2023)
PreCog: Exploring the Relation between Memorization and Performance in Pre-trained Language Models
by: Ranaldi, Leonardo, et al.
Published: (2023)
by: Ranaldi, Leonardo, et al.
Published: (2023)
Empowering Multi-step Reasoning across Languages via Tree-of-Thoughts
by: Ranaldi, Leonardo, et al.
Published: (2023)
by: Ranaldi, Leonardo, et al.
Published: (2023)
Enhancing Data Privacy in Large Language Models through Private Association Editing
by: Venditti, Davide, et al.
Published: (2024)
by: Venditti, Davide, et al.
Published: (2024)
Exploring Linguistic Properties of Monolingual BERTs with Typological Classification among Languages
by: Ruzzetti, Elena Sofia, et al.
Published: (2023)
by: Ruzzetti, Elena Sofia, et al.
Published: (2023)
Protoknowledge Shapes Behaviour of LLMs in Downstream Tasks: Memorization and Generalization with Knowledge Graphs
by: Ranaldi, Federico, et al.
Published: (2025)
by: Ranaldi, Federico, et al.
Published: (2025)
The Dark Side of the Language: Pre-trained Transformers in the DarkNet
by: Ranaldi, Leonardo, et al.
Published: (2022)
by: Ranaldi, Leonardo, et al.
Published: (2022)
HANS, are you clever? Clever Hans Effect Analysis of Neural Systems
by: Ranaldi, Leonardo, et al.
Published: (2023)
by: Ranaldi, Leonardo, et al.
Published: (2023)
Animate, or Inanimate, That is the Question for Large Language Models
by: Ranaldi, Leonardo, et al.
Published: (2024)
by: Ranaldi, Leonardo, et al.
Published: (2024)
Improving Multilingual Retrieval-Augmented Language Models through Dialectic Reasoning Argumentations
by: Ranaldi, Leonardo, et al.
Published: (2025)
by: Ranaldi, Leonardo, et al.
Published: (2025)
Every time I fire a conversational designer, the performance of the dialog system goes down
by: Xompero, Giancarlo A., et al.
Published: (2021)
by: Xompero, Giancarlo A., et al.
Published: (2021)
Neuro-symbolic Syntactic Parsing: Shaping a Neural Network with the CYK Algorithm
by: Zanzotto, Fabio Massimo, et al.
Published: (2026)
by: Zanzotto, Fabio Massimo, et al.
Published: (2026)
Abstract Activation Spaces for Content-Invariant Reasoning in Large Language Models
by: Maraia, Gabriele, et al.
Published: (2026)
by: Maraia, Gabriele, et al.
Published: (2026)
Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models
by: Ruzzetti, Elena Sofia, et al.
Published: (2025)
by: Ruzzetti, Elena Sofia, et al.
Published: (2025)
When Large Language Models contradict humans? Large Language Models' Sycophantic Behaviour
by: Ranaldi, Leonardo, et al.
Published: (2023)
by: Ranaldi, Leonardo, et al.
Published: (2023)
Empowering Cross-lingual Abilities of Instruction-tuned Large Language Models by Translation-following demonstrations
by: Ranaldi, Leonardo, et al.
Published: (2023)
by: Ranaldi, Leonardo, et al.
Published: (2023)
Preserving Privacy in Large Language Models: A Survey on Current Threats and Solutions
by: Miranda, Michele, et al.
Published: (2024)
by: Miranda, Michele, et al.
Published: (2024)
Limitarianism: The Case Against Extreme Wealth, by IngridRobeyns (Allen Lane (Penguin Books), First published in Great Britain, 2024, 336.
by: Marco Ranaldi
Published: (2024)
by: Marco Ranaldi
Published: (2024)
Self-Refine Instruction-Tuning for Aligning Reasoning in Language Models
by: Ranaldi, Leonardo, et al.
Published: (2024)
by: Ranaldi, Leonardo, et al.
Published: (2024)
Eliciting Critical Reasoning in Retrieval-Augmented Language Models via Contrastive Explanations
by: Ranaldi, Leonardo, et al.
Published: (2024)
by: Ranaldi, Leonardo, et al.
Published: (2024)
Improving Chain-of-Thought Reasoning via Quasi-Symbolic Abstractions
by: Ranaldi, Leonardo, et al.
Published: (2025)
by: Ranaldi, Leonardo, et al.
Published: (2025)
Multilingual Retrieval-Augmented Generation for Knowledge-Intensive Task
by: Ranaldi, Leonardo, et al.
Published: (2025)
by: Ranaldi, Leonardo, et al.
Published: (2025)
Less is KEN: a Universal and Simple Non-Parametric Pruning Algorithm for Large Language Models
by: Mastromattei, Michele, et al.
Published: (2024)
by: Mastromattei, Michele, et al.
Published: (2024)
Dissecting Clinical Reasoning in Language Models: A Comparative Study of Prompts and Model Adaptation Strategies
by: Jullien, Mael, et al.
Published: (2025)
by: Jullien, Mael, et al.
Published: (2025)
Linguistic Fingerprint in Transformer Models: How Language Variation Influences Parameter Selection in Irony Detection
by: Mastromattei, Michele, et al.
Published: (2024)
by: Mastromattei, Michele, et al.
Published: (2024)
WhoFi: Deep Person Re-Identification via Wi-Fi Channel Signal Encoding
by: Avola, Danilo, et al.
Published: (2025)
by: Avola, Danilo, et al.
Published: (2025)
An Investigation of Ear-EEG Signals for a Novel Biometric Authentication System
by: Avola, Danilo, et al.
Published: (2025)
by: Avola, Danilo, et al.
Published: (2025)
Bi-GRU Based Deception Detection using EEG Signals
by: Avola, Danilo, et al.
Published: (2025)
by: Avola, Danilo, et al.
Published: (2025)
SING-SQL: A Synthetic Data Generation Framework for In-Domain Text-to-SQL Translation
by: Caferoğlu, Hasan Alp, et al.
Published: (2025)
by: Caferoğlu, Hasan Alp, et al.
Published: (2025)
Digital Shielding for Cross-Domain Wi-Fi Signal Adaptation using Relativistic Average Generative Adversarial Network
by: Avola, Danilo, et al.
Published: (2025)
by: Avola, Danilo, et al.
Published: (2025)
Transformer-Based Person Identification via Wi-Fi CSI Amplitude and Phase Perturbations
by: Avola, Danilo, et al.
Published: (2025)
by: Avola, Danilo, et al.
Published: (2025)
EPI-SQL: Enhancing Text-to-SQL Translation with Error-Prevention Instructions
by: Liu, Xiping, et al.
Published: (2024)
by: Liu, Xiping, et al.
Published: (2024)
Benchmarking of EEG Analysis Techniques for Parkinson's Disease Diagnosis: A Comparison between Traditional ML Methods and Foundation DL Methods
by: Avola, Danilo, et al.
Published: (2025)
by: Avola, Danilo, et al.
Published: (2025)
Padrão e ritmo de aquisição das habilidades motoras de lactentes pré-termo nos quatro primeiros meses de idade corrigida
by: Elaine P. Raniero
Published: (2010)
by: Elaine P. Raniero
Published: (2010)
GBV-SQL: Guided Generation and SQL2Text Back-Translation Validation for Multi-Agent Text2SQL
by: Chen, Daojun, et al.
Published: (2025)
by: Chen, Daojun, et al.
Published: (2025)
Boundaries and Limits of the Structural Chaos of Prime Numbers
by: Romagnoli, Federico
Published: (2026)
by: Romagnoli, Federico
Published: (2026)
Fabbisogni professionali e formativi nei settori moda, metalmeccanico ed agro-alimentare della provincia di Chieti
by: Romagnoli, Federico
Published: (2007)
by: Romagnoli, Federico
Published: (2007)
Indagine sulle competenze dei lavoratori della provincia di Teramo
by: Romagnoli, Federico
Published: (2006)
by: Romagnoli, Federico
Published: (2006)
Relativistic limits on the discretization and temporal resolution of a quantum clock
by: Favalli, Tommaso
Published: (2025)
by: Favalli, Tommaso
Published: (2025)
Similar Items
-
MeMo: Towards Language Models with Associative Memory Mechanisms
by: Zanzotto, Fabio Massimo, et al.
Published: (2025) -
A Trip Towards Fairness: Bias and De-Biasing in Large Language Models
by: Ranaldi, Leonardo, et al.
Published: (2023) -
PreCog: Exploring the Relation between Memorization and Performance in Pre-trained Language Models
by: Ranaldi, Leonardo, et al.
Published: (2023) -
Empowering Multi-step Reasoning across Languages via Tree-of-Thoughts
by: Ranaldi, Leonardo, et al.
Published: (2023) -
Enhancing Data Privacy in Large Language Models through Private Association Editing
by: Venditti, Davide, et al.
Published: (2024)