Auto prompt sql: a resource-efficient architecture for text-to-sql translation in constrained environments
Fuente:
arXiv
Salvato in:
| Autori principali: | Tang, Zetong, Ma, Qian, Wu, Di |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
EvidenceMap: Learning Evidence Analysis to Unleash the Power of Small Language Models for Biomedical Question Answering
di: Zong, Chang, et al.
Pubblicazione: (2025)
di: Zong, Chang, et al.
Pubblicazione: (2025)
Mediator: Memory-efficient LLM Merging with Less Parameter Conflicts and Uncertainty Based Routing
di: Lai, Kunfeng, et al.
Pubblicazione: (2025)
di: Lai, Kunfeng, et al.
Pubblicazione: (2025)
$\text{Memory}^3$: Language Modeling with Explicit Memory
di: Yang, Hongkang, et al.
Pubblicazione: (2024)
di: Yang, Hongkang, et al.
Pubblicazione: (2024)
Pay Attention to What You Need
di: Gao, Yifei, et al.
Pubblicazione: (2023)
di: Gao, Yifei, et al.
Pubblicazione: (2023)
DYNAMAX: Dynamic computing for Transformers and Mamba based architectures
di: Nogales, Miguel, et al.
Pubblicazione: (2025)
di: Nogales, Miguel, et al.
Pubblicazione: (2025)
OPENXRD: A Comprehensive Benchmark Framework for LLM/MLLM XRD Question Answering
di: Vosoughi, Ali, et al.
Pubblicazione: (2025)
di: Vosoughi, Ali, et al.
Pubblicazione: (2025)
HInter: Exposing Hidden Intersectional Bias in Large Language Models
di: Souani, Badr, et al.
Pubblicazione: (2025)
di: Souani, Badr, et al.
Pubblicazione: (2025)
Meaning-infused grammar: Gradient Acceptability Shapes the Geometric Representations of Constructions in LLMs
di: Rakshit, Supantho, et al.
Pubblicazione: (2025)
di: Rakshit, Supantho, et al.
Pubblicazione: (2025)
T-VEC: A Telecom-Specific Vectorization Model with Enhanced Semantic Understanding via Deep Triplet Loss Fine-Tuning
di: Ethiraj, Vignesh, et al.
Pubblicazione: (2025)
di: Ethiraj, Vignesh, et al.
Pubblicazione: (2025)
One SPACE to Rule Them All: Jointly Mitigating Factuality and Faithfulness Hallucinations in LLMs
di: Wang, Pengbo, et al.
Pubblicazione: (2025)
di: Wang, Pengbo, et al.
Pubblicazione: (2025)
Japanese Tort-case Dataset for Rationale-supported Legal Judgment Prediction
di: Yamada, Hiroaki, et al.
Pubblicazione: (2023)
di: Yamada, Hiroaki, et al.
Pubblicazione: (2023)
Crossing Linguistic Horizons: Finetuning and Comprehensive Evaluation of Vietnamese Large Language Models
di: Truong, Sang T., et al.
Pubblicazione: (2024)
di: Truong, Sang T., et al.
Pubblicazione: (2024)
Curriculum Recommendations Using Transformer Base Model with InfoNCE Loss And Language Switching Method
di: Xu, Xiaonan, et al.
Pubblicazione: (2024)
di: Xu, Xiaonan, et al.
Pubblicazione: (2024)
A Primer on Large Language Models and their Limitations
di: Johnson, Sandra, et al.
Pubblicazione: (2024)
di: Johnson, Sandra, et al.
Pubblicazione: (2024)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
di: Otal, Hakan T., et al.
Pubblicazione: (2024)
di: Otal, Hakan T., et al.
Pubblicazione: (2024)
Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring
di: Aksoy, Sinan G., et al.
Pubblicazione: (2026)
di: Aksoy, Sinan G., et al.
Pubblicazione: (2026)
Forging GEMs: Advancing Greek NLP through Quality-Based Corpus Curation
di: Apostolopoulou, Alexandra, et al.
Pubblicazione: (2025)
di: Apostolopoulou, Alexandra, et al.
Pubblicazione: (2025)
LLMs as Deceptive Agents: How Role-Based Prompting Induces Semantic Ambiguity in Puzzle Tasks
di: Yoo, Seunghyun
Pubblicazione: (2025)
di: Yoo, Seunghyun
Pubblicazione: (2025)
A Language Model-Driven Semi-Supervised Ensemble Framework for Illicit Market Detection Across Deep/Dark Web and Social Platforms
di: Yazdanjue, Navid, et al.
Pubblicazione: (2025)
di: Yazdanjue, Navid, et al.
Pubblicazione: (2025)
Evaluating Named Entity Recognition: A comparative analysis of mono- and multilingual transformer models on a novel Brazilian corporate earnings call transcripts dataset
di: Abilio, Ramon, et al.
Pubblicazione: (2024)
di: Abilio, Ramon, et al.
Pubblicazione: (2024)
An Iterative Optimizing Framework for Radiology Report Summarization with ChatGPT
di: Ma, Chong, et al.
Pubblicazione: (2023)
di: Ma, Chong, et al.
Pubblicazione: (2023)
Do LLMs have a Gender (Entropy) Bias?
di: Prabhune, Sonal, et al.
Pubblicazione: (2025)
di: Prabhune, Sonal, et al.
Pubblicazione: (2025)
LayerTracer: A Joint Task-Particle and Vulnerable-Layer Analysis framework for Arbitrary Large Language Model Architectures
di: Wu, Yuhang, et al.
Pubblicazione: (2026)
di: Wu, Yuhang, et al.
Pubblicazione: (2026)
Measuring Faithfulness and Abstention: An Automated Pipeline for Evaluating LLM-Generated 3-ply Case-Based Legal Arguments
di: Zhang, Li, et al.
Pubblicazione: (2025)
di: Zhang, Li, et al.
Pubblicazione: (2025)
The Illusion of Role Separation: Hidden Shortcuts in LLM Role Learning (and How to Fix Them)
di: Wang, Zihao, et al.
Pubblicazione: (2025)
di: Wang, Zihao, et al.
Pubblicazione: (2025)
SVDq: 1.25-bit and 410x Key Cache Compression for LLM Attention
di: Yankun, Hong, et al.
Pubblicazione: (2025)
di: Yankun, Hong, et al.
Pubblicazione: (2025)
Data filtering methods for training language models
di: Shevchenko, Egor, et al.
Pubblicazione: (2026)
di: Shevchenko, Egor, et al.
Pubblicazione: (2026)
Pretraining and Updates of Domain-Specific LLM: A Case Study in the Japanese Business Domain
di: Takahashi, Kosuke, et al.
Pubblicazione: (2024)
di: Takahashi, Kosuke, et al.
Pubblicazione: (2024)
LEGAL-UQA: A Low-Resource Urdu-English Dataset for Legal Question Answering
di: Faisal, Faizan, et al.
Pubblicazione: (2024)
di: Faisal, Faizan, et al.
Pubblicazione: (2024)
Transforming and Combining Rewards for Aligning Large Language Models
di: Wang, Zihao, et al.
Pubblicazione: (2024)
di: Wang, Zihao, et al.
Pubblicazione: (2024)
Can Out-of-Distribution Evaluations Uncover Reliance on Shortcuts? A Case Study in Question Answering
di: Štefánik, Michal, et al.
Pubblicazione: (2025)
di: Štefánik, Michal, et al.
Pubblicazione: (2025)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
di: Wang, Youkang, et al.
Pubblicazione: (2025)
di: Wang, Youkang, et al.
Pubblicazione: (2025)
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models
di: Safavi-Naini, Seyed Amir Ahmad, et al.
Pubblicazione: (2024)
di: Safavi-Naini, Seyed Amir Ahmad, et al.
Pubblicazione: (2024)
Empirical analysis of binding precedent efficiency in Brazilian Supreme Court via case classification
di: Tinarrage, Raphaël, et al.
Pubblicazione: (2024)
di: Tinarrage, Raphaël, et al.
Pubblicazione: (2024)
Clinical information extraction for Low-resource languages with Few-shot learning using Pre-trained language models and Prompting
di: Richter-Pechanski, Phillip, et al.
Pubblicazione: (2024)
di: Richter-Pechanski, Phillip, et al.
Pubblicazione: (2024)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
di: Miliani, Martina, et al.
Pubblicazione: (2025)
di: Miliani, Martina, et al.
Pubblicazione: (2025)
Context Aware Lemmatization and Morphological Tagging Method in Turkish
di: Sayallar, Cagri
Pubblicazione: (2025)
di: Sayallar, Cagri
Pubblicazione: (2025)
MetaCheckGPT -- A Multi-task Hallucination Detector Using LLM Uncertainty and Meta-models
di: Mehta, Rahul, et al.
Pubblicazione: (2024)
di: Mehta, Rahul, et al.
Pubblicazione: (2024)
Prompting Encoder Models for Zero-Shot Classification: A Cross-Domain Study in Italian
di: Auriemma, Serena, et al.
Pubblicazione: (2024)
di: Auriemma, Serena, et al.
Pubblicazione: (2024)
Raw Text is All you Need: Knowledge-intensive Multi-turn Instruction Tuning for Large Language Model
di: Hou, Xia, et al.
Pubblicazione: (2024)
di: Hou, Xia, et al.
Pubblicazione: (2024)
Documenti analoghi
-
EvidenceMap: Learning Evidence Analysis to Unleash the Power of Small Language Models for Biomedical Question Answering
di: Zong, Chang, et al.
Pubblicazione: (2025) -
Mediator: Memory-efficient LLM Merging with Less Parameter Conflicts and Uncertainty Based Routing
di: Lai, Kunfeng, et al.
Pubblicazione: (2025) -
$\text{Memory}^3$: Language Modeling with Explicit Memory
di: Yang, Hongkang, et al.
Pubblicazione: (2024) -
Pay Attention to What You Need
di: Gao, Yifei, et al.
Pubblicazione: (2023) -
DYNAMAX: Dynamic computing for Transformers and Mamba based architectures
di: Nogales, Miguel, et al.
Pubblicazione: (2025)