Impact of Model Size on Fine-tuned LLM Performance in Data-to-Text Generation: A State-of-the-Art Investigation
Fuente:
arXiv
Salvato in:
| Autori principali: | Mahapatra, Joy, Garain, Utpal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An Extensive Evaluation of Factual Consistency in Large Language Models for Data-to-Text Generation
di: Mahapatra, Joy, et al.
Pubblicazione: (2024)
di: Mahapatra, Joy, et al.
Pubblicazione: (2024)
Factual Inconsistency in Data-to-Text Generation Scales Exponentially with LLM Size: A Statistical Validation
di: Mahapatra, Joy, et al.
Pubblicazione: (2025)
di: Mahapatra, Joy, et al.
Pubblicazione: (2025)
sudoLLM: On Multi-role Alignment of Language Models
di: Saha, Soumadeep, et al.
Pubblicazione: (2025)
di: Saha, Soumadeep, et al.
Pubblicazione: (2025)
Language Models are Crossword Solvers
di: Saha, Soumadeep, et al.
Pubblicazione: (2024)
di: Saha, Soumadeep, et al.
Pubblicazione: (2024)
Data-efficient LLM Fine-tuning for Code Generation
di: Lv, Weijie, et al.
Pubblicazione: (2025)
di: Lv, Weijie, et al.
Pubblicazione: (2025)
mdok of KInIT: Robustly Fine-tuned LLM for Binary and Multiclass AI-Generated Text Detection
di: Macko, Dominik
Pubblicazione: (2025)
di: Macko, Dominik
Pubblicazione: (2025)
Stage-wise Fine-tuning for Graph-to-Text Generation
di: Wang, Qingyun, et al.
Pubblicazione: (2021)
di: Wang, Qingyun, et al.
Pubblicazione: (2021)
Investigating the Representation of Backchannels and Fillers in Fine-tuned Language Models
di: Wang, Yu, et al.
Pubblicazione: (2025)
di: Wang, Yu, et al.
Pubblicazione: (2025)
KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?
di: Saha, Soumadeep, et al.
Pubblicazione: (2025)
di: Saha, Soumadeep, et al.
Pubblicazione: (2025)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
Souper-Model: How Simple Arithmetic Unlocks State-of-the-Art LLM Performance
di: Maiti, Shalini, et al.
Pubblicazione: (2025)
di: Maiti, Shalini, et al.
Pubblicazione: (2025)
Building a Family of Data Augmentation Models for Low-cost LLM Fine-tuning on the Cloud
di: Yue, Yuanhao, et al.
Pubblicazione: (2024)
di: Yue, Yuanhao, et al.
Pubblicazione: (2024)
On the Effects of Fine-tuning Language Models for Text-Based Reinforcement Learning
di: Gruppi, Mauricio, et al.
Pubblicazione: (2024)
di: Gruppi, Mauricio, et al.
Pubblicazione: (2024)
Are LLMs Effective Backbones for Fine-tuning? An Experimental Investigation of Supervised LLMs on Chinese Short Text Matching
di: Liu, Shulin, et al.
Pubblicazione: (2024)
di: Liu, Shulin, et al.
Pubblicazione: (2024)
Region Mixup
di: Saha, Saptarshi, et al.
Pubblicazione: (2024)
di: Saha, Saptarshi, et al.
Pubblicazione: (2024)
A General Framework to Enhance Fine-tuning-based LLM Unlearning
di: Ren, Jie, et al.
Pubblicazione: (2025)
di: Ren, Jie, et al.
Pubblicazione: (2025)
Can Perplexity Predict Fine-tuning Performance? An Investigation of Tokenization Effects on Sequential Language Models for Nepali
di: Luitel, Nishant, et al.
Pubblicazione: (2024)
di: Luitel, Nishant, et al.
Pubblicazione: (2024)
Increasing the Robustness of the Fine-tuned Multilingual Machine-Generated Text Detectors
di: Macko, Dominik, et al.
Pubblicazione: (2025)
di: Macko, Dominik, et al.
Pubblicazione: (2025)
Improving Text Embeddings for Smaller Language Models Using Contrastive Fine-tuning
di: Ukarapol, Trapoom, et al.
Pubblicazione: (2024)
di: Ukarapol, Trapoom, et al.
Pubblicazione: (2024)
State of the Art in Text Classification for South Slavic Languages: Fine-Tuning or Prompting?
di: Pungeršek, Taja Kuzman, et al.
Pubblicazione: (2025)
di: Pungeršek, Taja Kuzman, et al.
Pubblicazione: (2025)
Fine-tuning Large Language Models for Multigenerator, Multidomain, and Multilingual Machine-Generated Text Detection
di: Xiong, Feng, et al.
Pubblicazione: (2024)
di: Xiong, Feng, et al.
Pubblicazione: (2024)
LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning
di: Huang, Wei, et al.
Pubblicazione: (2026)
di: Huang, Wei, et al.
Pubblicazione: (2026)
LoFiT: Localized Fine-tuning on LLM Representations
di: Yin, Fangcong, et al.
Pubblicazione: (2024)
di: Yin, Fangcong, et al.
Pubblicazione: (2024)
An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model is not a General Substitute for GPT-4
di: Huang, Hui, et al.
Pubblicazione: (2024)
di: Huang, Hui, et al.
Pubblicazione: (2024)
Token-level Data Selection for Safe LLM Fine-tuning
di: Li, Yanping, et al.
Pubblicazione: (2026)
di: Li, Yanping, et al.
Pubblicazione: (2026)
Two-stage LLM Fine-tuning with Less Specialization and More Generalization
di: Wang, Yihan, et al.
Pubblicazione: (2022)
di: Wang, Yihan, et al.
Pubblicazione: (2022)
The Harder The Better: Maintaining Supervised Fine-tuning Generalization with Less but Harder Data
di: Shang, Zhaoyang, et al.
Pubblicazione: (2025)
di: Shang, Zhaoyang, et al.
Pubblicazione: (2025)
AI Generated Text Detection Using Instruction Fine-tuned Large Language and Transformer-Based Models
di: Guggilla, Chinnappa, et al.
Pubblicazione: (2025)
di: Guggilla, Chinnappa, et al.
Pubblicazione: (2025)
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
di: Zhang, Zhexin, et al.
Pubblicazione: (2025)
di: Zhang, Zhexin, et al.
Pubblicazione: (2025)
FourierKAN outperforms MLP on Text Classification Head Fine-tuning
di: Imran, Abdullah Al, et al.
Pubblicazione: (2024)
di: Imran, Abdullah Al, et al.
Pubblicazione: (2024)
Intention-Adaptive LLM Fine-Tuning for Text Revision Generation
di: Liu, Zhexiong, et al.
Pubblicazione: (2026)
di: Liu, Zhexiong, et al.
Pubblicazione: (2026)
MEDVOC: Vocabulary Adaptation for Fine-tuning Pre-trained Language Models on Medical Text Summarization
di: Balde, Gunjan, et al.
Pubblicazione: (2024)
di: Balde, Gunjan, et al.
Pubblicazione: (2024)
How Green are Neural Language Models? Analyzing Energy Consumption in Text Summarization Fine-tuning
di: Rehman, Tohida, et al.
Pubblicazione: (2025)
di: Rehman, Tohida, et al.
Pubblicazione: (2025)
Evaluating The Impact of Stimulus Quality in Investigations of LLM Language Performance
di: Pistotti, Timothy, et al.
Pubblicazione: (2025)
di: Pistotti, Timothy, et al.
Pubblicazione: (2025)
Sensitivity of Small Language Models to Fine-tuning Data Contamination
di: Scaria, Nicy, et al.
Pubblicazione: (2025)
di: Scaria, Nicy, et al.
Pubblicazione: (2025)
Agent Data Protocol: Unifying Datasets for Diverse, Effective Fine-tuning of LLM Agents
di: Song, Yueqi, et al.
Pubblicazione: (2025)
di: Song, Yueqi, et al.
Pubblicazione: (2025)
What Makes and Breaks Safety Fine-tuning? A Mechanistic Study
di: Jain, Samyak, et al.
Pubblicazione: (2024)
di: Jain, Samyak, et al.
Pubblicazione: (2024)
Investigating the Impact of Data Selection Strategies on Language Model Performance
di: Gu, Jiayao, et al.
Pubblicazione: (2025)
di: Gu, Jiayao, et al.
Pubblicazione: (2025)
Fine-tuning Done Right in Model Editing
di: Yang, Wanli, et al.
Pubblicazione: (2025)
di: Yang, Wanli, et al.
Pubblicazione: (2025)
Table-LLM-Specialist: Language Model Specialists for Tables using Iterative Generator-Validator Fine-tuning
di: Xing, Junjie, et al.
Pubblicazione: (2024)
di: Xing, Junjie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
An Extensive Evaluation of Factual Consistency in Large Language Models for Data-to-Text Generation
di: Mahapatra, Joy, et al.
Pubblicazione: (2024) -
Factual Inconsistency in Data-to-Text Generation Scales Exponentially with LLM Size: A Statistical Validation
di: Mahapatra, Joy, et al.
Pubblicazione: (2025) -
sudoLLM: On Multi-role Alignment of Language Models
di: Saha, Soumadeep, et al.
Pubblicazione: (2025) -
Language Models are Crossword Solvers
di: Saha, Soumadeep, et al.
Pubblicazione: (2024) -
Data-efficient LLM Fine-tuning for Code Generation
di: Lv, Weijie, et al.
Pubblicazione: (2025)