Plug-in and Fine-tuning: Bridging the Gap between Small Language Models and Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Kyeonghyun, Jang, Jinhee, Choi, Juhwan, Lee, Yoonji, Jin, Kyohoon, Kim, YoungBin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GPTs Are Multilingual Annotators for Sequence Generation Tasks
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
Multi-News+: Cost-efficient Dataset Cleansing via LLM-based Data Annotation
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
SUMMPILOT: Bridging Efficiency and Customization for Interactive Summarization System
von: Yun, JungMin, et al.
Veröffentlicht: (2026)
von: Yun, JungMin, et al.
Veröffentlicht: (2026)
Adverb Is the Key: Simple Text Data Augmentation with Adverb Deletion
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
Beyond Single-User Dialogue: Assessing Multi-User Dialogue State Tracking Capabilities of Large Language Models
von: Song, Sangmin, et al.
Veröffentlicht: (2025)
von: Song, Sangmin, et al.
Veröffentlicht: (2025)
Colorful Cutout: Enhancing Image Data Augmentation with Curriculum Learning
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
Medal Matters: Probing LLMs' Failure Cases Through Olympic Rankings
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
VolDoGer: LLM-assisted Datasets for Domain Generalization in Vision-Language Tasks
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
CoBA: Counterbias Text Augmentation for Mitigating Various Spurious Correlations via Semantic Triples
von: Jin, Kyohoon, et al.
Veröffentlicht: (2025)
von: Jin, Kyohoon, et al.
Veröffentlicht: (2025)
UniGen: Universal Domain Generalization for Sentiment Classification via Zero-shot Dataset Generation
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
SoftEDA: Rethinking Rule-Based Data Augmentation with Soft Labels
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
AutoAugment Is What You Need: Enhancing Rule-based Augmentation Methods in Low-resource Regimes
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
von: Choi, Juhwan, et al.
Veröffentlicht: (2024)
Strategic Data Ordering: Enhancing Large Language Model Performance through Curriculum Learning
von: Kim, Jisu, et al.
Veröffentlicht: (2024)
von: Kim, Jisu, et al.
Veröffentlicht: (2024)
Enhancing Effectiveness and Robustness in a Low-Resource Regime via Decision-Boundary-aware Data Augmentation
von: Jin, Kyohoon, et al.
Veröffentlicht: (2024)
von: Jin, Kyohoon, et al.
Veröffentlicht: (2024)
Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models
von: Kim, Mingyeong, et al.
Veröffentlicht: (2026)
von: Kim, Mingyeong, et al.
Veröffentlicht: (2026)
LLM Agents at the Roundtable: A Multi-Perspective and Dialectical Reasoning Framework for Essay Scoring
von: Jang, Jinhee, et al.
Veröffentlicht: (2025)
von: Jang, Jinhee, et al.
Veröffentlicht: (2025)
VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
Model Fusion through Bayesian Optimization in Language Model Fine-Tuning
von: Jang, Chaeyun, et al.
Veröffentlicht: (2024)
von: Jang, Chaeyun, et al.
Veröffentlicht: (2024)
Aligning with Your Own Voice: Self-Corrected Preference Learning for Hallucination Mitigation in LVLMs
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
Developing a Pragmatic Benchmark for Assessing Korean Legal Language Understanding in Large Language Models
von: Kim, Yeeun, et al.
Veröffentlicht: (2024)
von: Kim, Yeeun, et al.
Veröffentlicht: (2024)
Sensitivity of Small Language Models to Fine-tuning Data Contamination
von: Scaria, Nicy, et al.
Veröffentlicht: (2025)
von: Scaria, Nicy, et al.
Veröffentlicht: (2025)
Delving into Multilingual Ethical Bias: The MSQAD with Statistical Hypothesis Tests for Large Language Models
von: Yu, Seunguk, et al.
Veröffentlicht: (2025)
von: Yu, Seunguk, et al.
Veröffentlicht: (2025)
FairQE: Multi-Agent Framework for Mitigating Gender Bias in Translation Quality Estimation
von: Jang, Jinhee, et al.
Veröffentlicht: (2026)
von: Jang, Jinhee, et al.
Veröffentlicht: (2026)
Small Language Models Fine-tuned to Coordinate Larger Language Models improve Complex Reasoning
von: Juneja, Gurusha, et al.
Veröffentlicht: (2023)
von: Juneja, Gurusha, et al.
Veröffentlicht: (2023)
Semi-supervised Fine-tuning for Large Language Models
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
PEMA: An Offsite-Tunable Plug-in External Memory Adaptation for Language Models
von: Kim, HyunJin, et al.
Veröffentlicht: (2023)
von: Kim, HyunJin, et al.
Veröffentlicht: (2023)
Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models
von: Kwon, JuneHyoung, et al.
Veröffentlicht: (2026)
von: Kwon, JuneHyoung, et al.
Veröffentlicht: (2026)
Fine-tuning Small Language Models as Efficient Enterprise Search Relevance Labelers
von: Kang, Yue, et al.
Veröffentlicht: (2026)
von: Kang, Yue, et al.
Veröffentlicht: (2026)
Beyond Learning: A Training-Free Alternative to Model Adaptation
von: Yoon, Namkyung, et al.
Veröffentlicht: (2026)
von: Yoon, Namkyung, et al.
Veröffentlicht: (2026)
Federated Co-tuning Framework for Large and Small Language Models
von: Fan, Tao, et al.
Veröffentlicht: (2024)
von: Fan, Tao, et al.
Veröffentlicht: (2024)
Control Token with Dense Passage Retrieval
von: Lee, Juhwan, et al.
Veröffentlicht: (2024)
von: Lee, Juhwan, et al.
Veröffentlicht: (2024)
Bridging the Gap between Expert and Language Models: Concept-guided Chess Commentary Generation and Evaluation
von: Kim, Jaechang, et al.
Veröffentlicht: (2024)
von: Kim, Jaechang, et al.
Veröffentlicht: (2024)
Eliciting Instruction-tuned Code Language Models' Capabilities to Utilize Auxiliary Function for Code Generation
von: Lee, Seonghyeon, et al.
Veröffentlicht: (2024)
von: Lee, Seonghyeon, et al.
Veröffentlicht: (2024)
SmallToLarge (S2L): Scalable Data Selection for Fine-tuning Large Language Models by Summarizing Training Trajectories of Small Models
von: Yang, Yu, et al.
Veröffentlicht: (2024)
von: Yang, Yu, et al.
Veröffentlicht: (2024)
Information Guided Regularization for Fine-tuning Language Models
von: Sharma, Mandar, et al.
Veröffentlicht: (2024)
von: Sharma, Mandar, et al.
Veröffentlicht: (2024)
Prompting Strategies for Language Model-Based Item Generation in K-12 Education: Bridging the Gap Between Small and Large Language Models
von: Amini, Mohammad, et al.
Veröffentlicht: (2025)
von: Amini, Mohammad, et al.
Veröffentlicht: (2025)
How Well Do Large Language Models Truly Ground?
von: Lee, Hyunji, et al.
Veröffentlicht: (2023)
von: Lee, Hyunji, et al.
Veröffentlicht: (2023)
Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models
von: Kang, Jaehoon, et al.
Veröffentlicht: (2026)
von: Kang, Jaehoon, et al.
Veröffentlicht: (2026)
KnowTuning: Knowledge-aware Fine-tuning for Large Language Models
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
JudgeLM: Fine-tuned Large Language Models are Scalable Judges
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
GPTs Are Multilingual Annotators for Sequence Generation Tasks
von: Choi, Juhwan, et al.
Veröffentlicht: (2024) -
Multi-News+: Cost-efficient Dataset Cleansing via LLM-based Data Annotation
von: Choi, Juhwan, et al.
Veröffentlicht: (2024) -
SUMMPILOT: Bridging Efficiency and Customization for Interactive Summarization System
von: Yun, JungMin, et al.
Veröffentlicht: (2026) -
Adverb Is the Key: Simple Text Data Augmentation with Adverb Deletion
von: Choi, Juhwan, et al.
Veröffentlicht: (2024) -
Beyond Single-User Dialogue: Assessing Multi-User Dialogue State Tracking Capabilities of Large Language Models
von: Song, Sangmin, et al.
Veröffentlicht: (2025)