Agent Fine-tuning through Distillation for Domain-specific LLMs in Microdomains
Fuente:
arXiv
Guardado en:
| Autores principales: | Xue, Yawen, Tsunokake, Masaya, Koreeda, Yuta, Amin, Ekant Muljibhai, Sumiyoshi, Takashi, Sogawa, Yasuhiro |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Is Micro Domain-Adaptive Pre-Training Effective for Real-World Operations? Multi-Step Evaluation Reveals Potential and Bottlenecks
por: Tsunokake, Masaya, et al.
Publicado: (2026)
por: Tsunokake, Masaya, et al.
Publicado: (2026)
Building Domain-Specific Small Language Models via Guided Data Generation
por: Kumar, Aman, et al.
Publicado: (2025)
por: Kumar, Aman, et al.
Publicado: (2025)
Web Page Classification using LLMs for Crawling Support
por: Sasazawa, Yuichi, et al.
Publicado: (2025)
por: Sasazawa, Yuichi, et al.
Publicado: (2025)
Fine-tuning and Utilization Methods of Domain-specific LLMs
por: Jeong, Cheonsu
Publicado: (2024)
por: Jeong, Cheonsu
Publicado: (2024)
GFlowNet Fine-tuning for Diverse Correct Solutions in Mathematical Reasoning Tasks
por: Takase, Ryoichi, et al.
Publicado: (2024)
por: Takase, Ryoichi, et al.
Publicado: (2024)
Acquiring Bidirectionality via Large and Small Language Models
por: Goto, Takumi, et al.
Publicado: (2024)
por: Goto, Takumi, et al.
Publicado: (2024)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
por: Zhou, Huichi, et al.
Publicado: (2025)
por: Zhou, Huichi, et al.
Publicado: (2025)
KaFT: Knowledge-aware Fine-tuning for Boosting LLMs' Domain-specific Question-Answering Performance
por: Zhong, Qihuang, et al.
Publicado: (2025)
por: Zhong, Qihuang, et al.
Publicado: (2025)
Fine-tuning Large Language Models for Domain-specific Machine Translation
por: Zheng, Jiawei, et al.
Publicado: (2024)
por: Zheng, Jiawei, et al.
Publicado: (2024)
Layout Generation Agents with Large Language Models
por: Sasazawa, Yuichi, et al.
Publicado: (2024)
por: Sasazawa, Yuichi, et al.
Publicado: (2024)
Affordably Fine-tuned LLMs Provide Better Answers to Course-specific MCQs
por: Raimondi, Bianca, et al.
Publicado: (2025)
por: Raimondi, Bianca, et al.
Publicado: (2025)
Breaking Free Transformer Models: Task-specific Context Attribution Promises Improved Generalizability Without Fine-tuning Pre-trained LLMs
por: Tytarenko, Stepan, et al.
Publicado: (2024)
por: Tytarenko, Stepan, et al.
Publicado: (2024)
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
por: Zhang, Zhexin, et al.
Publicado: (2025)
por: Zhang, Zhexin, et al.
Publicado: (2025)
Relational Knowledge Distillation Using Fine-tuned Function Vectors
por: Kang, Andrea, et al.
Publicado: (2026)
por: Kang, Andrea, et al.
Publicado: (2026)
Large Language Model for Multi-Domain Translation: Benchmarking and Domain CoT Fine-tuning
por: Hu, Tianxiang, et al.
Publicado: (2024)
por: Hu, Tianxiang, et al.
Publicado: (2024)
Exploring Intrinsic Language-specific Subspaces in Fine-tuning Multilingual Neural Machine Translation
por: Cao, Zhe, et al.
Publicado: (2024)
por: Cao, Zhe, et al.
Publicado: (2024)
Topic Modeling with Fine-tuning LLMs and Bag of Sentences
por: Schneider, Johannes
Publicado: (2024)
por: Schneider, Johannes
Publicado: (2024)
Beyond Fine-tuning: Unleashing the Potential of Continuous Pretraining for Clinical LLMs
por: Christophe, Clément, et al.
Publicado: (2024)
por: Christophe, Clément, et al.
Publicado: (2024)
Polarization of Autonomous Generative AI Agents Under Echo Chambers
por: Ohagi, Masaya
Publicado: (2024)
por: Ohagi, Masaya
Publicado: (2024)
Dynamic Self-Distillation via Previous Mini-batches for Fine-tuning Small Language Models
por: Fu, Yao, et al.
Publicado: (2024)
por: Fu, Yao, et al.
Publicado: (2024)
Can Pre-training Indicators Reliably Predict Fine-tuning Outcomes of LLMs?
por: Zeng, Hansi, et al.
Publicado: (2025)
por: Zeng, Hansi, et al.
Publicado: (2025)
Aloe: A Family of Fine-tuned Open Healthcare LLMs
por: Gururajan, Ashwin Kumar, et al.
Publicado: (2024)
por: Gururajan, Ashwin Kumar, et al.
Publicado: (2024)
Are LLMs Effective Backbones for Fine-tuning? An Experimental Investigation of Supervised LLMs on Chinese Short Text Matching
por: Liu, Shulin, et al.
Publicado: (2024)
por: Liu, Shulin, et al.
Publicado: (2024)
Question-Answering System for Bangla: Fine-tuning BERT-Bangla for a Closed Domain
por: Roy, Subal Chandra, et al.
Publicado: (2024)
por: Roy, Subal Chandra, et al.
Publicado: (2024)
RLSF: Fine-tuning LLMs via Symbolic Feedback
por: Jha, Piyush, et al.
Publicado: (2024)
por: Jha, Piyush, et al.
Publicado: (2024)
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning
por: Zhong, Qihuang, et al.
Publicado: (2025)
por: Zhong, Qihuang, et al.
Publicado: (2025)
From 'Showgirls' to 'Performers': Fine-tuning with Gender-inclusive Language for Bias Reduction in LLMs
por: Bartl, Marion, et al.
Publicado: (2024)
por: Bartl, Marion, et al.
Publicado: (2024)
Do Domain-specific Experts exist in MoE-based LLMs?
por: Do, Giang, et al.
Publicado: (2026)
por: Do, Giang, et al.
Publicado: (2026)
ITERTL: An Iterative Framework for Fine-tuning LLMs for RTL Code Generation
por: Wu, Peiyang, et al.
Publicado: (2024)
por: Wu, Peiyang, et al.
Publicado: (2024)
Fine-tuning Strategies for Domain Specific Question Answering under Low Annotation Budget Constraints
por: Guo, Kunpeng, et al.
Publicado: (2024)
por: Guo, Kunpeng, et al.
Publicado: (2024)
Dental Severity Assessment through Few-shot Learning and SBERT Fine-tuning
por: Dehghani, Mohammad
Publicado: (2024)
por: Dehghani, Mohammad
Publicado: (2024)
Targeted Efficient Fine-tuning: Optimizing Parameter Updates with Data-Driven Sample Selection
por: Dong, Ming, et al.
Publicado: (2024)
por: Dong, Ming, et al.
Publicado: (2024)
Towards Objective Fine-tuning: How LLMs' Prior Knowledge Causes Potential Poor Calibration?
por: Wang, Ziming, et al.
Publicado: (2025)
por: Wang, Ziming, et al.
Publicado: (2025)
ObfuscaTune: Obfuscated Offsite Fine-tuning and Inference of Proprietary LLMs on Private Datasets
por: Frikha, Ahmed, et al.
Publicado: (2024)
por: Frikha, Ahmed, et al.
Publicado: (2024)
Can We Use Probing to Better Understand Fine-tuning and Knowledge Distillation of the BERT NLU?
por: Hościłowicz, Jakub, et al.
Publicado: (2023)
por: Hościłowicz, Jakub, et al.
Publicado: (2023)
Agent Data Protocol: Unifying Datasets for Diverse, Effective Fine-tuning of LLM Agents
por: Song, Yueqi, et al.
Publicado: (2025)
por: Song, Yueqi, et al.
Publicado: (2025)
Can Visual Encoder Learn to See Arrows?
por: Terashita, Naoyuki, et al.
Publicado: (2025)
por: Terashita, Naoyuki, et al.
Publicado: (2025)
TagRouter: Learning Route to LLMs through Tags for Open-Domain Text Generation Tasks
por: Chen, Zhou, et al.
Publicado: (2025)
por: Chen, Zhou, et al.
Publicado: (2025)
Towards Context-Robust LLMs: A Gated Representation Fine-tuning Approach
por: Zeng, Shenglai, et al.
Publicado: (2025)
por: Zeng, Shenglai, et al.
Publicado: (2025)
Robust and Efficient Fine-tuning of LLMs with Bayesian Reparameterization of Low-Rank Adaptation
por: Sengupta, Ayan, et al.
Publicado: (2024)
por: Sengupta, Ayan, et al.
Publicado: (2024)
Ejemplares similares
-
Is Micro Domain-Adaptive Pre-Training Effective for Real-World Operations? Multi-Step Evaluation Reveals Potential and Bottlenecks
por: Tsunokake, Masaya, et al.
Publicado: (2026) -
Building Domain-Specific Small Language Models via Guided Data Generation
por: Kumar, Aman, et al.
Publicado: (2025) -
Web Page Classification using LLMs for Crawling Support
por: Sasazawa, Yuichi, et al.
Publicado: (2025) -
Fine-tuning and Utilization Methods of Domain-specific LLMs
por: Jeong, Cheonsu
Publicado: (2024) -
GFlowNet Fine-tuning for Diverse Correct Solutions in Mathematical Reasoning Tasks
por: Takase, Ryoichi, et al.
Publicado: (2024)