A New Pipeline For Generating Instruction Dataset via RAG and Self Fine-Tuning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Chih-Wei, Lee, Yu-Kai, Tsai, Yin-Te |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hyacinth6B: A large language model for Traditional Chinese
von: Song, Chih-Wei, et al.
Veröffentlicht: (2024)
von: Song, Chih-Wei, et al.
Veröffentlicht: (2024)
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025)
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025)
Phased Instruction Fine-Tuning for Large Language Models
von: Pang, Wei, et al.
Veröffentlicht: (2024)
von: Pang, Wei, et al.
Veröffentlicht: (2024)
CrisisSense-LLM: Instruction Fine-Tuned Large Language Model for Multi-label Social Media Text Classification in Disaster Informatics
von: Yin, Kai, et al.
Veröffentlicht: (2024)
von: Yin, Kai, et al.
Veröffentlicht: (2024)
Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications
von: Sturm, Jakob, et al.
Veröffentlicht: (2026)
von: Sturm, Jakob, et al.
Veröffentlicht: (2026)
Multilevel Analysis of Cryptocurrency News using RAG Approach with Fine-Tuned Mistral Large Language Model
von: Pavlyshenko, Bohdan M.
Veröffentlicht: (2025)
von: Pavlyshenko, Bohdan M.
Veröffentlicht: (2025)
Fine-Tuning MedGemma for Clinical Captioning to Enhance Multimodal RAG over Malaysia CPGs
von: Zun, Lee Qi, et al.
Veröffentlicht: (2025)
von: Zun, Lee Qi, et al.
Veröffentlicht: (2025)
Towards Automatic Continual Learning: A Self-Adaptive Framework for Continual Instruction Tuning
von: Lin, Peiyi, et al.
Veröffentlicht: (2025)
von: Lin, Peiyi, et al.
Veröffentlicht: (2025)
Aya Dataset: An Open-Access Collection for Multilingual Instruction Tuning
von: Singh, Shivalika, et al.
Veröffentlicht: (2024)
von: Singh, Shivalika, et al.
Veröffentlicht: (2024)
Generative Representational Instruction Tuning
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
Generalization-Enhanced Code Vulnerability Detection via Multi-Task Instruction Fine-Tuning
von: Du, Xiaohu, et al.
Veröffentlicht: (2024)
von: Du, Xiaohu, et al.
Veröffentlicht: (2024)
AgentBank: Towards Generalized LLM Agents via Fine-Tuning on 50000+ Interaction Trajectories
von: Song, Yifan, et al.
Veröffentlicht: (2024)
von: Song, Yifan, et al.
Veröffentlicht: (2024)
TACOS: Open Tagging and Comparative Scoring for Instruction Fine-Tuning Data Selection
von: He, Xixiang, et al.
Veröffentlicht: (2025)
von: He, Xixiang, et al.
Veröffentlicht: (2025)
Balancing Truthfulness and Informativeness with Uncertainty-Aware Instruction Fine-Tuning
von: Wu, Tianyi, et al.
Veröffentlicht: (2025)
von: Wu, Tianyi, et al.
Veröffentlicht: (2025)
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG
von: Chen, Xinxi, et al.
Veröffentlicht: (2024)
von: Chen, Xinxi, et al.
Veröffentlicht: (2024)
CF-RAG: A Dataset and Method for Carbon Footprint QA Using Retrieval-Augmented Generation
von: Zhao, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zhao, Kaiwen, et al.
Veröffentlicht: (2025)
LuxInstruct: A Cross-Lingual Instruction Tuning Dataset For Luxembourgish
von: Philippy, Fred, et al.
Veröffentlicht: (2025)
von: Philippy, Fred, et al.
Veröffentlicht: (2025)
Fine-tuning with RAG for Improving LLM Learning of New Skills
von: Ibrahim, Humaid, et al.
Veröffentlicht: (2025)
von: Ibrahim, Humaid, et al.
Veröffentlicht: (2025)
Should We Fine-Tune or RAG? Evaluating Different Techniques to Adapt LLMs for Dialogue
von: Alghisi, Simone, et al.
Veröffentlicht: (2024)
von: Alghisi, Simone, et al.
Veröffentlicht: (2024)
Team Trifecta at Factify5WQA: Setting the Standard in Fact Verification with Fine-Tuning
von: Chiang, Shang-Hsuan, et al.
Veröffentlicht: (2024)
von: Chiang, Shang-Hsuan, et al.
Veröffentlicht: (2024)
A Comparative Analysis of Instruction Fine-Tuning LLMs for Financial Text Classification
von: Fatemi, Sorouralsadat, et al.
Veröffentlicht: (2024)
von: Fatemi, Sorouralsadat, et al.
Veröffentlicht: (2024)
GraphGPT: Graph Instruction Tuning for Large Language Models
von: Tang, Jiabin, et al.
Veröffentlicht: (2023)
von: Tang, Jiabin, et al.
Veröffentlicht: (2023)
Capability Instruction Tuning: A New Paradigm for Dynamic LLM Routing
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2025)
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2025)
Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines
von: Song, Hwanjun
Veröffentlicht: (2026)
von: Song, Hwanjun
Veröffentlicht: (2026)
SIFT-50M: A Large-Scale Multilingual Dataset for Speech Instruction Fine-Tuning
von: Pandey, Prabhat, et al.
Veröffentlicht: (2025)
von: Pandey, Prabhat, et al.
Veröffentlicht: (2025)
Retrieval-Augmented Fine-Tuning With Preference Optimization For Visual Program Generation
von: Kang, Deokhyung, et al.
Veröffentlicht: (2025)
von: Kang, Deokhyung, et al.
Veröffentlicht: (2025)
PAFT: Prompt-Agnostic Fine-Tuning
von: Wei, Chenxing, et al.
Veröffentlicht: (2025)
von: Wei, Chenxing, et al.
Veröffentlicht: (2025)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2024)
von: Köksal, Abdullatif, et al.
Veröffentlicht: (2024)
A Preliminary Study of RAG for Taiwanese Historical Archives
von: Lin, Claire, et al.
Veröffentlicht: (2025)
von: Lin, Claire, et al.
Veröffentlicht: (2025)
Contrastive Instruction Tuning
von: Yan, Tianyi Lorena, et al.
Veröffentlicht: (2024)
von: Yan, Tianyi Lorena, et al.
Veröffentlicht: (2024)
HIRAG: Hierarchical-Thought Instruction-Tuning Retrieval-Augmented Generation
von: Jiao, YiHan, et al.
Veröffentlicht: (2025)
von: Jiao, YiHan, et al.
Veröffentlicht: (2025)
DrugRAG: Enhancing Pharmacy LLM Performance Through A Novel Retrieval-Augmented Generation Pipeline
von: Kazemzadeh, Houman, et al.
Veröffentlicht: (2025)
von: Kazemzadeh, Houman, et al.
Veröffentlicht: (2025)
EvalYaks: Instruction Tuning Datasets and LoRA Fine-tuned Models for Automated Scoring of CEFR B2 Speaking Assessment Transcripts
von: Scaria, Nicy, et al.
Veröffentlicht: (2024)
von: Scaria, Nicy, et al.
Veröffentlicht: (2024)
Controllable Text Generation in the Instruction-Tuning Era
von: Ashok, Dhananjay, et al.
Veröffentlicht: (2024)
von: Ashok, Dhananjay, et al.
Veröffentlicht: (2024)
Safeguarding RAG Pipelines with GMTP: A Gradient-based Masked Token Probability Method for Poisoned Document Detection
von: Kim, San, et al.
Veröffentlicht: (2025)
von: Kim, San, et al.
Veröffentlicht: (2025)
Instruction Fine-Tuning: Does Prompt Loss Matter?
von: Huerta-Enochian, Mathew, et al.
Veröffentlicht: (2024)
von: Huerta-Enochian, Mathew, et al.
Veröffentlicht: (2024)
Routing-Aligned Fine-Tuning for Multilingual Downstream Tasks in Mixture-of-Experts Models
von: Deng, Guanzhi, et al.
Veröffentlicht: (2026)
von: Deng, Guanzhi, et al.
Veröffentlicht: (2026)
Unveiling the Impact of Coding Data Instruction Fine-Tuning on Large Language Models Reasoning
von: Zhang, Xinlu, et al.
Veröffentlicht: (2024)
von: Zhang, Xinlu, et al.
Veröffentlicht: (2024)
Scaling Data Diversity for Fine-Tuning Language Models in Human Alignment
von: Song, Feifan, et al.
Veröffentlicht: (2024)
von: Song, Feifan, et al.
Veröffentlicht: (2024)
DeFine: A Decomposed and Fine-Grained Annotated Dataset for Long-form Article Generation
von: Wang, Ming, et al.
Veröffentlicht: (2025)
von: Wang, Ming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Hyacinth6B: A large language model for Traditional Chinese
von: Song, Chih-Wei, et al.
Veröffentlicht: (2024) -
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
von: Chih, Yu Cheng, et al.
Veröffentlicht: (2025) -
Phased Instruction Fine-Tuning for Large Language Models
von: Pang, Wei, et al.
Veröffentlicht: (2024) -
CrisisSense-LLM: Instruction Fine-Tuned Large Language Model for Multi-label Social Media Text Classification in Disaster Informatics
von: Yin, Kai, et al.
Veröffentlicht: (2024) -
Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications
von: Sturm, Jakob, et al.
Veröffentlicht: (2026)