A New Pipeline For Generating Instruction Dataset via RAG and Self Fine-Tuning
Fuente:
arXiv
Guardado en:
| Autores principales: | Song, Chih-Wei, Lee, Yu-Kai, Tsai, Yin-Te |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hyacinth6B: A large language model for Traditional Chinese
por: Song, Chih-Wei, et al.
Publicado: (2024)
por: Song, Chih-Wei, et al.
Publicado: (2024)
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
por: Chih, Yu Cheng, et al.
Publicado: (2025)
por: Chih, Yu Cheng, et al.
Publicado: (2025)
Phased Instruction Fine-Tuning for Large Language Models
por: Pang, Wei, et al.
Publicado: (2024)
por: Pang, Wei, et al.
Publicado: (2024)
CrisisSense-LLM: Instruction Fine-Tuned Large Language Model for Multi-label Social Media Text Classification in Disaster Informatics
por: Yin, Kai, et al.
Publicado: (2024)
por: Yin, Kai, et al.
Publicado: (2024)
Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications
por: Sturm, Jakob, et al.
Publicado: (2026)
por: Sturm, Jakob, et al.
Publicado: (2026)
Multilevel Analysis of Cryptocurrency News using RAG Approach with Fine-Tuned Mistral Large Language Model
por: Pavlyshenko, Bohdan M.
Publicado: (2025)
por: Pavlyshenko, Bohdan M.
Publicado: (2025)
Fine-Tuning MedGemma for Clinical Captioning to Enhance Multimodal RAG over Malaysia CPGs
por: Zun, Lee Qi, et al.
Publicado: (2025)
por: Zun, Lee Qi, et al.
Publicado: (2025)
Towards Automatic Continual Learning: A Self-Adaptive Framework for Continual Instruction Tuning
por: Lin, Peiyi, et al.
Publicado: (2025)
por: Lin, Peiyi, et al.
Publicado: (2025)
Aya Dataset: An Open-Access Collection for Multilingual Instruction Tuning
por: Singh, Shivalika, et al.
Publicado: (2024)
por: Singh, Shivalika, et al.
Publicado: (2024)
Generative Representational Instruction Tuning
por: Muennighoff, Niklas, et al.
Publicado: (2024)
por: Muennighoff, Niklas, et al.
Publicado: (2024)
Generalization-Enhanced Code Vulnerability Detection via Multi-Task Instruction Fine-Tuning
por: Du, Xiaohu, et al.
Publicado: (2024)
por: Du, Xiaohu, et al.
Publicado: (2024)
AgentBank: Towards Generalized LLM Agents via Fine-Tuning on 50000+ Interaction Trajectories
por: Song, Yifan, et al.
Publicado: (2024)
por: Song, Yifan, et al.
Publicado: (2024)
TACOS: Open Tagging and Comparative Scoring for Instruction Fine-Tuning Data Selection
por: He, Xixiang, et al.
Publicado: (2025)
por: He, Xixiang, et al.
Publicado: (2025)
Balancing Truthfulness and Informativeness with Uncertainty-Aware Instruction Fine-Tuning
por: Wu, Tianyi, et al.
Publicado: (2025)
por: Wu, Tianyi, et al.
Publicado: (2025)
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG
por: Chen, Xinxi, et al.
Publicado: (2024)
por: Chen, Xinxi, et al.
Publicado: (2024)
CF-RAG: A Dataset and Method for Carbon Footprint QA Using Retrieval-Augmented Generation
por: Zhao, Kaiwen, et al.
Publicado: (2025)
por: Zhao, Kaiwen, et al.
Publicado: (2025)
LuxInstruct: A Cross-Lingual Instruction Tuning Dataset For Luxembourgish
por: Philippy, Fred, et al.
Publicado: (2025)
por: Philippy, Fred, et al.
Publicado: (2025)
Fine-tuning with RAG for Improving LLM Learning of New Skills
por: Ibrahim, Humaid, et al.
Publicado: (2025)
por: Ibrahim, Humaid, et al.
Publicado: (2025)
Should We Fine-Tune or RAG? Evaluating Different Techniques to Adapt LLMs for Dialogue
por: Alghisi, Simone, et al.
Publicado: (2024)
por: Alghisi, Simone, et al.
Publicado: (2024)
Team Trifecta at Factify5WQA: Setting the Standard in Fact Verification with Fine-Tuning
por: Chiang, Shang-Hsuan, et al.
Publicado: (2024)
por: Chiang, Shang-Hsuan, et al.
Publicado: (2024)
A Comparative Analysis of Instruction Fine-Tuning LLMs for Financial Text Classification
por: Fatemi, Sorouralsadat, et al.
Publicado: (2024)
por: Fatemi, Sorouralsadat, et al.
Publicado: (2024)
GraphGPT: Graph Instruction Tuning for Large Language Models
por: Tang, Jiabin, et al.
Publicado: (2023)
por: Tang, Jiabin, et al.
Publicado: (2023)
Capability Instruction Tuning: A New Paradigm for Dynamic LLM Routing
por: Zhang, Yi-Kai, et al.
Publicado: (2025)
por: Zhang, Yi-Kai, et al.
Publicado: (2025)
Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines
por: Song, Hwanjun
Publicado: (2026)
por: Song, Hwanjun
Publicado: (2026)
SIFT-50M: A Large-Scale Multilingual Dataset for Speech Instruction Fine-Tuning
por: Pandey, Prabhat, et al.
Publicado: (2025)
por: Pandey, Prabhat, et al.
Publicado: (2025)
Retrieval-Augmented Fine-Tuning With Preference Optimization For Visual Program Generation
por: Kang, Deokhyung, et al.
Publicado: (2025)
por: Kang, Deokhyung, et al.
Publicado: (2025)
PAFT: Prompt-Agnostic Fine-Tuning
por: Wei, Chenxing, et al.
Publicado: (2025)
por: Wei, Chenxing, et al.
Publicado: (2025)
MURI: High-Quality Instruction Tuning Datasets for Low-Resource Languages via Reverse Instructions
por: Köksal, Abdullatif, et al.
Publicado: (2024)
por: Köksal, Abdullatif, et al.
Publicado: (2024)
A Preliminary Study of RAG for Taiwanese Historical Archives
por: Lin, Claire, et al.
Publicado: (2025)
por: Lin, Claire, et al.
Publicado: (2025)
Contrastive Instruction Tuning
por: Yan, Tianyi Lorena, et al.
Publicado: (2024)
por: Yan, Tianyi Lorena, et al.
Publicado: (2024)
HIRAG: Hierarchical-Thought Instruction-Tuning Retrieval-Augmented Generation
por: Jiao, YiHan, et al.
Publicado: (2025)
por: Jiao, YiHan, et al.
Publicado: (2025)
DrugRAG: Enhancing Pharmacy LLM Performance Through A Novel Retrieval-Augmented Generation Pipeline
por: Kazemzadeh, Houman, et al.
Publicado: (2025)
por: Kazemzadeh, Houman, et al.
Publicado: (2025)
EvalYaks: Instruction Tuning Datasets and LoRA Fine-tuned Models for Automated Scoring of CEFR B2 Speaking Assessment Transcripts
por: Scaria, Nicy, et al.
Publicado: (2024)
por: Scaria, Nicy, et al.
Publicado: (2024)
Controllable Text Generation in the Instruction-Tuning Era
por: Ashok, Dhananjay, et al.
Publicado: (2024)
por: Ashok, Dhananjay, et al.
Publicado: (2024)
Safeguarding RAG Pipelines with GMTP: A Gradient-based Masked Token Probability Method for Poisoned Document Detection
por: Kim, San, et al.
Publicado: (2025)
por: Kim, San, et al.
Publicado: (2025)
Instruction Fine-Tuning: Does Prompt Loss Matter?
por: Huerta-Enochian, Mathew, et al.
Publicado: (2024)
por: Huerta-Enochian, Mathew, et al.
Publicado: (2024)
Routing-Aligned Fine-Tuning for Multilingual Downstream Tasks in Mixture-of-Experts Models
por: Deng, Guanzhi, et al.
Publicado: (2026)
por: Deng, Guanzhi, et al.
Publicado: (2026)
Unveiling the Impact of Coding Data Instruction Fine-Tuning on Large Language Models Reasoning
por: Zhang, Xinlu, et al.
Publicado: (2024)
por: Zhang, Xinlu, et al.
Publicado: (2024)
Scaling Data Diversity for Fine-Tuning Language Models in Human Alignment
por: Song, Feifan, et al.
Publicado: (2024)
por: Song, Feifan, et al.
Publicado: (2024)
DeFine: A Decomposed and Fine-Grained Annotated Dataset for Long-form Article Generation
por: Wang, Ming, et al.
Publicado: (2025)
por: Wang, Ming, et al.
Publicado: (2025)
Ejemplares similares
-
Hyacinth6B: A large language model for Traditional Chinese
por: Song, Chih-Wei, et al.
Publicado: (2024) -
Low-Resource Fine-Tuning for Multi-Task Structured Information Extraction with a Billion-Parameter Instruction-Tuned Model
por: Chih, Yu Cheng, et al.
Publicado: (2025) -
Phased Instruction Fine-Tuning for Large Language Models
por: Pang, Wei, et al.
Publicado: (2024) -
CrisisSense-LLM: Instruction Fine-Tuned Large Language Model for Multi-label Social Media Text Classification in Disaster Informatics
por: Yin, Kai, et al.
Publicado: (2024) -
Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications
por: Sturm, Jakob, et al.
Publicado: (2026)