Balancing Continuous Pre-Training and Instruction Fine-Tuning: Optimizing Instruction-Following in LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Jindal, Ishan, Badrinath, Chandana, Bharti, Pranjal, Vinay, Lakkidi, Sharma, Sachin Dev |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Path of Least Resistance: Guiding LLM Reasoning Trajectories with Prefix Consensus
di: Jindal, Ishan, et al.
Pubblicazione: (2026)
di: Jindal, Ishan, et al.
Pubblicazione: (2026)
Instruction Following without Instruction Tuning
di: Hewitt, John, et al.
Pubblicazione: (2024)
di: Hewitt, John, et al.
Pubblicazione: (2024)
FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale
di: Patel, Ajay, et al.
Pubblicazione: (2026)
di: Patel, Ajay, et al.
Pubblicazione: (2026)
Assessing and Mitigating Data Memorization Risks in Fine-Tuned Large Language Models
di: Ramakrishnan, Badrinath, et al.
Pubblicazione: (2025)
di: Ramakrishnan, Badrinath, et al.
Pubblicazione: (2025)
Balancing Truthfulness and Informativeness with Uncertainty-Aware Instruction Fine-Tuning
di: Wu, Tianyi, et al.
Pubblicazione: (2025)
di: Wu, Tianyi, et al.
Pubblicazione: (2025)
The Instruction Gap: LLMs get lost in Following Instruction
di: Tripathi, Vishesh, et al.
Pubblicazione: (2025)
di: Tripathi, Vishesh, et al.
Pubblicazione: (2025)
From Language Modeling to Instruction Following: Understanding the Behavior Shift in LLMs after Instruction Tuning
di: Wu, Xuansheng, et al.
Pubblicazione: (2023)
di: Wu, Xuansheng, et al.
Pubblicazione: (2023)
Non-instructional Fine-tuning: Enabling Instruction-Following Capabilities in Pre-trained Language Models without Instruction-Following Data
di: Xie, Juncheng, et al.
Pubblicazione: (2024)
di: Xie, Juncheng, et al.
Pubblicazione: (2024)
Enhancing and Assessing Instruction-Following with Fine-Grained Instruction Variants
di: Yang, Jiuding, et al.
Pubblicazione: (2024)
di: Yang, Jiuding, et al.
Pubblicazione: (2024)
Training with Pseudo-Code for Instruction Following
di: Kumar, Prince, et al.
Pubblicazione: (2025)
di: Kumar, Prince, et al.
Pubblicazione: (2025)
UPDESH: Synthesizing Grounded Instruction Tuning Data for 13 Indic Languages
di: Chitale, Pranjal A., et al.
Pubblicazione: (2025)
di: Chitale, Pranjal A., et al.
Pubblicazione: (2025)
Inverse IFEval: Can LLMs Unlearn Stubborn Training Conventions to Follow Real Instructions?
di: Zhang, Qinyan, et al.
Pubblicazione: (2025)
di: Zhang, Qinyan, et al.
Pubblicazione: (2025)
A Comparative Analysis of Instruction Fine-Tuning LLMs for Financial Text Classification
di: Fatemi, Sorouralsadat, et al.
Pubblicazione: (2024)
di: Fatemi, Sorouralsadat, et al.
Pubblicazione: (2024)
Empowering Persian LLMs for Instruction Following: A Novel Dataset and Training Approach
di: Mokhtarabadi, Hojjat, et al.
Pubblicazione: (2024)
di: Mokhtarabadi, Hojjat, et al.
Pubblicazione: (2024)
Fine-Tuning on Noisy Instructions: Effects on Generalization and Performance
di: Alajrami, Ahmed, et al.
Pubblicazione: (2025)
di: Alajrami, Ahmed, et al.
Pubblicazione: (2025)
The Atomic Instruction Gap: Instruction-Tuned LLMs Struggle with Simple, Self-Contained Directives
di: Lim, Henry, et al.
Pubblicazione: (2025)
di: Lim, Henry, et al.
Pubblicazione: (2025)
Reverse Preference Optimization for Complex Instruction Following
di: Huang, Xiang, et al.
Pubblicazione: (2025)
di: Huang, Xiang, et al.
Pubblicazione: (2025)
Mortgage Language Model: Domain-Adaptive Pretraining with Residual Instruction, Alignment Tuning, and Task-Specific Routing
di: Jain, Manish, et al.
Pubblicazione: (2025)
di: Jain, Manish, et al.
Pubblicazione: (2025)
Prompt, Translate, Fine-Tune, Re-Initialize, or Instruction-Tune? Adapting LLMs for In-Context Learning in Low-Resource Languages
di: Toukmaji, Christopher, et al.
Pubblicazione: (2025)
di: Toukmaji, Christopher, et al.
Pubblicazione: (2025)
The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
di: Wallace, Eric, et al.
Pubblicazione: (2024)
di: Wallace, Eric, et al.
Pubblicazione: (2024)
Does Instruction Tuning Make LLMs More Consistent?
di: Fierro, Constanza, et al.
Pubblicazione: (2024)
di: Fierro, Constanza, et al.
Pubblicazione: (2024)
Instruction-Tuning LLMs for Event Extraction with Annotation Guidelines
di: Srivastava, Saurabh, et al.
Pubblicazione: (2025)
di: Srivastava, Saurabh, et al.
Pubblicazione: (2025)
Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning
di: Du, Yanrui, et al.
Pubblicazione: (2024)
di: Du, Yanrui, et al.
Pubblicazione: (2024)
GIFT: Guided Fine-Tuning and Transfer for Enhancing Instruction-Tuned Language Models
di: Ruan, Zhiwen, et al.
Pubblicazione: (2026)
di: Ruan, Zhiwen, et al.
Pubblicazione: (2026)
Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following
di: He, Yun, et al.
Pubblicazione: (2024)
di: He, Yun, et al.
Pubblicazione: (2024)
When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs
di: Li, Xiaomin, et al.
Pubblicazione: (2025)
di: Li, Xiaomin, et al.
Pubblicazione: (2025)
DELIFT: Data Efficient Language model Instruction Fine Tuning
di: Agarwal, Ishika, et al.
Pubblicazione: (2024)
di: Agarwal, Ishika, et al.
Pubblicazione: (2024)
ArgInstruct: Specialized Instruction Fine-Tuning for Computational Argumentation
di: Stahl, Maja, et al.
Pubblicazione: (2025)
di: Stahl, Maja, et al.
Pubblicazione: (2025)
Dynamics of Instruction Fine-Tuning for Chinese Large Language Models
di: Song, Chiyu, et al.
Pubblicazione: (2023)
di: Song, Chiyu, et al.
Pubblicazione: (2023)
Thinking LLMs: General Instruction Following with Thought Generation
di: Wu, Tianhao, et al.
Pubblicazione: (2024)
di: Wu, Tianhao, et al.
Pubblicazione: (2024)
Phased Instruction Fine-Tuning for Large Language Models
di: Pang, Wei, et al.
Pubblicazione: (2024)
di: Pang, Wei, et al.
Pubblicazione: (2024)
SwitchCIT: Switching for Continual Instruction Tuning
di: Wu, Xinbo, et al.
Pubblicazione: (2024)
di: Wu, Xinbo, et al.
Pubblicazione: (2024)
Is In-Context Learning Sufficient for Instruction Following in LLMs?
di: Zhao, Hao, et al.
Pubblicazione: (2024)
di: Zhao, Hao, et al.
Pubblicazione: (2024)
Learning To See But Forgetting To Follow: Visual Instruction Tuning Makes LLMs More Prone To Jailbreak Attacks
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)
di: Pantazopoulos, Georgios, et al.
Pubblicazione: (2024)
Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Approaches
di: Yousefiramandi, Amirhossein, et al.
Pubblicazione: (2025)
di: Yousefiramandi, Amirhossein, et al.
Pubblicazione: (2025)
IOPO: Empowering LLMs with Complex Instruction Following via Input-Output Preference Optimization
di: Zhang, Xinghua, et al.
Pubblicazione: (2024)
di: Zhang, Xinghua, et al.
Pubblicazione: (2024)
Instruction Tuning and CoT Prompting for Contextual Medical QA with LLMs
di: Le, Chenqian, et al.
Pubblicazione: (2025)
di: Le, Chenqian, et al.
Pubblicazione: (2025)
Tool Calling for Arabic LLMs: Data Strategies and Instruction Tuning
di: Ersoy, Asim, et al.
Pubblicazione: (2025)
di: Ersoy, Asim, et al.
Pubblicazione: (2025)
Instruction Pre-Training: Language Models are Supervised Multitask Learners
di: Cheng, Daixuan, et al.
Pubblicazione: (2024)
di: Cheng, Daixuan, et al.
Pubblicazione: (2024)
Instruction Tuning With Loss Over Instructions
di: Shi, Zhengyan, et al.
Pubblicazione: (2024)
di: Shi, Zhengyan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
The Path of Least Resistance: Guiding LLM Reasoning Trajectories with Prefix Consensus
di: Jindal, Ishan, et al.
Pubblicazione: (2026) -
Instruction Following without Instruction Tuning
di: Hewitt, John, et al.
Pubblicazione: (2024) -
FineInstructions: Scaling Synthetic Instructions to Pre-Training Scale
di: Patel, Ajay, et al.
Pubblicazione: (2026) -
Assessing and Mitigating Data Memorization Risks in Fine-Tuned Large Language Models
di: Ramakrishnan, Badrinath, et al.
Pubblicazione: (2025) -
Balancing Truthfulness and Informativeness with Uncertainty-Aware Instruction Fine-Tuning
di: Wu, Tianyi, et al.
Pubblicazione: (2025)