Refine Large Language Model Fine-tuning via Instruction Vector
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Gangwei, Li, Zhaoyi, Lian, Defu, Wei, Ying |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unlocking the Power of Function Vectors for Characterizing and Mitigating Catastrophic Forgetting in Continual Instruction Tuning
von: Jiang, Gangwei, et al.
Veröffentlicht: (2025)
von: Jiang, Gangwei, et al.
Veröffentlicht: (2025)
What Makes a Good Reasoning Chain? Uncovering Structural Patterns in Long Chain-of-Thought Reasoning
von: Jiang, Gangwei, et al.
Veröffentlicht: (2025)
von: Jiang, Gangwei, et al.
Veröffentlicht: (2025)
Mitigate Negative Transfer with Similarity Heuristic Lifelong Prompt Tuning
von: Wu, Chenyuan, et al.
Veröffentlicht: (2024)
von: Wu, Chenyuan, et al.
Veröffentlicht: (2024)
Scaling Reasoning Hop Exposes Weaknesses: Demystifying and Improving Hop Generalization in Large Language Models
von: Li, Zhaoyi, et al.
Veröffentlicht: (2026)
von: Li, Zhaoyi, et al.
Veröffentlicht: (2026)
InstructAV: Instruction Fine-tuning Large Language Models for Authorship Verification
von: Hu, Yujia, et al.
Veröffentlicht: (2024)
von: Hu, Yujia, et al.
Veröffentlicht: (2024)
Demystifying Instruction Mixing for Fine-tuning Large Language Models
von: Wang, Renxi, et al.
Veröffentlicht: (2023)
von: Wang, Renxi, et al.
Veröffentlicht: (2023)
Internal Value Alignment in Large Language Models through Controlled Value Vector Activation
von: Jin, Haoran, et al.
Veröffentlicht: (2025)
von: Jin, Haoran, et al.
Veröffentlicht: (2025)
Model Specific Task Similarity for Vision Language Model Selection via Layer Conductance
von: Yang, Wei, et al.
Veröffentlicht: (2026)
von: Yang, Wei, et al.
Veröffentlicht: (2026)
Semi-supervised Fine-tuning for Large Language Models
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
When Large Language Models Meet Personalization: Perspectives of Challenges and Opportunities
von: Chen, Jin, et al.
Veröffentlicht: (2023)
von: Chen, Jin, et al.
Veröffentlicht: (2023)
Towards Green AI in Fine-tuning Large Language Models via Adaptive Backpropagation
von: Huang, Kai, et al.
Veröffentlicht: (2023)
von: Huang, Kai, et al.
Veröffentlicht: (2023)
RecExplainer: Aligning Large Language Models for Explaining Recommendation Models
von: Lei, Yuxuan, et al.
Veröffentlicht: (2023)
von: Lei, Yuxuan, et al.
Veröffentlicht: (2023)
Invariant Representation via Decoupling Style and Spurious Features from Images
von: Li, Ruimeng, et al.
Veröffentlicht: (2023)
von: Li, Ruimeng, et al.
Veröffentlicht: (2023)
Privacy-preserving Fine-tuning of Large Language Models through Flatness
von: Chen, Tiejin, et al.
Veröffentlicht: (2024)
von: Chen, Tiejin, et al.
Veröffentlicht: (2024)
Panacea: Mitigating Harmful Fine-tuning for Large Language Models via Post-fine-tuning Perturbation
von: Wang, Yibo, et al.
Veröffentlicht: (2025)
von: Wang, Yibo, et al.
Veröffentlicht: (2025)
Demystifying Design Choices of Reinforcement Fine-tuning: A Batched Contextual Bandit Learning Perspective
von: Xie, Hong, et al.
Veröffentlicht: (2026)
von: Xie, Hong, et al.
Veröffentlicht: (2026)
Recommender AI Agent: Integrating Large Language Models for Interactive Recommendations
von: Huang, Xu, et al.
Veröffentlicht: (2023)
von: Huang, Xu, et al.
Veröffentlicht: (2023)
Phased Instruction Fine-Tuning for Large Language Models
von: Pang, Wei, et al.
Veröffentlicht: (2024)
von: Pang, Wei, et al.
Veröffentlicht: (2024)
AI Generated Text Detection Using Instruction Fine-tuned Large Language and Transformer-Based Models
von: Guggilla, Chinnappa, et al.
Veröffentlicht: (2025)
von: Guggilla, Chinnappa, et al.
Veröffentlicht: (2025)
Unlocking Large Language Model's Planning Capabilities with Maximum Diversity Fine-tuning
von: Li, Wenjun, et al.
Veröffentlicht: (2024)
von: Li, Wenjun, et al.
Veröffentlicht: (2024)
Speechworthy Instruction-tuned Language Models
von: Cho, Hyundong, et al.
Veröffentlicht: (2024)
von: Cho, Hyundong, et al.
Veröffentlicht: (2024)
Difference Vector Equalization for Robust Fine-tuning of Vision-Language Models
von: Suzuki, Satoshi, et al.
Veröffentlicht: (2025)
von: Suzuki, Satoshi, et al.
Veröffentlicht: (2025)
Understanding and Patching Compositional Reasoning in LLMs
von: Li, Zhaoyi, et al.
Veröffentlicht: (2024)
von: Li, Zhaoyi, et al.
Veröffentlicht: (2024)
Learning to Substitute Components for Compositional Generalization
von: Li, Zhaoyi, et al.
Veröffentlicht: (2025)
von: Li, Zhaoyi, et al.
Veröffentlicht: (2025)
You Only Fine-tune Once: Many-Shot In-Context Fine-Tuning for Large Language Models
von: He, Wenchong, et al.
Veröffentlicht: (2025)
von: He, Wenchong, et al.
Veröffentlicht: (2025)
SecureReviewer: Enhancing Large Language Models for Secure Code Review through Secure-aware Fine-tuning
von: Liu, Fang, et al.
Veröffentlicht: (2025)
von: Liu, Fang, et al.
Veröffentlicht: (2025)
AlpaCare:Instruction-tuned Large Language Models for Medical Application
von: Zhang, Xinlu, et al.
Veröffentlicht: (2023)
von: Zhang, Xinlu, et al.
Veröffentlicht: (2023)
CommonIT: Commonality-Aware Instruction Tuning for Large Language Models via Data Partitions
von: Rao, Jun, et al.
Veröffentlicht: (2024)
von: Rao, Jun, et al.
Veröffentlicht: (2024)
A Study of Backdoors in Instruction Fine-tuned Language Models
von: Raghuram, Jayaram, et al.
Veröffentlicht: (2024)
von: Raghuram, Jayaram, et al.
Veröffentlicht: (2024)
Selecting Large Language Model to Fine-tune via Rectified Scaling Law
von: Lin, Haowei, et al.
Veröffentlicht: (2024)
von: Lin, Haowei, et al.
Veröffentlicht: (2024)
Your Instructions Are Not Always Helpful: Assessing the Efficacy of Instruction Fine-tuning for Software Vulnerability Detection
von: Yusuf, Imam Nur Bani, et al.
Veröffentlicht: (2024)
von: Yusuf, Imam Nur Bani, et al.
Veröffentlicht: (2024)
Fine-tuning Large Language Model for Automated Algorithm Design
von: Liu, Fei, et al.
Veröffentlicht: (2025)
von: Liu, Fei, et al.
Veröffentlicht: (2025)
Privacy in Fine-tuning Large Language Models: Attacks, Defenses, and Future Directions
von: Du, Hao, et al.
Veröffentlicht: (2024)
von: Du, Hao, et al.
Veröffentlicht: (2024)
Pre-train and Fine-tune: Recommenders as Large Models
von: Jiang, Zhenhao, et al.
Veröffentlicht: (2025)
von: Jiang, Zhenhao, et al.
Veröffentlicht: (2025)
Rethinking Reinforcement fine-tuning of LLMs: A Multi-armed Bandit Learning Perspective
von: Hu, Xiao, et al.
Veröffentlicht: (2026)
von: Hu, Xiao, et al.
Veröffentlicht: (2026)
Booster: Tackling Harmful Fine-tuning for Large Language Models via Attenuating Harmful Perturbation
von: Huang, Tiansheng, et al.
Veröffentlicht: (2024)
von: Huang, Tiansheng, et al.
Veröffentlicht: (2024)
KnowTuning: Knowledge-aware Fine-tuning for Large Language Models
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
JudgeLM: Fine-tuned Large Language Models are Scalable Judges
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
Dhati+: Fine-tuned Large Language Models for Arabic Subjectivity Evaluation
von: Bellaouar, Slimane, et al.
Veröffentlicht: (2025)
von: Bellaouar, Slimane, et al.
Veröffentlicht: (2025)
Compositional Subspace Representation Fine-tuning for Adaptive Large Language Models
von: Zhou, Andy
Veröffentlicht: (2025)
von: Zhou, Andy
Veröffentlicht: (2025)
Ähnliche Einträge
-
Unlocking the Power of Function Vectors for Characterizing and Mitigating Catastrophic Forgetting in Continual Instruction Tuning
von: Jiang, Gangwei, et al.
Veröffentlicht: (2025) -
What Makes a Good Reasoning Chain? Uncovering Structural Patterns in Long Chain-of-Thought Reasoning
von: Jiang, Gangwei, et al.
Veröffentlicht: (2025) -
Mitigate Negative Transfer with Similarity Heuristic Lifelong Prompt Tuning
von: Wu, Chenyuan, et al.
Veröffentlicht: (2024) -
Scaling Reasoning Hop Exposes Weaknesses: Demystifying and Improving Hop Generalization in Large Language Models
von: Li, Zhaoyi, et al.
Veröffentlicht: (2026) -
InstructAV: Instruction Fine-tuning Large Language Models for Authorship Verification
von: Hu, Yujia, et al.
Veröffentlicht: (2024)