PILLOW: Enhancing Efficient Instruction Fine-tuning via Prompt Matching
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qi, Zhenting, Tan, Xiaoyu, Shi, Shaojie, Qu, Chao, Xu, Yinghui, Qi, Yuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AURORA:Automated Training Framework of Universal Process Reward Models via Ensemble Prompting and Reverse Verification
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2025)
Thought-Like-Pro: Enhancing Reasoning of Large Language Models through Self-Driven Prolog-based Chain-of-Thought
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024)
Struct-X: Enhancing Large Language Models Reasoning with Structured Data
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024)
SCP-116K: A High-Quality Problem-Solution Dataset and a Generalized Pipeline for Automated Extraction in the Higher Education Science Domain
von: Lu, Dakuan, et al.
Veröffentlicht: (2025)
von: Lu, Dakuan, et al.
Veröffentlicht: (2025)
PARA: Parameter-Efficient Fine-tuning with Prompt Aware Representation Adjustment
von: Liu, Zequan, et al.
Veröffentlicht: (2025)
von: Liu, Zequan, et al.
Veröffentlicht: (2025)
Towards Collaborative Intelligence: Propagating Intentions and Reasoning for Multi-Agent Coordination with Large Language Models
von: Qiu, Xihe, et al.
Veröffentlicht: (2024)
von: Qiu, Xihe, et al.
Veröffentlicht: (2024)
Reflective Personalization Optimization: A Post-hoc Rewriting Framework for Black-Box Large Language Models
von: Hao, Teqi, et al.
Veröffentlicht: (2025)
von: Hao, Teqi, et al.
Veröffentlicht: (2025)
Enhancing Complex Instruction Following for Large Language Models with Mixture-of-Contexts Fine-tuning
von: Lu, Yuheng, et al.
Veröffentlicht: (2025)
von: Lu, Yuheng, et al.
Veröffentlicht: (2025)
DePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning
von: Shi, Zhengxiang, et al.
Veröffentlicht: (2023)
von: Shi, Zhengxiang, et al.
Veröffentlicht: (2023)
Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning
von: Tan, Zhaoxuan, et al.
Veröffentlicht: (2024)
von: Tan, Zhaoxuan, et al.
Veröffentlicht: (2024)
CogniDual Framework: Self-Training Large Language Models within a Dual-System Theoretical Framework for Improving Cognitive Tasks
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
ADePT: Adaptive Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning
von: Tang, Pengwei, et al.
Veröffentlicht: (2025)
von: Tang, Pengwei, et al.
Veröffentlicht: (2025)
PEDRO: Parameter-Efficient Fine-tuning with Prompt DEpenDent Representation MOdification
von: Xie, Tianfang, et al.
Veröffentlicht: (2024)
von: Xie, Tianfang, et al.
Veröffentlicht: (2024)
FIPO: Free-form Instruction-oriented Prompt Optimization with Preference Dataset and Modular Fine-tuning Schema
von: Lu, Junru, et al.
Veröffentlicht: (2024)
von: Lu, Junru, et al.
Veröffentlicht: (2024)
Resource-Efficient Adaptation of Large Language Models for Text Embeddings via Prompt Engineering and Contrastive Fine-tuning
von: Roth, Benedikt, et al.
Veröffentlicht: (2025)
von: Roth, Benedikt, et al.
Veröffentlicht: (2025)
Fine-tuning Large Language Models with Sequential Instructions
von: Hu, Hanxu, et al.
Veröffentlicht: (2024)
von: Hu, Hanxu, et al.
Veröffentlicht: (2024)
Learning or Self-aligning? Rethinking Instruction Fine-tuning
von: Ren, Mengjie, et al.
Veröffentlicht: (2024)
von: Ren, Mengjie, et al.
Veröffentlicht: (2024)
Utility-Diversity Aware Online Batch Selection for LLM Supervised Fine-tuning
von: Zou, Heming, et al.
Veröffentlicht: (2025)
von: Zou, Heming, et al.
Veröffentlicht: (2025)
MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching
von: Qu, Changle, et al.
Veröffentlicht: (2026)
von: Qu, Changle, et al.
Veröffentlicht: (2026)
Prompt-tuning for Clickbait Detection via Text Summarization
von: Deng, Haoxiang, et al.
Veröffentlicht: (2024)
von: Deng, Haoxiang, et al.
Veröffentlicht: (2024)
ElaLoRA: Elastic & Learnable Low-Rank Adaptation for Efficient Model Fine-Tuning
von: Chang, Huandong, et al.
Veröffentlicht: (2025)
von: Chang, Huandong, et al.
Veröffentlicht: (2025)
On the Loss of Context-awareness in General Instruction Fine-tuning
von: Wang, Yihan, et al.
Veröffentlicht: (2024)
von: Wang, Yihan, et al.
Veröffentlicht: (2024)
Fine-tuning Done Right in Model Editing
von: Yang, Wanli, et al.
Veröffentlicht: (2025)
von: Yang, Wanli, et al.
Veröffentlicht: (2025)
A General Framework to Enhance Fine-tuning-based LLM Unlearning
von: Ren, Jie, et al.
Veröffentlicht: (2025)
von: Ren, Jie, et al.
Veröffentlicht: (2025)
GraphIF: Enhancing Multi-Turn Instruction Following for Large Language Models with Relation Graph Prompt
von: Li, Zhenhe, et al.
Veröffentlicht: (2025)
von: Li, Zhenhe, et al.
Veröffentlicht: (2025)
Follow My Instruction and Spill the Beans: Scalable Data Extraction from Retrieval-Augmented Generation Systems
von: Qi, Zhenting, et al.
Veröffentlicht: (2024)
von: Qi, Zhenting, et al.
Veröffentlicht: (2024)
LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning
von: Zhang, Longteng, et al.
Veröffentlicht: (2023)
von: Zhang, Longteng, et al.
Veröffentlicht: (2023)
Enhancing Multimodal Large Language Models for Ancient Chinese Character Evolution Analysis via Glyph-Driven Fine-Tuning
von: Song, Rui, et al.
Veröffentlicht: (2026)
von: Song, Rui, et al.
Veröffentlicht: (2026)
Prompt-tuning with Attribute Guidance for Low-resource Entity Matching
von: Liu, Lihui, et al.
Veröffentlicht: (2026)
von: Liu, Lihui, et al.
Veröffentlicht: (2026)
Promoting Equality in Large Language Models: Identifying and Mitigating the Implicit Bias based on Bayesian Theory
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
von: Deng, Yongxin, et al.
Veröffentlicht: (2024)
DoG-Instruct: Towards Premium Instruction-Tuning Data via Text-Grounded Instruction Wrapping
von: Chen, Yongrui, et al.
Veröffentlicht: (2023)
von: Chen, Yongrui, et al.
Veröffentlicht: (2023)
Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment
von: Wang, Jiongxiao, et al.
Veröffentlicht: (2024)
von: Wang, Jiongxiao, et al.
Veröffentlicht: (2024)
Are LLMs Effective Backbones for Fine-tuning? An Experimental Investigation of Supervised LLMs on Chinese Short Text Matching
von: Liu, Shulin, et al.
Veröffentlicht: (2024)
von: Liu, Shulin, et al.
Veröffentlicht: (2024)
UMB@PerAnsSumm 2025: Enhancing Perspective-Aware Summarization with Prompt Optimization and Supervised Fine-Tuning
von: Qi, Kristin, et al.
Veröffentlicht: (2025)
von: Qi, Kristin, et al.
Veröffentlicht: (2025)
MultiLingPoT: Enhancing Mathematical Reasoning with Multilingual Program Fine-tuning
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
von: Li, Nianqi, et al.
Veröffentlicht: (2024)
InsCL: A Data-efficient Continual Learning Paradigm for Fine-tuning Large Language Models with Instructions
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
Small Updates, Big Doubts: Does Parameter-Efficient Fine-tuning Enhance Hallucination Detection ?
von: Hu, Xu, et al.
Veröffentlicht: (2026)
von: Hu, Xu, et al.
Veröffentlicht: (2026)
Beyond Prompt Content: Enhancing LLM Performance via Content-Format Integrated Prompt Optimization
von: Liu, Yuanye, et al.
Veröffentlicht: (2025)
von: Liu, Yuanye, et al.
Veröffentlicht: (2025)
COIG-CQIA: Quality is All You Need for Chinese Instruction Fine-tuning
von: Bai, Yuelin, et al.
Veröffentlicht: (2024)
von: Bai, Yuelin, et al.
Veröffentlicht: (2024)
Fine-tuning with Hierarchical Prompting for Robust Propaganda Classification Across Annotation Schemas
von: Stähelin, Lukas, et al.
Veröffentlicht: (2026)
von: Stähelin, Lukas, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
AURORA:Automated Training Framework of Universal Process Reward Models via Ensemble Prompting and Reverse Verification
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2025) -
Thought-Like-Pro: Enhancing Reasoning of Large Language Models through Self-Driven Prolog-based Chain-of-Thought
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024) -
Struct-X: Enhancing Large Language Models Reasoning with Structured Data
von: Tan, Xiaoyu, et al.
Veröffentlicht: (2024) -
SCP-116K: A High-Quality Problem-Solution Dataset and a Generalized Pipeline for Automated Extraction in the Higher Education Science Domain
von: Lu, Dakuan, et al.
Veröffentlicht: (2025) -
PARA: Parameter-Efficient Fine-tuning with Prompt Aware Representation Adjustment
von: Liu, Zequan, et al.
Veröffentlicht: (2025)