From Instance Training to Instruction Learning: Task Adapters Generation from Instructions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liao, Huanxuan, He, Shizhu, Xu, Yao, Zhang, Yuanzhe, Hao, Yanchao, Liu, Shengping, Liu, Kang, Zhao, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Awakening Augmented Generation: Learning to Awaken Internal Knowledge of Large Language Models for Question Answering
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
DATA: Decomposed Attention-based Task Adaptation for Rehearsal-Free Continual Learning
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
$\textit{SKIntern}$: Internalizing Symbolic Knowledge for Distilling Better CoT Capabilities into Small Language Models
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
Shuttle Between the Instructions and the Parameters of Large Language Models
von: Sun, Wangtao, et al.
Veröffentlicht: (2025)
von: Sun, Wangtao, et al.
Veröffentlicht: (2025)
Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement
von: Tan, Yuqiao, et al.
Veröffentlicht: (2025)
von: Tan, Yuqiao, et al.
Veröffentlicht: (2025)
Beyond Hard and Soft: Hybrid Context Compression for Balancing Local and Global Information Retention
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
ControlLM: Crafting Diverse Personalities for Language Models
von: Weng, Yixuan, et al.
Veröffentlicht: (2024)
von: Weng, Yixuan, et al.
Veröffentlicht: (2024)
ResAdapt: Adaptive Resolution for Efficient Multimodal Reasoning
von: Liao, Huanxuan, et al.
Veröffentlicht: (2026)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2026)
SparK: Query-Aware Unstructured Sparsity with Recoverable KV Cache Channel Pruning
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
CITI: Enhancing Tool Utilizing Ability in Large Language Models without Sacrificing General Performance
von: Hao, Yupu, et al.
Veröffentlicht: (2024)
von: Hao, Yupu, et al.
Veröffentlicht: (2024)
Beyond Instruction Following: Evaluating Inferential Rule Following of Large Language Models
von: Sun, Wangtao, et al.
Veröffentlicht: (2024)
von: Sun, Wangtao, et al.
Veröffentlicht: (2024)
ExpNote: Black-box Large Language Models are Better Task Solvers with Experience Notebook
von: Sun, Wangtao, et al.
Veröffentlicht: (2023)
von: Sun, Wangtao, et al.
Veröffentlicht: (2023)
SC-Taxo: Hierarchical Taxonomy Generation under Semantic Consistency Constraints using Large Language Models
von: Cai, Shiqiang, et al.
Veröffentlicht: (2026)
von: Cai, Shiqiang, et al.
Veröffentlicht: (2026)
Evaluating Personalized Tool-Augmented LLMs from the Perspectives of Personalization and Proactivity
von: Hao, Yupu, et al.
Veröffentlicht: (2025)
von: Hao, Yupu, et al.
Veröffentlicht: (2025)
The Zero-Step Thinking: An Empirical Study of Mode Selection as Harder Early Exit in Reasoning Models
von: Tan, Yuqiao, et al.
Veröffentlicht: (2025)
von: Tan, Yuqiao, et al.
Veröffentlicht: (2025)
Neural Incompatibility: The Unbridgeable Gap of Cross-Scale Parametric Knowledge Transfer in Large Language Models
von: Tan, Yuqiao, et al.
Veröffentlicht: (2025)
von: Tan, Yuqiao, et al.
Veröffentlicht: (2025)
S3Eval: A Synthetic, Scalable, Systematic Evaluation Suite for Large Language Models
von: Lei, Fangyu, et al.
Veröffentlicht: (2023)
von: Lei, Fangyu, et al.
Veröffentlicht: (2023)
MoELoRA: Contrastive Learning Guided Mixture of Experts on Parameter-Efficient Fine-Tuning for Large Language Models
von: Luo, Tongxu, et al.
Veröffentlicht: (2024)
von: Luo, Tongxu, et al.
Veröffentlicht: (2024)
LogLM: From Task-based to Instruction-based Automated Log Analysis
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
von: Liu, Yilun, et al.
Veröffentlicht: (2024)
Instruction-Tuned, but Not More Verifiable Instruction-Following: A Cross-Task Diagnosis for LoRA Adapters
von: Zou, Junyi
Veröffentlicht: (2026)
von: Zou, Junyi
Veröffentlicht: (2026)
From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space
von: Tan, Yuqiao, et al.
Veröffentlicht: (2026)
von: Tan, Yuqiao, et al.
Veröffentlicht: (2026)
Find Parent then Label Children: A Two-stage Taxonomy Completion Method with Pre-trained Language Model
von: Xia, Fei, et al.
Veröffentlicht: (2024)
von: Xia, Fei, et al.
Veröffentlicht: (2024)
Generate-on-Graph: Treat LLM as both Agent and KG in Incomplete Knowledge Graph Question Answering
von: Xu, Yao, et al.
Veröffentlicht: (2024)
von: Xu, Yao, et al.
Veröffentlicht: (2024)
BP4ER: Bootstrap Prompting for Explicit Reasoning in Medical Dialogue Generation
von: He, Yuhong, et al.
Veröffentlicht: (2024)
von: He, Yuhong, et al.
Veröffentlicht: (2024)
Reasoning-Table: Exploring Reinforcement Learning for Table Reasoning
von: Lei, Fangyu, et al.
Veröffentlicht: (2025)
von: Lei, Fangyu, et al.
Veröffentlicht: (2025)
Instruction Embedding: Latent Representations of Instructions Towards Task Identification
von: Li, Yiwei, et al.
Veröffentlicht: (2024)
von: Li, Yiwei, et al.
Veröffentlicht: (2024)
LLaSA: Large Language and Structured Data Assistant
von: Xu, Yao, et al.
Veröffentlicht: (2024)
von: Xu, Yao, et al.
Veröffentlicht: (2024)
On the Paradoxical Interference between Instruction-Following and Task Solving
von: Qi, Yunjia, et al.
Veröffentlicht: (2026)
von: Qi, Yunjia, et al.
Veröffentlicht: (2026)
Graph-oriented Instruction Tuning of Large Language Models for Generic Graph Mining
von: Tan, Yanchao, et al.
Veröffentlicht: (2024)
von: Tan, Yanchao, et al.
Veröffentlicht: (2024)
ItD: Large Language Models Can Teach Themselves Induction through Deduction
von: Sun, Wangtao, et al.
Veröffentlicht: (2024)
von: Sun, Wangtao, et al.
Veröffentlicht: (2024)
S$^3$HQA: A Three-Stage Approach for Multi-hop Text-Table Hybrid Question Answering
von: Lei, Fangyu, et al.
Veröffentlicht: (2023)
von: Lei, Fangyu, et al.
Veröffentlicht: (2023)
Mastering Symbolic Operations: Augmenting Language Models with Compiled Neural Networks
von: Weng, Yixuan, et al.
Veröffentlicht: (2023)
von: Weng, Yixuan, et al.
Veröffentlicht: (2023)
Multilingual Knowledge Graph Completion from Pretrained Language Models with Knowledge Constraints
von: Song, Ran, et al.
Veröffentlicht: (2024)
von: Song, Ran, et al.
Veröffentlicht: (2024)
An Incomplete Loop: Instruction Inference, Instruction Following, and In-context Learning in Language Models
von: Liu, Emmy, et al.
Veröffentlicht: (2024)
von: Liu, Emmy, et al.
Veröffentlicht: (2024)
Semantic Pivots Enable Cross-Lingual Transfer in Large Language Models
von: He, Kaiyu, et al.
Veröffentlicht: (2025)
von: He, Kaiyu, et al.
Veröffentlicht: (2025)
Instruction Diversity Drives Generalization To Unseen Tasks
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
von: Zhang, Dylan, et al.
Veröffentlicht: (2024)
Efficient Data Learning for Open Information Extraction with Pre-trained Language Models
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Fan, Zhiyuan, et al.
Veröffentlicht: (2023)
Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN
von: Xu, Yao, et al.
Veröffentlicht: (2025)
von: Xu, Yao, et al.
Veröffentlicht: (2025)
MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training
von: Huang, Hui, et al.
Veröffentlicht: (2025)
von: Huang, Hui, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Awakening Augmented Generation: Learning to Awaken Internal Knowledge of Large Language Models for Question Answering
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024) -
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024) -
DATA: Decomposed Attention-based Task Adaptation for Rehearsal-Free Continual Learning
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025) -
$\textit{SKIntern}$: Internalizing Symbolic Knowledge for Distilling Better CoT Capabilities into Small Language Models
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024) -
Shuttle Between the Instructions and the Parameters of Large Language Models
von: Sun, Wangtao, et al.
Veröffentlicht: (2025)