Enhancing Knowledge Distillation of Large Language Models through Efficient Multi-Modal Distribution Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Tianyu, Zhang, Jiajun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models
von: Wang, Shuo, et al.
Veröffentlicht: (2024)
von: Wang, Shuo, et al.
Veröffentlicht: (2024)
Feature Alignment and Representation Transfer in Knowledge Distillation for Large Language Models
von: Yang, Junjie, et al.
Veröffentlicht: (2025)
von: Yang, Junjie, et al.
Veröffentlicht: (2025)
DDK: Distilling Domain Knowledge for Efficient Large Language Models
von: Liu, Jiaheng, et al.
Veröffentlicht: (2024)
von: Liu, Jiaheng, et al.
Veröffentlicht: (2024)
Cross-Modal Knowledge Distillation for Speech Large Language Models
von: Wang, Enzhi, et al.
Veröffentlicht: (2025)
von: Wang, Enzhi, et al.
Veröffentlicht: (2025)
Exploring and Enhancing the Transfer of Distribution in Knowledge Distillation for Autoregressive Language Models
von: Rao, Jun, et al.
Veröffentlicht: (2024)
von: Rao, Jun, et al.
Veröffentlicht: (2024)
MT-PATCHER: Selective and Extendable Knowledge Distillation from Large Language Models for Machine Translation
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
MTA: Multi-Granular Trajectory Alignment for Large Language Model Distillation
von: Chi, Pham Khanh, et al.
Veröffentlicht: (2026)
von: Chi, Pham Khanh, et al.
Veröffentlicht: (2026)
Enhancing Multilingual Capabilities of Large Language Models through Self-Distillation from Resource-Rich Languages
von: Zhang, Yuanchi, et al.
Veröffentlicht: (2024)
von: Zhang, Yuanchi, et al.
Veröffentlicht: (2024)
Enhancing Multi-hop Reasoning through Knowledge Erasure in Large Language Model Editing
von: Zhang, Mengqi, et al.
Veröffentlicht: (2024)
von: Zhang, Mengqi, et al.
Veröffentlicht: (2024)
Generative Multi-Modal Knowledge Retrieval with Large Language Models
von: Long, Xinwei, et al.
Veröffentlicht: (2024)
von: Long, Xinwei, et al.
Veröffentlicht: (2024)
Large Language Models are Limited in Out-of-Context Knowledge Reasoning
von: Hu, Peng, et al.
Veröffentlicht: (2024)
von: Hu, Peng, et al.
Veröffentlicht: (2024)
KALE: Enhancing Knowledge Manipulation in Large Language Models via Knowledge-aware Learning
von: Lv, Qitan, et al.
Veröffentlicht: (2026)
von: Lv, Qitan, et al.
Veröffentlicht: (2026)
Leveraging Large Language Models for Enhanced NLP Task Performance through Knowledge Distillation and Optimized Training Strategies
von: Huang, Yining, et al.
Veröffentlicht: (2024)
von: Huang, Yining, et al.
Veröffentlicht: (2024)
Contrastive Knowledge Transfer and Robust Optimization for Secure Alignment of Large Language Models
von: Zheng, Jiasen, et al.
Veröffentlicht: (2025)
von: Zheng, Jiasen, et al.
Veröffentlicht: (2025)
Knowledge Distillation of Black-Box Large Language Models
von: Chen, Hongzhan, et al.
Veröffentlicht: (2024)
von: Chen, Hongzhan, et al.
Veröffentlicht: (2024)
Multi-Aspect Knowledge Distillation for Language Model with Low-rank Factorization
von: Liu, Zihe, et al.
Veröffentlicht: (2026)
von: Liu, Zihe, et al.
Veröffentlicht: (2026)
Less is More: Selective Reflection for Compatible and Efficient Knowledge Distillation in Large Language Models
von: Liu, Lingyuan, et al.
Veröffentlicht: (2025)
von: Liu, Lingyuan, et al.
Veröffentlicht: (2025)
Dual-Space Knowledge Distillation for Large Language Models
von: Zhang, Songming, et al.
Veröffentlicht: (2024)
von: Zhang, Songming, et al.
Veröffentlicht: (2024)
Being Strong Progressively! Enhancing Knowledge Distillation of Large Language Models through a Curriculum Learning Framework
von: Liu, Lingyuan, et al.
Veröffentlicht: (2025)
von: Liu, Lingyuan, et al.
Veröffentlicht: (2025)
From Images to Words: Efficient Cross-Modal Knowledge Distillation to Language Models from Black-box Teachers
von: Sengupta, Ayan, et al.
Veröffentlicht: (2026)
von: Sengupta, Ayan, et al.
Veröffentlicht: (2026)
Efficient Knowledge Transfer in Multi-Task Learning through Task-Adaptive Low-Rank Representation
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
SRA: Span Representation Alignment for Large Language Model Distillation
von: Dao, Quoc Phong, et al.
Veröffentlicht: (2026)
von: Dao, Quoc Phong, et al.
Veröffentlicht: (2026)
Enhancing Romanian Offensive Language Detection through Knowledge Distillation, Multi-Task Learning, and Data Augmentation
von: Matei, Vlad-Cristian, et al.
Veröffentlicht: (2024)
von: Matei, Vlad-Cristian, et al.
Veröffentlicht: (2024)
Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration Rate
von: Huang, Qidong, et al.
Veröffentlicht: (2024)
von: Huang, Qidong, et al.
Veröffentlicht: (2024)
Direct Preference Knowledge Distillation for Large Language Models
von: Li, Yixing, et al.
Veröffentlicht: (2024)
von: Li, Yixing, et al.
Veröffentlicht: (2024)
A Survey on Knowledge Distillation of Large Language Models
von: Xu, Xiaohan, et al.
Veröffentlicht: (2024)
von: Xu, Xiaohan, et al.
Veröffentlicht: (2024)
Distribution Corrected Offline Data Distillation for Large Language Models
von: Zhang, Yumeng, et al.
Veröffentlicht: (2026)
von: Zhang, Yumeng, et al.
Veröffentlicht: (2026)
Distilling Rule-based Knowledge into Large Language Models
von: Yang, Wenkai, et al.
Veröffentlicht: (2023)
von: Yang, Wenkai, et al.
Veröffentlicht: (2023)
Knowledge Distillation for Temporal Knowledge Graph Reasoning with Large Language Models
von: Xing, Wang, et al.
Veröffentlicht: (2026)
von: Xing, Wang, et al.
Veröffentlicht: (2026)
Multi-Sense Embeddings for Language Models and Knowledge Distillation
von: Wang, Qitong, et al.
Veröffentlicht: (2025)
von: Wang, Qitong, et al.
Veröffentlicht: (2025)
KDFlow: A User-Friendly and Efficient Knowledge Distillation Framework for Large Language Models
von: Zhang, Songming, et al.
Veröffentlicht: (2026)
von: Zhang, Songming, et al.
Veröffentlicht: (2026)
Can Large Models Teach Student Models to Solve Mathematical Problems Like Human Beings? A Reasoning Distillation Method via Multi-LoRA Interaction
von: Li, Xinhe, et al.
Veröffentlicht: (2025)
von: Li, Xinhe, et al.
Veröffentlicht: (2025)
DRPruning: Efficient Large Language Model Pruning through Distributionally Robust Optimization
von: Deng, Hexuan, et al.
Veröffentlicht: (2024)
von: Deng, Hexuan, et al.
Veröffentlicht: (2024)
Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning
von: Yang, Zhaorui, et al.
Veröffentlicht: (2024)
von: Yang, Zhaorui, et al.
Veröffentlicht: (2024)
LLM-NEO: Parameter Efficient Knowledge Distillation for Large Language Models
von: Yang, Runming, et al.
Veröffentlicht: (2024)
von: Yang, Runming, et al.
Veröffentlicht: (2024)
SWITCH: Studying with Teacher for Knowledge Distillation of Large Language Models
von: Koo, Jahyun, et al.
Veröffentlicht: (2024)
von: Koo, Jahyun, et al.
Veröffentlicht: (2024)
Evolving Knowledge Distillation with Large Language Models and Active Learning
von: Liu, Chengyuan, et al.
Veröffentlicht: (2024)
von: Liu, Chengyuan, et al.
Veröffentlicht: (2024)
Progressively Modality Freezing for Multi-Modal Entity Alignment
von: Huang, Yani, et al.
Veröffentlicht: (2024)
von: Huang, Yani, et al.
Veröffentlicht: (2024)
Enhancing Large Language Models with Reliable Knowledge Graphs
von: Zhang, Qinggang
Veröffentlicht: (2025)
von: Zhang, Qinggang
Veröffentlicht: (2025)
Knowledge Distillation for Large Language Models
von: La Torre, Alejandro Paredes, et al.
Veröffentlicht: (2026)
von: La Torre, Alejandro Paredes, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models
von: Wang, Shuo, et al.
Veröffentlicht: (2024) -
Feature Alignment and Representation Transfer in Knowledge Distillation for Large Language Models
von: Yang, Junjie, et al.
Veröffentlicht: (2025) -
DDK: Distilling Domain Knowledge for Efficient Large Language Models
von: Liu, Jiaheng, et al.
Veröffentlicht: (2024) -
Cross-Modal Knowledge Distillation for Speech Large Language Models
von: Wang, Enzhi, et al.
Veröffentlicht: (2025) -
Exploring and Enhancing the Transfer of Distribution in Knowledge Distillation for Autoregressive Language Models
von: Rao, Jun, et al.
Veröffentlicht: (2024)