GOLD: Generalized Knowledge Distillation via Out-of-Distribution-Guided Language Data Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Gholami, Mohsen, Akbari, Mohammad, Hu, Cindy, Masrani, Vaden, Wang, Z. Jane, Zhang, Yong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Task-Agnostic Language Model Watermarking via High Entropy Passthrough Layers
por: Masrani, Vaden, et al.
Publicado: (2024)
por: Masrani, Vaden, et al.
Publicado: (2024)
CASP: Compression of Large Multimodal Models Based on Attention Sparsity
por: Gholami, Mohsen, et al.
Publicado: (2025)
por: Gholami, Mohsen, et al.
Publicado: (2025)
TAIA: Large Language Models are Out-of-Distribution Data Learners
por: Jiang, Shuyang, et al.
Publicado: (2024)
por: Jiang, Shuyang, et al.
Publicado: (2024)
Extracting General-use Transformers for Low-resource Languages via Knowledge Distillation
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2025)
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2025)
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model
por: Brown, Andrew, et al.
Publicado: (2024)
por: Brown, Andrew, et al.
Publicado: (2024)
Rainproof: An Umbrella To Shield Text Generators From Out-Of-Distribution Data
por: Darrin, Maxime, et al.
Publicado: (2022)
por: Darrin, Maxime, et al.
Publicado: (2022)
GOLD: Geometry Problem Solver with Natural Language Description
por: Zhang, Jiaxin, et al.
Publicado: (2024)
por: Zhang, Jiaxin, et al.
Publicado: (2024)
LLM-Guided Knowledge Distillation for Temporal Knowledge Graph Reasoning
por: Xing, Wang, et al.
Publicado: (2026)
por: Xing, Wang, et al.
Publicado: (2026)
Distilling Implicit Multimodal Knowledge into Large Language Models for Zero-Resource Dialogue Generation
por: Zhang, Bo, et al.
Publicado: (2024)
por: Zhang, Bo, et al.
Publicado: (2024)
Exploring and Enhancing the Transfer of Distribution in Knowledge Distillation for Autoregressive Language Models
por: Rao, Jun, et al.
Publicado: (2024)
por: Rao, Jun, et al.
Publicado: (2024)
LLM-based Privacy Data Augmentation Guided by Knowledge Distillation with a Distribution Tutor for Medical Text Classification
por: Song, Yiping, et al.
Publicado: (2024)
por: Song, Yiping, et al.
Publicado: (2024)
Out-of-Distribution Detection using Synthetic Data Generation
por: Abbas, Momin, et al.
Publicado: (2025)
por: Abbas, Momin, et al.
Publicado: (2025)
Distribution Corrected Offline Data Distillation for Large Language Models
por: Zhang, Yumeng, et al.
Publicado: (2026)
por: Zhang, Yumeng, et al.
Publicado: (2026)
$\mathcal{X}$-KD: General Experiential Knowledge Distillation for Large Language Models
por: Cai, Yuang, et al.
Publicado: (2026)
por: Cai, Yuang, et al.
Publicado: (2026)
On the Generalization vs Fidelity Paradox in Knowledge Distillation
por: Ramesh, Suhas Kamasetty, et al.
Publicado: (2025)
por: Ramesh, Suhas Kamasetty, et al.
Publicado: (2025)
COLLEAGUE.SKILL: Automated AI Skill Generation via Expert Knowledge Distillation
por: Zhou, Tianyi, et al.
Publicado: (2026)
por: Zhou, Tianyi, et al.
Publicado: (2026)
Scaling Knowledge Graph Construction through Synthetic Data Generation and Distillation
por: Choubey, Prafulla Kumar, et al.
Publicado: (2024)
por: Choubey, Prafulla Kumar, et al.
Publicado: (2024)
Multi-MLLM Knowledge Distillation for Out-of-Context News Detection
por: Gu, Yimeng, et al.
Publicado: (2025)
por: Gu, Yimeng, et al.
Publicado: (2025)
Enhancing Knowledge Distillation of Large Language Models through Efficient Multi-Modal Distribution Alignment
por: Peng, Tianyu, et al.
Publicado: (2024)
por: Peng, Tianyu, et al.
Publicado: (2024)
Large Language Models are Limited in Out-of-Context Knowledge Reasoning
por: Hu, Peng, et al.
Publicado: (2024)
por: Hu, Peng, et al.
Publicado: (2024)
MMIDR: Teaching Large Language Model to Interpret Multimodal Misinformation via Knowledge Distillation
por: Wang, Longzheng, et al.
Publicado: (2024)
por: Wang, Longzheng, et al.
Publicado: (2024)
Compact Language Models via Pruning and Knowledge Distillation
por: Muralidharan, Saurav, et al.
Publicado: (2024)
por: Muralidharan, Saurav, et al.
Publicado: (2024)
Knowledge Graph-Guided Retrieval Augmented Generation
por: Zhu, Xiangrong, et al.
Publicado: (2025)
por: Zhu, Xiangrong, et al.
Publicado: (2025)
A Dual-Space Framework for General Knowledge Distillation of Large Language Models
por: Zhang, Xue, et al.
Publicado: (2025)
por: Zhang, Xue, et al.
Publicado: (2025)
LLM-Oriented Token-Adaptive Knowledge Distillation
por: Xie, Xurong, et al.
Publicado: (2025)
por: Xie, Xurong, et al.
Publicado: (2025)
Pedagogically-Inspired Data Synthesis for Language Model Knowledge Distillation
por: He, Bowei, et al.
Publicado: (2026)
por: He, Bowei, et al.
Publicado: (2026)
Privacy-Preserving Reasoning with Knowledge-Distilled Parametric Retrieval Augmented Generation
por: Chen, Jinwen, et al.
Publicado: (2025)
por: Chen, Jinwen, et al.
Publicado: (2025)
EnrichEvent: Enriching Social Data with Contextual Information for Emerging Event Extraction
por: Esfahani, Mohammadali Sefidi, et al.
Publicado: (2023)
por: Esfahani, Mohammadali Sefidi, et al.
Publicado: (2023)
EGAD: Entropy-Guided Adaptive Distillation for Token-Level Knowledge Transfer
por: Zhang, Hao, et al.
Publicado: (2026)
por: Zhang, Hao, et al.
Publicado: (2026)
Does Knowledge Distillation Matter for Large Language Model based Bundle Generation?
por: Feng, Kaidong, et al.
Publicado: (2025)
por: Feng, Kaidong, et al.
Publicado: (2025)
DSG-KD: Knowledge Distillation from Domain-Specific to General Language Models
por: Cho, Sangyeon, et al.
Publicado: (2024)
por: Cho, Sangyeon, et al.
Publicado: (2024)
Knowledge Distillation for Temporal Knowledge Graph Reasoning with Large Language Models
por: Xing, Wang, et al.
Publicado: (2026)
por: Xing, Wang, et al.
Publicado: (2026)
Differentially Private Knowledge Distillation via Synthetic Text Generation
por: Flemings, James, et al.
Publicado: (2024)
por: Flemings, James, et al.
Publicado: (2024)
PromptKD: Distilling Student-Friendly Knowledge for Generative Language Models via Prompt Tuning
por: Kim, Gyeongman, et al.
Publicado: (2024)
por: Kim, Gyeongman, et al.
Publicado: (2024)
ELAD: Explanation-Guided Large Language Models Active Distillation
por: Zhang, Yifei, et al.
Publicado: (2024)
por: Zhang, Yifei, et al.
Publicado: (2024)
Routing Distilled Knowledge via Mixture of LoRA Experts for Large Language Model based Bundle Generation
por: Feng, Kaidong, et al.
Publicado: (2025)
por: Feng, Kaidong, et al.
Publicado: (2025)
DDK: Distilling Domain Knowledge for Efficient Large Language Models
por: Liu, Jiaheng, et al.
Publicado: (2024)
por: Liu, Jiaheng, et al.
Publicado: (2024)
Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation
por: Shen, Hanwen, et al.
Publicado: (2026)
por: Shen, Hanwen, et al.
Publicado: (2026)
Gumbel Distillation for Parallel Text Generation
por: Zhang, Chi, et al.
Publicado: (2026)
por: Zhang, Chi, et al.
Publicado: (2026)
Guiding Generative Storytelling with Knowledge Graphs
por: Pan, Zhijun, et al.
Publicado: (2025)
por: Pan, Zhijun, et al.
Publicado: (2025)
Ejemplares similares
-
Task-Agnostic Language Model Watermarking via High Entropy Passthrough Layers
por: Masrani, Vaden, et al.
Publicado: (2024) -
CASP: Compression of Large Multimodal Models Based on Attention Sparsity
por: Gholami, Mohsen, et al.
Publicado: (2025) -
TAIA: Large Language Models are Out-of-Distribution Data Learners
por: Jiang, Shuyang, et al.
Publicado: (2024) -
Extracting General-use Transformers for Low-resource Languages via Knowledge Distillation
por: Cruz, Jan Christian Blaise, et al.
Publicado: (2025) -
Generation, Distillation and Evaluation of Motivational Interviewing-Style Reflections with a Foundational Language Model
por: Brown, Andrew, et al.
Publicado: (2024)