Layer-Wise Structural Mapping for Efficient Domain Transfer in Language Model Distillation
Fuente:
Zenodo
Saved in:
| Main Author: | Quan, Xuehui |
|---|---|
| Format: | Recurso digital |
| Published: |
Zenodo
2024
|
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DistilMOS: Layer-Wise Self-Distillation For Self-Supervised Learning Model-Based MOS Prediction
by: Yang, Jianing, et al.
Published: (2026)
by: Yang, Jianing, et al.
Published: (2026)
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
by: Sung, Yi-Lin, et al.
Published: (2023)
by: Sung, Yi-Lin, et al.
Published: (2023)
DDK: Distilling Domain Knowledge for Efficient Large Language Models
by: Liu, Jiaheng, et al.
Published: (2024)
by: Liu, Jiaheng, et al.
Published: (2024)
Pixel-Wise Contrastive Distillation
by: Huang, Junqiang, et al.
Published: (2022)
by: Huang, Junqiang, et al.
Published: (2022)
Iterative Layer-wise Distillation for Efficient Compression of Large Language Models
by: Kovalev, Grigory, et al.
Published: (2025)
by: Kovalev, Grigory, et al.
Published: (2025)
Adaptive Layer-Wise Transformations for Post-Training Quantization of Large Language Models
by: Pham, Cuong, et al.
Published: (2025)
by: Pham, Cuong, et al.
Published: (2025)
TAID: Temporally Adaptive Interpolated Distillation for Efficient Knowledge Transfer in Language Models
by: Shing, Makoto, et al.
Published: (2025)
by: Shing, Makoto, et al.
Published: (2025)
Causal Distillation: Transferring Structured Explanations from Large to Compact Language Models
by: Muhebwa, Aggrey, et al.
Published: (2025)
by: Muhebwa, Aggrey, et al.
Published: (2025)
Choosing Wisely and Learning Deeply: Selective Cross-Modality Distillation via CLIP for Domain Generalization
by: Leng, Jixuan, et al.
Published: (2023)
by: Leng, Jixuan, et al.
Published: (2023)
Detail Consistent Stage-Wise Distillation for Efficient 3D MRI Segmentation
by: Fan, Mengchen, et al.
Published: (2026)
by: Fan, Mengchen, et al.
Published: (2026)
BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation
by: Chen, Baoyou, et al.
Published: (2026)
by: Chen, Baoyou, et al.
Published: (2026)
OutlierTune: Efficient Channel-Wise Quantization for Large Language Models
by: Wang, Jinguang, et al.
Published: (2024)
by: Wang, Jinguang, et al.
Published: (2024)
ROSA: Robust and Energy-Efficient Microring-Based Optical Neural Networks via Optical Shift-and-Add and Layer-Wise Hybrid Mapping
by: Zhang, Huifan, et al.
Published: (2026)
by: Zhang, Huifan, et al.
Published: (2026)
KVSharer: Efficient Inference via Layer-Wise Dissimilar KV Cache Sharing
by: Yang, Yifei, et al.
Published: (2024)
by: Yang, Yifei, et al.
Published: (2024)
MoLT: Mixture of Layer-Wise Tokens for Efficient Audio-Visual Learning
by: Rho, Kyeongha, et al.
Published: (2025)
by: Rho, Kyeongha, et al.
Published: (2025)
SymptomWise: A Deterministic Reasoning Layer for Reliable and Efficient AI Systems
by: Henry, Isaac, et al.
Published: (2026)
by: Henry, Isaac, et al.
Published: (2026)
Hybrid Layer-Wise ANN-SNN With Surrogate Spike Encoding-Decoding Structure
by: Luu, Nhan T., et al.
Published: (2025)
by: Luu, Nhan T., et al.
Published: (2025)
How Large Language Models Encode Context Knowledge? A Layer-Wise Probing Study
by: Ju, Tianjie, et al.
Published: (2024)
by: Ju, Tianjie, et al.
Published: (2024)
TEON: Tensorized Orthonormalization Beyond Layer-Wise Muon for Large Language Model Pre-Training
by: Zhang, Ruijie, et al.
Published: (2026)
by: Zhang, Ruijie, et al.
Published: (2026)
NAYER: Noisy Layer Data Generation for Efficient and Effective Data-free Knowledge Distillation
by: Tran, Minh-Tuan, et al.
Published: (2023)
by: Tran, Minh-Tuan, et al.
Published: (2023)
Efficient Verified Machine Unlearning For Distillation
by: Quan, Yijun, et al.
Published: (2025)
by: Quan, Yijun, et al.
Published: (2025)
Layer-Wise High-Impact Parameter Ratio Optimization in Post-Training Quantization for Large Language Models
by: Pham, Cuong, et al.
Published: (2025)
by: Pham, Cuong, et al.
Published: (2025)
The Quantum Sieve Tracer: A Hybrid Framework for Layer-Wise Activation Tracing in Large Language Models
by: Pan, Jonathan
Published: (2026)
by: Pan, Jonathan
Published: (2026)
Explainability-Inspired Layer-Wise Pruning of Deep Neural Networks for Efficient Object Detection
by: Shukla, Abhinav, et al.
Published: (2026)
by: Shukla, Abhinav, et al.
Published: (2026)
DiffPro: Joint Timestep and Layer-Wise Precision Optimization for Efficient Diffusion Inference
by: Amin, Farhana, et al.
Published: (2025)
by: Amin, Farhana, et al.
Published: (2025)
Depth-Wise Attention (DWAtt): A Layer Fusion Method for Data-Efficient Classification
by: ElNokrashy, Muhammad, et al.
Published: (2022)
by: ElNokrashy, Muhammad, et al.
Published: (2022)
Structured Knowledge Accumulation: An Autonomous Framework for Layer-Wise Entropy Reduction in Neural Learning
by: Quantiota, Bouarfa Mahi
Published: (2025)
by: Quantiota, Bouarfa Mahi
Published: (2025)
Patch-Wise Hypergraph Contrastive Learning with Dual Normal Distribution Weighting for Multi-Domain Stain Transfer
by: Wei, Haiyan, et al.
Published: (2025)
by: Wei, Haiyan, et al.
Published: (2025)
How to Alleviate Catastrophic Forgetting in LLMs Finetuning? Hierarchical Layer-Wise and Element-Wise Regularization
by: Song, Shezheng, et al.
Published: (2025)
by: Song, Shezheng, et al.
Published: (2025)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
by: Taniguchi, Rei, et al.
Published: (2026)
by: Taniguchi, Rei, et al.
Published: (2026)
MapDistill: Boosting Efficient Camera-based HD Map Construction via Camera-LiDAR Fusion Model Distillation
by: Hao, Xiaoshuai, et al.
Published: (2024)
by: Hao, Xiaoshuai, et al.
Published: (2024)
Transferring Backdoors between Large Language Models by Knowledge Distillation
by: Cheng, Pengzhou, et al.
Published: (2024)
by: Cheng, Pengzhou, et al.
Published: (2024)
Dynamics Distillation for Efficient and Transferable Control Learning
by: Gu, Xunjiang, et al.
Published: (2026)
by: Gu, Xunjiang, et al.
Published: (2026)
PLD: A Choice-Theoretic List-Wise Knowledge Distillation
by: Bassam, Ejafa, et al.
Published: (2025)
by: Bassam, Ejafa, et al.
Published: (2025)
PANDA: Prompt Transfer Meets Knowledge Distillation for Efficient Model Adaptation
by: Zhong, Qihuang, et al.
Published: (2022)
by: Zhong, Qihuang, et al.
Published: (2022)
pMixFed: Efficient Personalized Federated Learning through Adaptive Layer-Wise Mixup
by: Saadati, Yasaman, et al.
Published: (2025)
by: Saadati, Yasaman, et al.
Published: (2025)
FLAM: Frame-Wise Language-Audio Modeling
by: Wu, Yusong, et al.
Published: (2025)
by: Wu, Yusong, et al.
Published: (2025)
KVPruner: Structural Pruning for Faster and Memory-Efficient Large Language Models
by: Lv, Bo, et al.
Published: (2024)
by: Lv, Bo, et al.
Published: (2024)
Time-Series Learning for Proactive Fault Prediction in Distributed Systems with Deep Neural Structures
by: Wang, Yang, et al.
Published: (2025)
by: Wang, Yang, et al.
Published: (2025)
PromptDistill: Query-based Selective Token Retention in Intermediate Layers for Efficient Large Language Model Inference
by: Jin, Weisheng, et al.
Published: (2025)
by: Jin, Weisheng, et al.
Published: (2025)
Similar Items
-
DistilMOS: Layer-Wise Self-Distillation For Self-Supervised Learning Model-Based MOS Prediction
by: Yang, Jianing, et al.
Published: (2026) -
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
by: Sung, Yi-Lin, et al.
Published: (2023) -
DDK: Distilling Domain Knowledge for Efficient Large Language Models
by: Liu, Jiaheng, et al.
Published: (2024) -
Pixel-Wise Contrastive Distillation
by: Huang, Junqiang, et al.
Published: (2022) -
Iterative Layer-wise Distillation for Efficient Compression of Large Language Models
by: Kovalev, Grigory, et al.
Published: (2025)