Towards Understanding and Improving Knowledge Distillation for Neural Machine Translation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Songming, Liang, Yunlong, Wang, Shuaibo, Han, Wenjuan, Liu, Jian, Xu, Jinan, Chen, Yufeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dual-Space Knowledge Distillation for Large Language Models
von: Zhang, Songming, et al.
Veröffentlicht: (2024)
von: Zhang, Songming, et al.
Veröffentlicht: (2024)
A Dual-Space Framework for General Knowledge Distillation of Large Language Models
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
KDFlow: A User-Friendly and Efficient Knowledge Distillation Framework for Large Language Models
von: Zhang, Songming, et al.
Veröffentlicht: (2026)
von: Zhang, Songming, et al.
Veröffentlicht: (2026)
AlignDistil: Token-Level Language Model Alignment as Adaptive Policy Distillation
von: Zhang, Songming, et al.
Veröffentlicht: (2025)
von: Zhang, Songming, et al.
Veröffentlicht: (2025)
LCS: A Language Converter Strategy for Zero-Shot Neural Machine Translation
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
Towards Faster k-Nearest-Neighbor Machine Translation
von: Shi, Xiangyu, et al.
Veröffentlicht: (2023)
von: Shi, Xiangyu, et al.
Veröffentlicht: (2023)
Multilingual Knowledge Editing with Language-Agnostic Factual Neurons
von: Zhang, Xue, et al.
Veröffentlicht: (2024)
von: Zhang, Xue, et al.
Veröffentlicht: (2024)
Outdated Issue Aware Decoding for Reasoning Questions on Edited Knowledge
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)
Improving Long Text Understanding with Knowledge Distilled from Summarization Model
von: Liu, Yan, et al.
Veröffentlicht: (2024)
von: Liu, Yan, et al.
Veröffentlicht: (2024)
Align-to-Distill: Trainable Attention Alignment for Knowledge Distillation in Neural Machine Translation
von: Jin, Heegon, et al.
Veröffentlicht: (2024)
von: Jin, Heegon, et al.
Veröffentlicht: (2024)
Less, but Better: Efficient Multilingual Expansion for LLMs via Layer-wise Mixture-of-Experts
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
CM-Align: Consistency-based Multilingual Alignment for Large Language Models
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
Understanding and Analyzing Model Robustness and Knowledge-Transfer in Multilingual Neural Machine Translation using TX-Ray
von: Saxena, Vageesh, et al.
Veröffentlicht: (2024)
von: Saxena, Vageesh, et al.
Veröffentlicht: (2024)
Efficient Technical Term Translation: A Knowledge Distillation Approach for Parenthetical Terminology Translation
von: Myung, Jiyoon, et al.
Veröffentlicht: (2024)
von: Myung, Jiyoon, et al.
Veröffentlicht: (2024)
CycleDistill: Bootstrapping Machine Translation using LLMs with Cyclical Distillation
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
von: Halder, Deepon, et al.
Veröffentlicht: (2025)
Evaluating Explainable AI Attribution Methods in Neural Machine Translation via Attention-Guided Knowledge Distillation
von: Nourbakhsh, Aria, et al.
Veröffentlicht: (2026)
von: Nourbakhsh, Aria, et al.
Veröffentlicht: (2026)
TEaR: Improving LLM-based Machine Translation with Systematic Self-Refinement
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
von: Feng, Zhaopeng, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Machine Translation with Unstructured Knowledge
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
Think Natively: Unlocking Multilingual Reasoning with Consistency-Enhanced Reinforcement Learning
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
von: Zhang, Xue, et al.
Veröffentlicht: (2025)
Improving Non-autoregressive Machine Translation with Error Exposure and Consistency Regularization
von: Chen, Xinran, et al.
Veröffentlicht: (2024)
von: Chen, Xinran, et al.
Veröffentlicht: (2024)
PMMT: Preference Alignment in Multilingual Machine Translation via LLM Distillation
von: Sun, Shuqiao, et al.
Veröffentlicht: (2024)
von: Sun, Shuqiao, et al.
Veröffentlicht: (2024)
Distilling Closed-Source LLM's Knowledge for Locally Stable and Economic Biomedical Entity Linking
von: Ai, Yihao, et al.
Veröffentlicht: (2025)
von: Ai, Yihao, et al.
Veröffentlicht: (2025)
DRT: Deep Reasoning Translation via Long Chain-of-Thought
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
von: Wang, Jiaan, et al.
Veröffentlicht: (2024)
CANTONMT: Investigating Back-Translation and Model-Switch Mechanisms for Cantonese-English Neural Machine Translation
von: Hong, Kung Yin, et al.
Veröffentlicht: (2024)
von: Hong, Kung Yin, et al.
Veröffentlicht: (2024)
Contrastive Token Learning with Similarity Decay for Repetition Suppression in Machine Translation
von: Dai, Huangyu, et al.
Veröffentlicht: (2024)
von: Dai, Huangyu, et al.
Veröffentlicht: (2024)
LLM-Oriented Token-Adaptive Knowledge Distillation
von: Xie, Xurong, et al.
Veröffentlicht: (2025)
von: Xie, Xurong, et al.
Veröffentlicht: (2025)
Towards Cross-Cultural Machine Translation with Retrieval-Augmented Generation from Multilingual Knowledge Graphs
von: Conia, Simone, et al.
Veröffentlicht: (2024)
von: Conia, Simone, et al.
Veröffentlicht: (2024)
WTU-EVAL: A Whether-or-Not Tool Usage Evaluation Benchmark for Large Language Models
von: Ning, Kangyun, et al.
Veröffentlicht: (2024)
von: Ning, Kangyun, et al.
Veröffentlicht: (2024)
From Words to Worlds: Benchmarking Cross-Cultural Cultural Understanding in Machine Translation
von: Han, Bangju, et al.
Veröffentlicht: (2026)
von: Han, Bangju, et al.
Veröffentlicht: (2026)
Cross-Lingual Knowledge Editing in Large Language Models
von: Wang, Jiaan, et al.
Veröffentlicht: (2023)
von: Wang, Jiaan, et al.
Veröffentlicht: (2023)
On the Shortcut Learning in Multilingual Neural Machine Translation
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
On Instruction-Finetuning Neural Machine Translation Models
von: Raunak, Vikas, et al.
Veröffentlicht: (2024)
von: Raunak, Vikas, et al.
Veröffentlicht: (2024)
DistillSpec: Improving Speculative Decoding via Knowledge Distillation
von: Zhou, Yongchao, et al.
Veröffentlicht: (2023)
von: Zhou, Yongchao, et al.
Veröffentlicht: (2023)
ACT-MNMT Auto-Constriction Turning for Multilingual Neural Machine Translation
von: Dai, Shaojie, et al.
Veröffentlicht: (2024)
von: Dai, Shaojie, et al.
Veröffentlicht: (2024)
Pedagogically-Inspired Data Synthesis for Language Model Knowledge Distillation
von: He, Bowei, et al.
Veröffentlicht: (2026)
von: He, Bowei, et al.
Veröffentlicht: (2026)
Towards Improving Interpretability of Language Model Generation through a Structured Knowledge Discovery Approach
von: Liu, Shuqi, et al.
Veröffentlicht: (2025)
von: Liu, Shuqi, et al.
Veröffentlicht: (2025)
Design of an Open-Source Architecture for Neural Machine Translation
von: Lankford, Séamus, et al.
Veröffentlicht: (2024)
von: Lankford, Séamus, et al.
Veröffentlicht: (2024)
Towards Understanding Continual Factual Knowledge Acquisition of Language Models: From Theory to Algorithm
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
von: Wang, Haoyu, et al.
Veröffentlicht: (2026)
Warmup-Distill: Bridge the Distribution Mismatch between Teacher and Student before Knowledge Distillation
von: Sun, Zengkui, et al.
Veröffentlicht: (2025)
von: Sun, Zengkui, et al.
Veröffentlicht: (2025)
A Two-Stage Framework with Self-Supervised Distillation For Cross-Domain Text Classification
von: Feng, Yunlong, et al.
Veröffentlicht: (2023)
von: Feng, Yunlong, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Dual-Space Knowledge Distillation for Large Language Models
von: Zhang, Songming, et al.
Veröffentlicht: (2024) -
A Dual-Space Framework for General Knowledge Distillation of Large Language Models
von: Zhang, Xue, et al.
Veröffentlicht: (2025) -
KDFlow: A User-Friendly and Efficient Knowledge Distillation Framework for Large Language Models
von: Zhang, Songming, et al.
Veröffentlicht: (2026) -
AlignDistil: Token-Level Language Model Alignment as Adaptive Policy Distillation
von: Zhang, Songming, et al.
Veröffentlicht: (2025) -
LCS: A Language Converter Strategy for Zero-Shot Neural Machine Translation
von: Sun, Zengkui, et al.
Veröffentlicht: (2024)