Boosting LLM Translation Skills without General Ability Loss via Rationale Distillation
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Junhong, Zhao, Yang, Xu, Yangyifan, Liu, Bing, Zong, Chengqing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
F-MALLOC: Feed-forward Memory Allocation for Continual Learning in Neural Machine Translation
by: Wu, Junhong, et al.
Published: (2024)
by: Wu, Junhong, et al.
Published: (2024)
SimulPL: Aligning Human Preferences in Simultaneous Machine Translation
by: Yu, Donglei, et al.
Published: (2025)
by: Yu, Donglei, et al.
Published: (2025)
Hit the Sweet Spot! Span-Level Ensemble for Large Language Models
by: Xu, Yangyifan, et al.
Published: (2024)
by: Xu, Yangyifan, et al.
Published: (2024)
LADM: Long-context Training Data Selection with Attention-based Dependency Measurement for LLMs
by: Chen, Jianghao, et al.
Published: (2025)
by: Chen, Jianghao, et al.
Published: (2025)
Parallel Scaling Law: Unveiling Reasoning Generalization through A Cross-Linguistic Perspective
by: Yang, Wen, et al.
Published: (2025)
by: Yang, Wen, et al.
Published: (2025)
Language Imbalance Driven Rewarding for Multilingual Self-improving
by: Yang, Wen, et al.
Published: (2024)
by: Yang, Wen, et al.
Published: (2024)
Implicit Cross-Lingual Rewarding for Efficient Multilingual Preference Alignment
by: Yang, Wen, et al.
Published: (2025)
by: Yang, Wen, et al.
Published: (2025)
Bridging the Gap between Different Vocabularies for LLM Ensemble
by: Xu, Yangyifan, et al.
Published: (2024)
by: Xu, Yangyifan, et al.
Published: (2024)
Self-Modifying State Modeling for Simultaneous Machine Translation
by: Yu, Donglei, et al.
Published: (2024)
by: Yu, Donglei, et al.
Published: (2024)
TokAlign: Efficient Vocabulary Adaptation via Token Alignment
by: Li, Chong, et al.
Published: (2025)
by: Li, Chong, et al.
Published: (2025)
Bridging Relevance and Reasoning: Rationale Distillation in Retrieval-Augmented Generation
by: Jia, Pengyue, et al.
Published: (2024)
by: Jia, Pengyue, et al.
Published: (2024)
The Fine-Tuning Paradox: Boosting Translation Quality Without Sacrificing LLM Abilities
by: Stap, David, et al.
Published: (2024)
by: Stap, David, et al.
Published: (2024)
A Survey of Large Language Models in Discipline-specific Research: Challenges, Methods and Opportunities
by: Xiang, Lu, et al.
Published: (2025)
by: Xiang, Lu, et al.
Published: (2025)
Decoding the Multimodal Mind: Generalizable Brain-to-Text Translation via Multimodal Alignment and Adaptive Routing
by: Ye, Chunyu, et al.
Published: (2025)
by: Ye, Chunyu, et al.
Published: (2025)
RDRec: Rationale Distillation for LLM-based Recommendation
by: Wang, Xinfeng, et al.
Published: (2024)
by: Wang, Xinfeng, et al.
Published: (2024)
BLSP-Emo: Towards Empathetic Large Speech-Language Models
by: Wang, Chen, et al.
Published: (2024)
by: Wang, Chen, et al.
Published: (2024)
BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing
by: Wang, Chen, et al.
Published: (2023)
by: Wang, Chen, et al.
Published: (2023)
Structural Rationale Distillation via Reasoning Space Compression
by: Yang, Jialin, et al.
Published: (2026)
by: Yang, Jialin, et al.
Published: (2026)
Improving In-context Learning of Multilingual Generative Language Models with Cross-lingual Alignment
by: Li, Chong, et al.
Published: (2023)
by: Li, Chong, et al.
Published: (2023)
From Generic Empathy to Personalized Emotional Support: A Self-Evolution Framework for User Preference Alignment
by: Ye, Jing, et al.
Published: (2025)
by: Ye, Jing, et al.
Published: (2025)
TokAlign++: Advancing Vocabulary Adaptation via Better Token Alignment
by: Li, Chong, et al.
Published: (2026)
by: Li, Chong, et al.
Published: (2026)
Improving MLLM's Document Image Machine Translation via Synchronously Self-reviewing Its OCR Proficiency
by: Liang, Yupu, et al.
Published: (2025)
by: Liang, Yupu, et al.
Published: (2025)
Towards Efficient CoT Distillation: Self-Guided Rationale Selector for Better Performance with Fewer Rationales
by: Yan, Jianzhi, et al.
Published: (2025)
by: Yan, Jianzhi, et al.
Published: (2025)
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation
by: Liang, Yupu, et al.
Published: (2025)
by: Liang, Yupu, et al.
Published: (2025)
Improving LLM Abilities in Idiomatic Translation
by: Donthi, Sundesh, et al.
Published: (2024)
by: Donthi, Sundesh, et al.
Published: (2024)
SweetieChat: A Strategy-Enhanced Role-playing Framework for Diverse Scenarios Handling Emotional Support Agent
by: Ye, Jing, et al.
Published: (2024)
by: Ye, Jing, et al.
Published: (2024)
Okay, Let's Do This! Modeling Event Coreference with Generated Rationales and Knowledge Distillation
by: Nath, Abhijnan, et al.
Published: (2024)
by: Nath, Abhijnan, et al.
Published: (2024)
PMMT: Preference Alignment in Multilingual Machine Translation via LLM Distillation
by: Sun, Shuqiao, et al.
Published: (2024)
by: Sun, Shuqiao, et al.
Published: (2024)
Effective Distillation of Table-based Reasoning Ability from LLMs
by: Yang, Bohao, et al.
Published: (2023)
by: Yang, Bohao, et al.
Published: (2023)
Multilingual Non-Autoregressive Machine Translation without Knowledge Distillation
by: Huang, Chenyang, et al.
Published: (2025)
by: Huang, Chenyang, et al.
Published: (2025)
RM-Distiller: Exploiting Generative LLM for Reward Model Distillation
by: Zhou, Hongli, et al.
Published: (2026)
by: Zhou, Hongli, et al.
Published: (2026)
Rationale-guided Prompting for Knowledge-based Visual Question Answering
by: Hu, Zhongjian, et al.
Published: (2024)
by: Hu, Zhongjian, et al.
Published: (2024)
EmoHarbor: Evaluating Personalized Emotional Support by Simulating the User's Internal World
by: Ye, Jing, et al.
Published: (2026)
by: Ye, Jing, et al.
Published: (2026)
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
by: Huang, Jiazhen, et al.
Published: (2026)
by: Huang, Jiazhen, et al.
Published: (2026)
Enhanced Multimodal Aspect-Based Sentiment Analysis by LLM-Generated Rationales
by: Cao, Jun, et al.
Published: (2025)
by: Cao, Jun, et al.
Published: (2025)
Improve LLM-as-a-Judge Ability as a General Ability
by: Yu, Jiachen, et al.
Published: (2025)
by: Yu, Jiachen, et al.
Published: (2025)
Rationales Are Not Silver Bullets: Measuring the Impact of Rationales on Model Performance and Reliability
by: Zhu, Chiwei, et al.
Published: (2025)
by: Zhu, Chiwei, et al.
Published: (2025)
TriSum: Learning Summarization Ability from Large Language Models with Structured Rationale
by: Jiang, Pengcheng, et al.
Published: (2024)
by: Jiang, Pengcheng, et al.
Published: (2024)
TROVE: A Challenge for Fine-Grained Text Provenance via Source Sentence Tracing and Relationship Classification
by: Zhu, Junnan, et al.
Published: (2025)
by: Zhu, Junnan, et al.
Published: (2025)
CITI: Enhancing Tool Utilizing Ability in Large Language Models without Sacrificing General Performance
by: Hao, Yupu, et al.
Published: (2024)
by: Hao, Yupu, et al.
Published: (2024)
Similar Items
-
F-MALLOC: Feed-forward Memory Allocation for Continual Learning in Neural Machine Translation
by: Wu, Junhong, et al.
Published: (2024) -
SimulPL: Aligning Human Preferences in Simultaneous Machine Translation
by: Yu, Donglei, et al.
Published: (2025) -
Hit the Sweet Spot! Span-Level Ensemble for Large Language Models
by: Xu, Yangyifan, et al.
Published: (2024) -
LADM: Long-context Training Data Selection with Attention-based Dependency Measurement for LLMs
by: Chen, Jianghao, et al.
Published: (2025) -
Parallel Scaling Law: Unveiling Reasoning Generalization through A Cross-Linguistic Perspective
by: Yang, Wen, et al.
Published: (2025)