RL from Teacher-Model Refinement: Gradual Imitation Learning for Machine Translation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Lee, Dongyub Jude, Ye, Zhenyi, He, Pengcheng |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Iterative Translation Refinement with Large Language Models
par: Chen, Pinzhen, et autres
Publié: (2023)
par: Chen, Pinzhen, et autres
Publié: (2023)
TEaR: Improving LLM-based Machine Translation with Systematic Self-Refinement
par: Feng, Zhaopeng, et autres
Publié: (2024)
par: Feng, Zhaopeng, et autres
Publié: (2024)
Translate-and-Revise: Boosting Large Language Models for Constrained Translation
par: Huang, Pengcheng, et autres
Publié: (2024)
par: Huang, Pengcheng, et autres
Publié: (2024)
GRATH: Gradual Self-Truthifying for Large Language Models
par: Chen, Weixin, et autres
Publié: (2024)
par: Chen, Weixin, et autres
Publié: (2024)
Translation of Multifaceted Data without Re-Training of Machine Translation Systems
par: Moon, Hyeonseok, et autres
Publié: (2024)
par: Moon, Hyeonseok, et autres
Publié: (2024)
Adaptive Few-shot Prompting for Machine Translation with Pre-trained Language Models
par: Tang, Lei, et autres
Publié: (2025)
par: Tang, Lei, et autres
Publié: (2025)
On the Shortcut Learning in Multilingual Neural Machine Translation
par: Wang, Wenxuan, et autres
Publié: (2024)
par: Wang, Wenxuan, et autres
Publié: (2024)
Planning without Search: Refining Frontier LLMs with Offline Goal-Conditioned RL
par: Hong, Joey, et autres
Publié: (2025)
par: Hong, Joey, et autres
Publié: (2025)
On Instruction-Finetuning Neural Machine Translation Models
par: Raunak, Vikas, et autres
Publié: (2024)
par: Raunak, Vikas, et autres
Publié: (2024)
Contextual Refinement of Translations: Large Language Models for Sentence and Document-Level Post-Editing
par: Koneru, Sai, et autres
Publié: (2023)
par: Koneru, Sai, et autres
Publié: (2023)
QE-EBM: Using Quality Estimators as Energy Loss for Machine Translation
par: Yoo, Gahyun, et autres
Publié: (2024)
par: Yoo, Gahyun, et autres
Publié: (2024)
Improving Machine Translation with Human Feedback: An Exploration of Quality Estimation as a Reward Model
par: He, Zhiwei, et autres
Publié: (2024)
par: He, Zhiwei, et autres
Publié: (2024)
Speaking at the Right Level: Literacy-Controlled Counterspeech Generation with RAG-RL
par: Song, Xiaoying, et autres
Publié: (2025)
par: Song, Xiaoying, et autres
Publié: (2025)
A Single Model Ensemble Framework for Neural Machine Translation using Pivot Translation
par: Oh, Seokjin, et autres
Publié: (2025)
par: Oh, Seokjin, et autres
Publié: (2025)
Language Model Distillation: A Temporal Difference Imitation Learning Perspective
par: Yu, Zishun, et autres
Publié: (2025)
par: Yu, Zishun, et autres
Publié: (2025)
Model Editing at Scale leads to Gradual and Catastrophic Forgetting
par: Gupta, Akshat, et autres
Publié: (2024)
par: Gupta, Akshat, et autres
Publié: (2024)
CANTONMT: Investigating Back-Translation and Model-Switch Mechanisms for Cantonese-English Neural Machine Translation
par: Hong, Kung Yin, et autres
Publié: (2024)
par: Hong, Kung Yin, et autres
Publié: (2024)
Gradual Forgetting: Logarithmic Compression for Extending Transformer Context Windows
par: Dickson, Billy, et autres
Publié: (2025)
par: Dickson, Billy, et autres
Publié: (2025)
Gradually Excavating External Knowledge for Implicit Complex Question Answering
par: Liu, Chang, et autres
Publié: (2026)
par: Liu, Chang, et autres
Publié: (2026)
Sociotechnical Effects of Machine Translation
par: Moorkens, Joss, et autres
Publié: (2025)
par: Moorkens, Joss, et autres
Publié: (2025)
Learning When to Translate for Multilingual Reasoning
par: Kang, Deokhyung, et autres
Publié: (2026)
par: Kang, Deokhyung, et autres
Publié: (2026)
PartialFormer: Modeling Part Instead of Whole for Machine Translation
par: Zheng, Tong, et autres
Publié: (2023)
par: Zheng, Tong, et autres
Publié: (2023)
Unlocking Reasoning Capability on Machine Translation in Large Language Models
par: Rajaee, Sara, et autres
Publié: (2026)
par: Rajaee, Sara, et autres
Publié: (2026)
NeoAMT: Neologism-Aware Agentic Machine Translation with Reinforcement Learning
par: Miao, Zhongtao, et autres
Publié: (2026)
par: Miao, Zhongtao, et autres
Publié: (2026)
Machine Translation Advancements of Low-Resource Indian Languages by Transfer Learning
par: Wei, Bin, et autres
Publié: (2024)
par: Wei, Bin, et autres
Publié: (2024)
Contrastive Token Learning with Similarity Decay for Repetition Suppression in Machine Translation
par: Dai, Huangyu, et autres
Publié: (2024)
par: Dai, Huangyu, et autres
Publié: (2024)
Critique-RL: Training Language Models for Critiquing through Two-Stage Reinforcement Learning
par: Xi, Zhiheng, et autres
Publié: (2025)
par: Xi, Zhiheng, et autres
Publié: (2025)
s3: You Don't Need That Much Data to Train a Search Agent via RL
par: Jiang, Pengcheng, et autres
Publié: (2025)
par: Jiang, Pengcheng, et autres
Publié: (2025)
Towards Cross-Cultural Machine Translation with Retrieval-Augmented Generation from Multilingual Knowledge Graphs
par: Conia, Simone, et autres
Publié: (2024)
par: Conia, Simone, et autres
Publié: (2024)
Transformer-Encoder Trees for Efficient Multilingual Machine Translation and Speech Translation
par: Guan, Yiwen, et autres
Publié: (2025)
par: Guan, Yiwen, et autres
Publié: (2025)
Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation
par: Lan, Zhibin, et autres
Publié: (2024)
par: Lan, Zhibin, et autres
Publié: (2024)
Scaling Laws of Decoder-Only Models on the Multilingual Machine Translation Task
par: Caillaut, Gaëtan, et autres
Publié: (2024)
par: Caillaut, Gaëtan, et autres
Publié: (2024)
Attention Mechanism and Context Modeling System for Text Mining Machine Translation
par: Zhang, Yuwei, et autres
Publié: (2024)
par: Zhang, Yuwei, et autres
Publié: (2024)
Generating Gender Alternatives in Machine Translation
par: Garg, Sarthak, et autres
Publié: (2024)
par: Garg, Sarthak, et autres
Publié: (2024)
Word Alignment as Preference for Machine Translation
par: Wu, Qiyu, et autres
Publié: (2024)
par: Wu, Qiyu, et autres
Publié: (2024)
Interplay of Machine Translation, Diacritics, and Diacritization
par: Chen, Wei-Rui, et autres
Publié: (2024)
par: Chen, Wei-Rui, et autres
Publié: (2024)
Glancing Future for Simultaneous Machine Translation
par: Guo, Shoutao, et autres
Publié: (2023)
par: Guo, Shoutao, et autres
Publié: (2023)
Neural Machine Translation of Clinical Text: An Empirical Investigation into Multilingual Pre-Trained Language Models and Transfer-Learning
par: Han, Lifeng, et autres
Publié: (2023)
par: Han, Lifeng, et autres
Publié: (2023)
Improving Low-Resource Machine Translation via Round-Trip Reinforcement Learning
par: Attia, Ahmed, et autres
Publié: (2026)
par: Attia, Ahmed, et autres
Publié: (2026)
OPSD Compresses What RLVR Teaches: A Post-RL Compaction Stage for Reasoning Models
par: Kim, Jaehoon, et autres
Publié: (2026)
par: Kim, Jaehoon, et autres
Publié: (2026)
Documents similaires
-
Iterative Translation Refinement with Large Language Models
par: Chen, Pinzhen, et autres
Publié: (2023) -
TEaR: Improving LLM-based Machine Translation with Systematic Self-Refinement
par: Feng, Zhaopeng, et autres
Publié: (2024) -
Translate-and-Revise: Boosting Large Language Models for Constrained Translation
par: Huang, Pengcheng, et autres
Publié: (2024) -
GRATH: Gradual Self-Truthifying for Large Language Models
par: Chen, Weixin, et autres
Publié: (2024) -
Translation of Multifaceted Data without Re-Training of Machine Translation Systems
par: Moon, Hyeonseok, et autres
Publié: (2024)