PEGRL: Improving Machine Translation by Post-Editing Guided Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shen, Yunzhi, Zhou, Hao, Huang, Xin, Han, Xue, Feng, Junlan, Huang, Shujian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-Improving Multilingual Long Reasoning via Translation-Reasoning Integrated Training
von: Liu, Junxiao, et al.
Veröffentlicht: (2026)
von: Liu, Junxiao, et al.
Veröffentlicht: (2026)
Align to the Pivot: Dual Alignment with Self-Feedback for Multilingual Math Reasoning
von: Zhao, Chunxu, et al.
Veröffentlicht: (2026)
von: Zhao, Chunxu, et al.
Veröffentlicht: (2026)
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
von: Zhou, Hao, et al.
Veröffentlicht: (2024)
von: Zhou, Hao, et al.
Veröffentlicht: (2024)
Understanding LLMs' Cross-Lingual Context Retrieval: How Good It Is And Where It Comes From
von: Gao, Changjiang, et al.
Veröffentlicht: (2025)
von: Gao, Changjiang, et al.
Veröffentlicht: (2025)
Large Language Models Are Cross-Lingual Knowledge-Free Reasoners
von: Hu, Peng, et al.
Veröffentlicht: (2024)
von: Hu, Peng, et al.
Veröffentlicht: (2024)
Investigating and Scaling up Code-Switching for Multilingual Language Model Pre-Training
von: Wang, Zhijun, et al.
Veröffentlicht: (2025)
von: Wang, Zhijun, et al.
Veröffentlicht: (2025)
Getting More from Less: Large Language Models are Good Spontaneous Multilingual Learners
von: Zhang, Shimao, et al.
Veröffentlicht: (2024)
von: Zhang, Shimao, et al.
Veröffentlicht: (2024)
Extend Adversarial Policy Against Neural Machine Translation via Unknown Token
von: Zou, Wei, et al.
Veröffentlicht: (2025)
von: Zou, Wei, et al.
Veröffentlicht: (2025)
Alleviating Distribution Shift in Synthetic Data for Machine Translation Quality Estimation
von: Geng, Xiang, et al.
Veröffentlicht: (2025)
von: Geng, Xiang, et al.
Veröffentlicht: (2025)
Eliciting the Translation Ability of Large Language Models via Multilingual Finetuning with Translation Instructions
von: Li, Jiahuan, et al.
Veröffentlicht: (2023)
von: Li, Jiahuan, et al.
Veröffentlicht: (2023)
Self-Evolution Knowledge Distillation for LLM-based Machine Translation
von: Song, Yuncheng, et al.
Veröffentlicht: (2024)
von: Song, Yuncheng, et al.
Veröffentlicht: (2024)
GRRM: Group Relative Reward Modeling for Machine Translation
von: Yang, Sen, et al.
Veröffentlicht: (2026)
von: Yang, Sen, et al.
Veröffentlicht: (2026)
MT-PATCHER: Selective and Extendable Knowledge Distillation from Large Language Models for Machine Translation
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
TAPO: Translation Augmented Policy Optimization for Multilingual Mathematical Reasoning
von: Huang, Xu, et al.
Veröffentlicht: (2026)
von: Huang, Xu, et al.
Veröffentlicht: (2026)
Lost in the Source Language: How Large Language Models Evaluate the Quality of Machine Translation
von: Huang, Xu, et al.
Veröffentlicht: (2024)
von: Huang, Xu, et al.
Veröffentlicht: (2024)
TranslationCorrect: A Unified Framework for Machine Translation Post-Editing with Predictive Error Assistance
von: Wasti, Syed Mekael, et al.
Veröffentlicht: (2025)
von: Wasti, Syed Mekael, et al.
Veröffentlicht: (2025)
EDT: Improving Large Language Models' Generation by Entropy-based Dynamic Temperature Sampling
von: Zhang, Shimao, et al.
Veröffentlicht: (2024)
von: Zhang, Shimao, et al.
Veröffentlicht: (2024)
Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis
von: Zhu, Wenhao, et al.
Veröffentlicht: (2023)
von: Zhu, Wenhao, et al.
Veröffentlicht: (2023)
Hindsight Quality Prediction Experiments in Multi-Candidate Human-Post-Edited Machine Translation
von: Marmonier, Malik, et al.
Veröffentlicht: (2026)
von: Marmonier, Malik, et al.
Veröffentlicht: (2026)
RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation
von: Li, Tianjiao, et al.
Veröffentlicht: (2025)
von: Li, Tianjiao, et al.
Veröffentlicht: (2025)
Multilingual Non-Autoregressive Machine Translation without Knowledge Distillation
von: Huang, Chenyang, et al.
Veröffentlicht: (2025)
von: Huang, Chenyang, et al.
Veröffentlicht: (2025)
LLaMAX2: Your Translation-Enhanced Model also Performs Well in Reasoning
von: Gao, Changjiang, et al.
Veröffentlicht: (2025)
von: Gao, Changjiang, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Edit-Based Non-Autoregressive Neural Machine Translation
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint
von: Chen, Zhipeng, et al.
Veröffentlicht: (2024)
von: Chen, Zhipeng, et al.
Veröffentlicht: (2024)
Question Translation Training for Better Multilingual Reasoning
von: Zhu, Wenhao, et al.
Veröffentlicht: (2024)
von: Zhu, Wenhao, et al.
Veröffentlicht: (2024)
Integrating Multi-scale Contextualized Information for Byte-based Neural Machine Translation
von: Huang, Langlin, et al.
Veröffentlicht: (2024)
von: Huang, Langlin, et al.
Veröffentlicht: (2024)
Asymmetric Conflict and Synergy in Post-training for LLM-based Multilingual Machine Translation
von: Zheng, Tong, et al.
Veröffentlicht: (2025)
von: Zheng, Tong, et al.
Veröffentlicht: (2025)
Ara-HOPE: Human-Centric Post-Editing Evaluation for Dialectal Arabic to Modern Standard Arabic Translation
von: Alabdullah, Abdullah, et al.
Veröffentlicht: (2025)
von: Alabdullah, Abdullah, et al.
Veröffentlicht: (2025)
Improving Low-Resource Machine Translation via Round-Trip Reinforcement Learning
von: Attia, Ahmed, et al.
Veröffentlicht: (2026)
von: Attia, Ahmed, et al.
Veröffentlicht: (2026)
ExpLang: Improved Exploration and Exploitation in LLM Reasoning with On-Policy Thinking Language Selection
von: Gao, Changjiang, et al.
Veröffentlicht: (2026)
von: Gao, Changjiang, et al.
Veröffentlicht: (2026)
A Data-Efficient Path to Multilingual LLMs: Language Expansion via Post-training PARAM$Δ$ Integration into Upcycled MoE
von: Zhou, Hao, et al.
Veröffentlicht: (2026)
von: Zhou, Hao, et al.
Veröffentlicht: (2026)
The Power of Question Translation Training in Multilingual Reasoning: Broadened Scope and Deepened Insights
von: Zhu, Wenhao, et al.
Veröffentlicht: (2024)
von: Zhu, Wenhao, et al.
Veröffentlicht: (2024)
EnAnchored-X2X: English-Anchored Optimization for Many-to-Many Translation
von: Yang, Sen, et al.
Veröffentlicht: (2025)
von: Yang, Sen, et al.
Veröffentlicht: (2025)
Guiding Large Language Models to Post-Edit Machine Translation with Error Annotations
von: Ki, Dayeon, et al.
Veröffentlicht: (2024)
von: Ki, Dayeon, et al.
Veröffentlicht: (2024)
PRIM: Towards Practical In-Image Multilingual Machine Translation
von: Tian, Yanzhi, et al.
Veröffentlicht: (2025)
von: Tian, Yanzhi, et al.
Veröffentlicht: (2025)
Trans-Zero: Self-Play Incentivizes Large Language Models for Multilingual Translation Without Parallel Data
von: Zou, Wei, et al.
Veröffentlicht: (2025)
von: Zou, Wei, et al.
Veröffentlicht: (2025)
MoCE: Adaptive Mixture of Contextualization Experts for Byte-based Neural Machine Translation
von: Huang, Langlin, et al.
Veröffentlicht: (2024)
von: Huang, Langlin, et al.
Veröffentlicht: (2024)
Efficient Machine Translation Corpus Generation: Integrating Human-in-the-Loop Post-Editing with Large Language Models
von: Yuksel, Kamer Ali, et al.
Veröffentlicht: (2025)
von: Yuksel, Kamer Ali, et al.
Veröffentlicht: (2025)
Formality is Favored: Unraveling the Learning Preferences of Large Language Models on Data with Conflicting Knowledge
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
von: Li, Jiahuan, et al.
Veröffentlicht: (2024)
Faster Machine Translation Ensembling with Reinforcement Learning and Competitive Correction
von: Prasad, Kritarth, et al.
Veröffentlicht: (2025)
von: Prasad, Kritarth, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Self-Improving Multilingual Long Reasoning via Translation-Reasoning Integrated Training
von: Liu, Junxiao, et al.
Veröffentlicht: (2026) -
Align to the Pivot: Dual Alignment with Self-Feedback for Multilingual Math Reasoning
von: Zhao, Chunxu, et al.
Veröffentlicht: (2026) -
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
von: Zhou, Hao, et al.
Veröffentlicht: (2024) -
Understanding LLMs' Cross-Lingual Context Retrieval: How Good It Is And Where It Comes From
von: Gao, Changjiang, et al.
Veröffentlicht: (2025) -
Large Language Models Are Cross-Lingual Knowledge-Free Reasoners
von: Hu, Peng, et al.
Veröffentlicht: (2024)