PEGRL: Improving Machine Translation by Post-Editing Guided Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Shen, Yunzhi, Zhou, Hao, Huang, Xin, Han, Xue, Feng, Junlan, Huang, Shujian |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Self-Improving Multilingual Long Reasoning via Translation-Reasoning Integrated Training
di: Liu, Junxiao, et al.
Pubblicazione: (2026)
di: Liu, Junxiao, et al.
Pubblicazione: (2026)
Align to the Pivot: Dual Alignment with Self-Feedback for Multilingual Math Reasoning
di: Zhao, Chunxu, et al.
Pubblicazione: (2026)
di: Zhao, Chunxu, et al.
Pubblicazione: (2026)
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
di: Zhou, Hao, et al.
Pubblicazione: (2024)
di: Zhou, Hao, et al.
Pubblicazione: (2024)
Understanding LLMs' Cross-Lingual Context Retrieval: How Good It Is And Where It Comes From
di: Gao, Changjiang, et al.
Pubblicazione: (2025)
di: Gao, Changjiang, et al.
Pubblicazione: (2025)
Large Language Models Are Cross-Lingual Knowledge-Free Reasoners
di: Hu, Peng, et al.
Pubblicazione: (2024)
di: Hu, Peng, et al.
Pubblicazione: (2024)
Investigating and Scaling up Code-Switching for Multilingual Language Model Pre-Training
di: Wang, Zhijun, et al.
Pubblicazione: (2025)
di: Wang, Zhijun, et al.
Pubblicazione: (2025)
Getting More from Less: Large Language Models are Good Spontaneous Multilingual Learners
di: Zhang, Shimao, et al.
Pubblicazione: (2024)
di: Zhang, Shimao, et al.
Pubblicazione: (2024)
Extend Adversarial Policy Against Neural Machine Translation via Unknown Token
di: Zou, Wei, et al.
Pubblicazione: (2025)
di: Zou, Wei, et al.
Pubblicazione: (2025)
Alleviating Distribution Shift in Synthetic Data for Machine Translation Quality Estimation
di: Geng, Xiang, et al.
Pubblicazione: (2025)
di: Geng, Xiang, et al.
Pubblicazione: (2025)
Eliciting the Translation Ability of Large Language Models via Multilingual Finetuning with Translation Instructions
di: Li, Jiahuan, et al.
Pubblicazione: (2023)
di: Li, Jiahuan, et al.
Pubblicazione: (2023)
Self-Evolution Knowledge Distillation for LLM-based Machine Translation
di: Song, Yuncheng, et al.
Pubblicazione: (2024)
di: Song, Yuncheng, et al.
Pubblicazione: (2024)
GRRM: Group Relative Reward Modeling for Machine Translation
di: Yang, Sen, et al.
Pubblicazione: (2026)
di: Yang, Sen, et al.
Pubblicazione: (2026)
MT-PATCHER: Selective and Extendable Knowledge Distillation from Large Language Models for Machine Translation
di: Li, Jiahuan, et al.
Pubblicazione: (2024)
di: Li, Jiahuan, et al.
Pubblicazione: (2024)
TAPO: Translation Augmented Policy Optimization for Multilingual Mathematical Reasoning
di: Huang, Xu, et al.
Pubblicazione: (2026)
di: Huang, Xu, et al.
Pubblicazione: (2026)
Lost in the Source Language: How Large Language Models Evaluate the Quality of Machine Translation
di: Huang, Xu, et al.
Pubblicazione: (2024)
di: Huang, Xu, et al.
Pubblicazione: (2024)
TranslationCorrect: A Unified Framework for Machine Translation Post-Editing with Predictive Error Assistance
di: Wasti, Syed Mekael, et al.
Pubblicazione: (2025)
di: Wasti, Syed Mekael, et al.
Pubblicazione: (2025)
EDT: Improving Large Language Models' Generation by Entropy-based Dynamic Temperature Sampling
di: Zhang, Shimao, et al.
Pubblicazione: (2024)
di: Zhang, Shimao, et al.
Pubblicazione: (2024)
Multilingual Machine Translation with Large Language Models: Empirical Results and Analysis
di: Zhu, Wenhao, et al.
Pubblicazione: (2023)
di: Zhu, Wenhao, et al.
Pubblicazione: (2023)
Hindsight Quality Prediction Experiments in Multi-Candidate Human-Post-Edited Machine Translation
di: Marmonier, Malik, et al.
Pubblicazione: (2026)
di: Marmonier, Malik, et al.
Pubblicazione: (2026)
RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation
di: Li, Tianjiao, et al.
Pubblicazione: (2025)
di: Li, Tianjiao, et al.
Pubblicazione: (2025)
Multilingual Non-Autoregressive Machine Translation without Knowledge Distillation
di: Huang, Chenyang, et al.
Pubblicazione: (2025)
di: Huang, Chenyang, et al.
Pubblicazione: (2025)
LLaMAX2: Your Translation-Enhanced Model also Performs Well in Reasoning
di: Gao, Changjiang, et al.
Pubblicazione: (2025)
di: Gao, Changjiang, et al.
Pubblicazione: (2025)
Reinforcement Learning for Edit-Based Non-Autoregressive Neural Machine Translation
di: Wang, Hao, et al.
Pubblicazione: (2024)
di: Wang, Hao, et al.
Pubblicazione: (2024)
Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
Question Translation Training for Better Multilingual Reasoning
di: Zhu, Wenhao, et al.
Pubblicazione: (2024)
di: Zhu, Wenhao, et al.
Pubblicazione: (2024)
Integrating Multi-scale Contextualized Information for Byte-based Neural Machine Translation
di: Huang, Langlin, et al.
Pubblicazione: (2024)
di: Huang, Langlin, et al.
Pubblicazione: (2024)
Asymmetric Conflict and Synergy in Post-training for LLM-based Multilingual Machine Translation
di: Zheng, Tong, et al.
Pubblicazione: (2025)
di: Zheng, Tong, et al.
Pubblicazione: (2025)
Ara-HOPE: Human-Centric Post-Editing Evaluation for Dialectal Arabic to Modern Standard Arabic Translation
di: Alabdullah, Abdullah, et al.
Pubblicazione: (2025)
di: Alabdullah, Abdullah, et al.
Pubblicazione: (2025)
Improving Low-Resource Machine Translation via Round-Trip Reinforcement Learning
di: Attia, Ahmed, et al.
Pubblicazione: (2026)
di: Attia, Ahmed, et al.
Pubblicazione: (2026)
ExpLang: Improved Exploration and Exploitation in LLM Reasoning with On-Policy Thinking Language Selection
di: Gao, Changjiang, et al.
Pubblicazione: (2026)
di: Gao, Changjiang, et al.
Pubblicazione: (2026)
A Data-Efficient Path to Multilingual LLMs: Language Expansion via Post-training PARAM$Δ$ Integration into Upcycled MoE
di: Zhou, Hao, et al.
Pubblicazione: (2026)
di: Zhou, Hao, et al.
Pubblicazione: (2026)
The Power of Question Translation Training in Multilingual Reasoning: Broadened Scope and Deepened Insights
di: Zhu, Wenhao, et al.
Pubblicazione: (2024)
di: Zhu, Wenhao, et al.
Pubblicazione: (2024)
EnAnchored-X2X: English-Anchored Optimization for Many-to-Many Translation
di: Yang, Sen, et al.
Pubblicazione: (2025)
di: Yang, Sen, et al.
Pubblicazione: (2025)
Guiding Large Language Models to Post-Edit Machine Translation with Error Annotations
di: Ki, Dayeon, et al.
Pubblicazione: (2024)
di: Ki, Dayeon, et al.
Pubblicazione: (2024)
PRIM: Towards Practical In-Image Multilingual Machine Translation
di: Tian, Yanzhi, et al.
Pubblicazione: (2025)
di: Tian, Yanzhi, et al.
Pubblicazione: (2025)
Trans-Zero: Self-Play Incentivizes Large Language Models for Multilingual Translation Without Parallel Data
di: Zou, Wei, et al.
Pubblicazione: (2025)
di: Zou, Wei, et al.
Pubblicazione: (2025)
MoCE: Adaptive Mixture of Contextualization Experts for Byte-based Neural Machine Translation
di: Huang, Langlin, et al.
Pubblicazione: (2024)
di: Huang, Langlin, et al.
Pubblicazione: (2024)
Efficient Machine Translation Corpus Generation: Integrating Human-in-the-Loop Post-Editing with Large Language Models
di: Yuksel, Kamer Ali, et al.
Pubblicazione: (2025)
di: Yuksel, Kamer Ali, et al.
Pubblicazione: (2025)
Formality is Favored: Unraveling the Learning Preferences of Large Language Models on Data with Conflicting Knowledge
di: Li, Jiahuan, et al.
Pubblicazione: (2024)
di: Li, Jiahuan, et al.
Pubblicazione: (2024)
Faster Machine Translation Ensembling with Reinforcement Learning and Competitive Correction
di: Prasad, Kritarth, et al.
Pubblicazione: (2025)
di: Prasad, Kritarth, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Self-Improving Multilingual Long Reasoning via Translation-Reasoning Integrated Training
di: Liu, Junxiao, et al.
Pubblicazione: (2026) -
Align to the Pivot: Dual Alignment with Self-Feedback for Multilingual Math Reasoning
di: Zhao, Chunxu, et al.
Pubblicazione: (2026) -
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
di: Zhou, Hao, et al.
Pubblicazione: (2024) -
Understanding LLMs' Cross-Lingual Context Retrieval: How Good It Is And Where It Comes From
di: Gao, Changjiang, et al.
Pubblicazione: (2025) -
Large Language Models Are Cross-Lingual Knowledge-Free Reasoners
di: Hu, Peng, et al.
Pubblicazione: (2024)