Reinforcement Learning for Edit-Based Non-Autoregressive Neural Machine Translation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Hao, Morimura, Tetsuro, Honda, Ukyo, Kawahara, Daisuke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the True Distribution Approximation of Minimum Bayes-Risk Decoding
von: Ohashi, Atsumoto, et al.
Veröffentlicht: (2024)
von: Ohashi, Atsumoto, et al.
Veröffentlicht: (2024)
Generating Diverse and High-Quality Texts by Minimum Bayes Risk Decoding
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
Model-Based Minimum Bayes Risk Decoding for Text Generation
von: Jinnai, Yuu, et al.
Veröffentlicht: (2023)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2023)
Exploring Explanations Improves the Robustness of In-Context Learning
von: Honda, Ukyo, et al.
Veröffentlicht: (2025)
von: Honda, Ukyo, et al.
Veröffentlicht: (2025)
Toward LLMs Beyond English-Centric Development
von: Takase, Sho, et al.
Veröffentlicht: (2026)
von: Takase, Sho, et al.
Veröffentlicht: (2026)
Distilling Many-Shot In-Context Learning into a Cheat Sheet
von: Honda, Ukyo, et al.
Veröffentlicht: (2025)
von: Honda, Ukyo, et al.
Veröffentlicht: (2025)
Annotation-Efficient Language Model Alignment via Diverse and Representative Response Texts
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
Kanbun-LM: Reading and Translating Classical Chinese in Japanese Methods by Language Models
von: Wang, Hao, et al.
Veröffentlicht: (2023)
von: Wang, Hao, et al.
Veröffentlicht: (2023)
Does Self-Consistency Improve the Recall of Encyclopedic Knowledge?
von: Hoshino, Sho, et al.
Veröffentlicht: (2026)
von: Hoshino, Sho, et al.
Veröffentlicht: (2026)
Revisiting the Capacity Gap in Chain-of-Thought Distillation from a Practical Perspective
von: Kajitsuka, Tokio, et al.
Veröffentlicht: (2026)
von: Kajitsuka, Tokio, et al.
Veröffentlicht: (2026)
Multilingual Non-Autoregressive Machine Translation without Knowledge Distillation
von: Huang, Chenyang, et al.
Veröffentlicht: (2025)
von: Huang, Chenyang, et al.
Veröffentlicht: (2025)
Exploring the Relationship Between Diversity and Quality in Ad Text Generation
von: Aoki, Yoichi, et al.
Veröffentlicht: (2025)
von: Aoki, Yoichi, et al.
Veröffentlicht: (2025)
Not Eliminate but Aggregate: Post-Hoc Control over Mixture-of-Experts to Address Shortcut Shifts in Natural Language Understanding
von: Honda, Ukyo, et al.
Veröffentlicht: (2024)
von: Honda, Ukyo, et al.
Veröffentlicht: (2024)
Shared Latent Space by Both Languages in Non-Autoregressive Neural Machine Translation
von: Heo, DongNyeong, et al.
Veröffentlicht: (2023)
von: Heo, DongNyeong, et al.
Veröffentlicht: (2023)
On the Information Redundancy in Non-Autoregressive Translation
von: Wang, Zhihao, et al.
Veröffentlicht: (2024)
von: Wang, Zhihao, et al.
Veröffentlicht: (2024)
Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
A Single Linear Layer Yields Task-Adapted Low-Rank Matrices
von: Kim, Hwichan, et al.
Veröffentlicht: (2024)
von: Kim, Hwichan, et al.
Veröffentlicht: (2024)
Theoretical Guarantees for Minimum Bayes Risk Decoding
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025)
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025)
FaithCAMERA: Construction of a Faithful Dataset for Ad Text Generation
von: Kato, Akihiko, et al.
Veröffentlicht: (2024)
von: Kato, Akihiko, et al.
Veröffentlicht: (2024)
PEGRL: Improving Machine Translation by Post-Editing Guided Reinforcement Learning
von: Shen, Yunzhi, et al.
Veröffentlicht: (2026)
von: Shen, Yunzhi, et al.
Veröffentlicht: (2026)
Filtered Direct Preference Optimization
von: Morimura, Tetsuro, et al.
Veröffentlicht: (2024)
von: Morimura, Tetsuro, et al.
Veröffentlicht: (2024)
Should We Respect LLMs? A Cross-Lingual Study on the Influence of Prompt Politeness on LLM Performance
von: Yin, Ziqi, et al.
Veröffentlicht: (2024)
von: Yin, Ziqi, et al.
Veröffentlicht: (2024)
RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation
von: Li, Tianjiao, et al.
Veröffentlicht: (2025)
von: Li, Tianjiao, et al.
Veröffentlicht: (2025)
Learning to Translate Ambiguous Terminology by Preference Optimization on Post-Edits
von: Berger, Nathaniel, et al.
Veröffentlicht: (2025)
von: Berger, Nathaniel, et al.
Veröffentlicht: (2025)
Evaluation of Best-of-N Sampling Strategies for Language Model Alignment
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025)
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025)
Leveraging Diverse Modeling Contexts with Collaborating Learning for Neural Machine Translation
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
von: Liao, Yusheng, et al.
Veröffentlicht: (2024)
On the Shortcut Learning in Multilingual Neural Machine Translation
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2024)
HiEdit: Lifelong Model Editing with Hierarchical Reinforcement Learning
von: Wang, Yangfan, et al.
Veröffentlicht: (2026)
von: Wang, Yangfan, et al.
Veröffentlicht: (2026)
Evaluating Multimodal Large Language Models on Vertically Written Japanese Text
von: Sasagawa, Keito, et al.
Veröffentlicht: (2025)
von: Sasagawa, Keito, et al.
Veröffentlicht: (2025)
Faster Machine Translation Ensembling with Reinforcement Learning and Competitive Correction
von: Prasad, Kritarth, et al.
Veröffentlicht: (2025)
von: Prasad, Kritarth, et al.
Veröffentlicht: (2025)
Structured Document Translation via Format Reinforcement Learning
von: Song, Haiyue, et al.
Veröffentlicht: (2025)
von: Song, Haiyue, et al.
Veröffentlicht: (2025)
Guiding Large Language Models to Post-Edit Machine Translation with Error Annotations
von: Ki, Dayeon, et al.
Veröffentlicht: (2024)
von: Ki, Dayeon, et al.
Veröffentlicht: (2024)
Non-Fluent Synthetic Target-Language Data Improve Neural Machine Translation
von: Sánchez-Cartagena, Víctor M., et al.
Veröffentlicht: (2024)
von: Sánchez-Cartagena, Víctor M., et al.
Veröffentlicht: (2024)
Minority-Aware Satisfaction Estimation in Dialogue Systems via Preference-Adaptive Reinforcement Learning
von: Fu, Yahui, et al.
Veröffentlicht: (2025)
von: Fu, Yahui, et al.
Veröffentlicht: (2025)
EditGRPO: Reinforcement Learning with Post-Rollout Edits for Clinically Accurate Chest X-Ray Report Generation
von: Zhang, Kai, et al.
Veröffentlicht: (2025)
von: Zhang, Kai, et al.
Veröffentlicht: (2025)
Traveling Across Languages: Benchmarking Cross-Lingual Consistency in Multimodal LLMs
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
On Temperature-Constrained Non-Deterministic Machine Translation: Potential and Evaluation
von: Wang, Weichuan, et al.
Veröffentlicht: (2026)
von: Wang, Weichuan, et al.
Veröffentlicht: (2026)
The Box is in the Pen: Evaluating Commonsense Reasoning in Neural Machine Translation
von: He, Jie, et al.
Veröffentlicht: (2025)
von: He, Jie, et al.
Veröffentlicht: (2025)
Evaluating Structural Generalization in Neural Machine Translation
von: Kumon, Ryoma, et al.
Veröffentlicht: (2024)
von: Kumon, Ryoma, et al.
Veröffentlicht: (2024)
Beyond Scalar Scores: Reinforcement Learning for Error-Aware Quality Estimation of Machine Translation
von: Sindhujan, Archchana, et al.
Veröffentlicht: (2026)
von: Sindhujan, Archchana, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
On the True Distribution Approximation of Minimum Bayes-Risk Decoding
von: Ohashi, Atsumoto, et al.
Veröffentlicht: (2024) -
Generating Diverse and High-Quality Texts by Minimum Bayes Risk Decoding
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024) -
Model-Based Minimum Bayes Risk Decoding for Text Generation
von: Jinnai, Yuu, et al.
Veröffentlicht: (2023) -
Exploring Explanations Improves the Robustness of In-Context Learning
von: Honda, Ukyo, et al.
Veröffentlicht: (2025) -
Toward LLMs Beyond English-Centric Development
von: Takase, Sho, et al.
Veröffentlicht: (2026)