Theoretical Guarantees for Minimum Bayes Risk Decoding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ichihara, Yuki, Jinnai, Yuu, Ariu, Kaito, Morimura, Tetsuro, Uchibe, Eiji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Model-Based Minimum Bayes Risk Decoding for Text Generation
von: Jinnai, Yuu, et al.
Veröffentlicht: (2023)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2023)
Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
Evaluation of Best-of-N Sampling Strategies for Language Model Alignment
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025)
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025)
Hyperparameter-Free Approach for Faster Minimum Bayes Risk Decoding
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
Consensus Group Relative Policy Optimization for Text Generation
von: Ichihara, Yuki, et al.
Veröffentlicht: (2026)
von: Ichihara, Yuki, et al.
Veröffentlicht: (2026)
On the True Distribution Approximation of Minimum Bayes-Risk Decoding
von: Ohashi, Atsumoto, et al.
Veröffentlicht: (2024)
von: Ohashi, Atsumoto, et al.
Veröffentlicht: (2024)
Generating Diverse and High-Quality Texts by Minimum Bayes Risk Decoding
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
Filtered Direct Preference Optimization
von: Morimura, Tetsuro, et al.
Veröffentlicht: (2024)
von: Morimura, Tetsuro, et al.
Veröffentlicht: (2024)
Re-evaluating Minimum Bayes Risk Decoding for Automatic Speech Recognition
von: Jinnai, Yuu
Veröffentlicht: (2025)
von: Jinnai, Yuu
Veröffentlicht: (2025)
Document-Level Text Generation with Minimum Bayes Risk Decoding using Optimal Transport
von: Jinnai, Yuu
Veröffentlicht: (2025)
von: Jinnai, Yuu
Veröffentlicht: (2025)
MO-GRPO: Mitigating Reward Hacking of Group Relative Policy Optimization on Multi-Objective Problems
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025)
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025)
Does Cross-Cultural Alignment Change the Commonsense Morality of Language Models?
von: Jinnai, Yuu
Veröffentlicht: (2024)
von: Jinnai, Yuu
Veröffentlicht: (2024)
Do Large Language Models Know Folktales? A Case Study of Yokai in Japanese Folktales
von: Tsutsumi, Ayuto, et al.
Veröffentlicht: (2025)
von: Tsutsumi, Ayuto, et al.
Veröffentlicht: (2025)
Annotation-Efficient Language Model Alignment via Diverse and Representative Response Texts
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024)
Structure-Conditional Minimum Bayes Risk Decoding
von: Eikema, Bryan, et al.
Veröffentlicht: (2025)
von: Eikema, Bryan, et al.
Veröffentlicht: (2025)
Improving Minimum Bayes Risk Decoding with Multi-Prompt
von: Heineman, David, et al.
Veröffentlicht: (2024)
von: Heineman, David, et al.
Veröffentlicht: (2024)
Centroid-Based Efficient Minimum Bayes Risk Decoding
von: Deguchi, Hiroyuki, et al.
Veröffentlicht: (2024)
von: Deguchi, Hiroyuki, et al.
Veröffentlicht: (2024)
mbrs: A Library for Minimum Bayes Risk Decoding
von: Deguchi, Hiroyuki, et al.
Veröffentlicht: (2024)
von: Deguchi, Hiroyuki, et al.
Veröffentlicht: (2024)
Linear-time Minimum Bayes Risk Decoding with Reference Aggregation
von: Vamvas, Jannis, et al.
Veröffentlicht: (2024)
von: Vamvas, Jannis, et al.
Veröffentlicht: (2024)
Uncertainty-Aware Decoding with Minimum Bayes Risk
von: Daheim, Nico, et al.
Veröffentlicht: (2025)
von: Daheim, Nico, et al.
Veröffentlicht: (2025)
Mitigating Metric Bias in Minimum Bayes Risk Decoding
von: Kovacs, Geza, et al.
Veröffentlicht: (2024)
von: Kovacs, Geza, et al.
Veröffentlicht: (2024)
Agreement-Constrained Probabilistic Minimum Bayes Risk Decoding
von: Natsumi, Koki, et al.
Veröffentlicht: (2025)
von: Natsumi, Koki, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Edit-Based Non-Autoregressive Neural Machine Translation
von: Wang, Hao, et al.
Veröffentlicht: (2024)
von: Wang, Hao, et al.
Veröffentlicht: (2024)
Direct Preference Optimization for Neural Machine Translation with Minimum Bayes Risk Decoding
von: Yang, Guangyu, et al.
Veröffentlicht: (2023)
von: Yang, Guangyu, et al.
Veröffentlicht: (2023)
Enhancing Factuality through Consensus and Consistency in Summarization Using Minimum Bayes Risk Decoding
von: Soetedjo, Riza Setiawan, et al.
Veröffentlicht: (2026)
von: Soetedjo, Riza Setiawan, et al.
Veröffentlicht: (2026)
Return-Aligned Decision Transformer
von: Tanaka, Tsunehiko, et al.
Veröffentlicht: (2024)
von: Tanaka, Tsunehiko, et al.
Veröffentlicht: (2024)
Efficient Minimum Bayes Risk Decoding using Low-Rank Matrix Completion Algorithms
von: Trabelsi, Firas, et al.
Veröffentlicht: (2024)
von: Trabelsi, Firas, et al.
Veröffentlicht: (2024)
Diversity Explains Inference Scaling Laws: Through a Case Study of Minimum Bayes Risk Decoding
von: Kamigaito, Hidetaka, et al.
Veröffentlicht: (2024)
von: Kamigaito, Hidetaka, et al.
Veröffentlicht: (2024)
Unveiling the Power of Source: Source-based Minimum Bayes Risk Decoding for Neural Machine Translation
von: Lyu, Boxuan, et al.
Veröffentlicht: (2024)
von: Lyu, Boxuan, et al.
Veröffentlicht: (2024)
Chasing COMET: Leveraging Minimum Bayes Risk Decoding for Self-Improving Machine Translation
von: Guttmann, Kamil, et al.
Veröffentlicht: (2024)
von: Guttmann, Kamil, et al.
Veröffentlicht: (2024)
Minimum Bayes Risk Decoding for Error Span Detection in Reference-Free Automatic Machine Translation Evaluation
von: Lyu, Boxuan, et al.
Veröffentlicht: (2025)
von: Lyu, Boxuan, et al.
Veröffentlicht: (2025)
Better Instruction-Following Through Minimum Bayes Risk
von: Wu, Ian, et al.
Veröffentlicht: (2024)
von: Wu, Ian, et al.
Veröffentlicht: (2024)
Uncertainty Quantification for LLMs through Minimum Bayes Risk: Bridging Confidence and Consistency
von: Vashurin, Roman, et al.
Veröffentlicht: (2025)
von: Vashurin, Roman, et al.
Veröffentlicht: (2025)
Reliable Chain-of-Thought via Prefix Consistency
von: Iwase, Naoto, et al.
Veröffentlicht: (2026)
von: Iwase, Naoto, et al.
Veröffentlicht: (2026)
Reward-Punishment Reinforcement Learning with Maximum Entropy
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)
von: Wang, Jiexin, et al.
Veröffentlicht: (2024)
Guaranteeing Knowledge Integration with Joint Decoding for Retrieval-Augmented Generation
von: Zhao, Zhengyi, et al.
Veröffentlicht: (2026)
von: Zhao, Zhengyi, et al.
Veröffentlicht: (2026)
When Choices Become Risks: Safety Failures of Large Language Models under Multiple-Choice Constraints
von: Chen, Yuheng, et al.
Veröffentlicht: (2026)
von: Chen, Yuheng, et al.
Veröffentlicht: (2026)
Last Iterate Convergence in Monotone Mean Field Games
von: Isobe, Noboru, et al.
Veröffentlicht: (2024)
von: Isobe, Noboru, et al.
Veröffentlicht: (2024)
Learning from Delayed Feedback in Games via Extra Prediction
von: Fujimoto, Yuma, et al.
Veröffentlicht: (2025)
von: Fujimoto, Yuma, et al.
Veröffentlicht: (2025)
Linear Convergence in Games with Delayed Feedback via Extra Prediction
von: Fujimoto, Yuma, et al.
Veröffentlicht: (2026)
von: Fujimoto, Yuma, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Model-Based Minimum Bayes Risk Decoding for Text Generation
von: Jinnai, Yuu, et al.
Veröffentlicht: (2023) -
Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024) -
Evaluation of Best-of-N Sampling Strategies for Language Model Alignment
von: Ichihara, Yuki, et al.
Veröffentlicht: (2025) -
Hyperparameter-Free Approach for Faster Minimum Bayes Risk Decoding
von: Jinnai, Yuu, et al.
Veröffentlicht: (2024) -
Consensus Group Relative Policy Optimization for Text Generation
von: Ichihara, Yuki, et al.
Veröffentlicht: (2026)