Model-Based Minimum Bayes Risk Decoding for Text Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Jinnai, Yuu, Morimura, Tetsuro, Honda, Ukyo, Ariu, Kaito, Abe, Kenshi |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment
by: Jinnai, Yuu, et al.
Published: (2024)
by: Jinnai, Yuu, et al.
Published: (2024)
Generating Diverse and High-Quality Texts by Minimum Bayes Risk Decoding
by: Jinnai, Yuu, et al.
Published: (2024)
by: Jinnai, Yuu, et al.
Published: (2024)
On the True Distribution Approximation of Minimum Bayes-Risk Decoding
by: Ohashi, Atsumoto, et al.
Published: (2024)
by: Ohashi, Atsumoto, et al.
Published: (2024)
Hyperparameter-Free Approach for Faster Minimum Bayes Risk Decoding
by: Jinnai, Yuu, et al.
Published: (2024)
by: Jinnai, Yuu, et al.
Published: (2024)
Filtered Direct Preference Optimization
by: Morimura, Tetsuro, et al.
Published: (2024)
by: Morimura, Tetsuro, et al.
Published: (2024)
Theoretical Guarantees for Minimum Bayes Risk Decoding
by: Ichihara, Yuki, et al.
Published: (2025)
by: Ichihara, Yuki, et al.
Published: (2025)
Annotation-Efficient Language Model Alignment via Diverse and Representative Response Texts
by: Jinnai, Yuu, et al.
Published: (2024)
by: Jinnai, Yuu, et al.
Published: (2024)
Document-Level Text Generation with Minimum Bayes Risk Decoding using Optimal Transport
by: Jinnai, Yuu
Published: (2025)
by: Jinnai, Yuu
Published: (2025)
Evaluation of Best-of-N Sampling Strategies for Language Model Alignment
by: Ichihara, Yuki, et al.
Published: (2025)
by: Ichihara, Yuki, et al.
Published: (2025)
Re-evaluating Minimum Bayes Risk Decoding for Automatic Speech Recognition
by: Jinnai, Yuu
Published: (2025)
by: Jinnai, Yuu
Published: (2025)
Does Cross-Cultural Alignment Change the Commonsense Morality of Language Models?
by: Jinnai, Yuu
Published: (2024)
by: Jinnai, Yuu
Published: (2024)
Last Iterate Convergence in Monotone Mean Field Games
by: Isobe, Noboru, et al.
Published: (2024)
by: Isobe, Noboru, et al.
Published: (2024)
Reinforcement Learning for Edit-Based Non-Autoregressive Neural Machine Translation
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Exploring Explanations Improves the Robustness of In-Context Learning
by: Honda, Ukyo, et al.
Published: (2025)
by: Honda, Ukyo, et al.
Published: (2025)
Policy Gradient Algorithms with Monte Carlo Tree Learning for Non-Markov Decision Processes
by: Morimura, Tetsuro, et al.
Published: (2022)
by: Morimura, Tetsuro, et al.
Published: (2022)
Mitigating Metric Bias in Minimum Bayes Risk Decoding
by: Kovacs, Geza, et al.
Published: (2024)
by: Kovacs, Geza, et al.
Published: (2024)
Uncertainty-Aware Decoding with Minimum Bayes Risk
by: Daheim, Nico, et al.
Published: (2025)
by: Daheim, Nico, et al.
Published: (2025)
Consensus Group Relative Policy Optimization for Text Generation
by: Ichihara, Yuki, et al.
Published: (2026)
by: Ichihara, Yuki, et al.
Published: (2026)
Revisiting the Capacity Gap in Chain-of-Thought Distillation from a Practical Perspective
by: Kajitsuka, Tokio, et al.
Published: (2026)
by: Kajitsuka, Tokio, et al.
Published: (2026)
Agreement-Constrained Probabilistic Minimum Bayes Risk Decoding
by: Natsumi, Koki, et al.
Published: (2025)
by: Natsumi, Koki, et al.
Published: (2025)
Return-Aligned Decision Transformer
by: Tanaka, Tsunehiko, et al.
Published: (2024)
by: Tanaka, Tsunehiko, et al.
Published: (2024)
A Single Linear Layer Yields Task-Adapted Low-Rank Matrices
by: Kim, Hwichan, et al.
Published: (2024)
by: Kim, Hwichan, et al.
Published: (2024)
Unveiling the Power of Source: Source-based Minimum Bayes Risk Decoding for Neural Machine Translation
by: Lyu, Boxuan, et al.
Published: (2024)
by: Lyu, Boxuan, et al.
Published: (2024)
Do Large Language Models Know Folktales? A Case Study of Yokai in Japanese Folktales
by: Tsutsumi, Ayuto, et al.
Published: (2025)
by: Tsutsumi, Ayuto, et al.
Published: (2025)
Chasing COMET: Leveraging Minimum Bayes Risk Decoding for Self-Improving Machine Translation
by: Guttmann, Kamil, et al.
Published: (2024)
by: Guttmann, Kamil, et al.
Published: (2024)
Better Instruction-Following Through Minimum Bayes Risk
by: Wu, Ian, et al.
Published: (2024)
by: Wu, Ian, et al.
Published: (2024)
Memory Asymmetry Creates Heteroclinic Orbits to Nash Equilibrium in Learning in Zero-Sum Games
by: Fujimoto, Yuma, et al.
Published: (2023)
by: Fujimoto, Yuma, et al.
Published: (2023)
Learning from Delayed Feedback in Games via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2025)
by: Fujimoto, Yuma, et al.
Published: (2025)
Linear Convergence in Games with Delayed Feedback via Extra Prediction
by: Fujimoto, Yuma, et al.
Published: (2026)
by: Fujimoto, Yuma, et al.
Published: (2026)
Nash Equilibrium and Learning Dynamics in Three-Player Matching $m$-Action Games
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Synchronization in Learning in Periodic Zero-Sum Games Triggers Divergence from Nash Equilibrium
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Time-Varyingness in Auction Breaks Revenue Equivalence
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Global Behavior of Learning Dynamics in Zero-Sum Games with Memory Asymmetry
by: Fujimoto, Yuma, et al.
Published: (2024)
by: Fujimoto, Yuma, et al.
Published: (2024)
Minimum Bayes Risk Decoding for Error Span Detection in Reference-Free Automatic Machine Translation Evaluation
by: Lyu, Boxuan, et al.
Published: (2025)
by: Lyu, Boxuan, et al.
Published: (2025)
Exploring the Relationship Between Diversity and Quality in Ad Text Generation
by: Aoki, Yoichi, et al.
Published: (2025)
by: Aoki, Yoichi, et al.
Published: (2025)
Boosting Perturbed Gradient Ascent for Last-Iterate Convergence in Games
by: Abe, Kenshi, et al.
Published: (2024)
by: Abe, Kenshi, et al.
Published: (2024)
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders
by: Zhu, Xiaofeng, et al.
Published: (2024)
by: Zhu, Xiaofeng, et al.
Published: (2024)
Adaptively Perturbed Mirror Descent for Learning in Games
by: Abe, Kenshi, et al.
Published: (2023)
by: Abe, Kenshi, et al.
Published: (2023)
Toward LLMs Beyond English-Centric Development
by: Takase, Sho, et al.
Published: (2026)
by: Takase, Sho, et al.
Published: (2026)
Pipelined Decoder for Efficient Context-Aware Text Generation
by: Huang, Zixian, et al.
Published: (2025)
by: Huang, Zixian, et al.
Published: (2025)
Similar Items
-
Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment
by: Jinnai, Yuu, et al.
Published: (2024) -
Generating Diverse and High-Quality Texts by Minimum Bayes Risk Decoding
by: Jinnai, Yuu, et al.
Published: (2024) -
On the True Distribution Approximation of Minimum Bayes-Risk Decoding
by: Ohashi, Atsumoto, et al.
Published: (2024) -
Hyperparameter-Free Approach for Faster Minimum Bayes Risk Decoding
by: Jinnai, Yuu, et al.
Published: (2024) -
Filtered Direct Preference Optimization
by: Morimura, Tetsuro, et al.
Published: (2024)