Making, not Taking, the Best of N
Fuente:
arXiv
Guardado en:
| Autores principales: | Khairi, Ammar, D'souza, Daniel, Fadaee, Marzieh, Kreutzer, Julia |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
When Life Gives You Samples: The Benefits of Scaling up Inference Compute for Multilingual LLMs
por: Khairi, Ammar, et al.
Publicado: (2025)
por: Khairi, Ammar, et al.
Publicado: (2025)
Treasure Hunt: Real-time Targeting of the Long Tail using Training-Time Markers
por: D'souza, Daniel, et al.
Publicado: (2025)
por: D'souza, Daniel, et al.
Publicado: (2025)
Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
por: Kreutzer, Julia, et al.
Publicado: (2025)
por: Kreutzer, Julia, et al.
Publicado: (2025)
LLM See, LLM Do: Guiding Data Generation to Target Non-Differentiable Objectives
por: Shimabucoro, Luísa, et al.
Publicado: (2024)
por: Shimabucoro, Luísa, et al.
Publicado: (2024)
The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It
por: Yong, Zheng-Xin, et al.
Publicado: (2025)
por: Yong, Zheng-Xin, et al.
Publicado: (2025)
The Art of Asking: Multilingual Prompt Optimization for Synthetic Data
por: Mora, David, et al.
Publicado: (2025)
por: Mora, David, et al.
Publicado: (2025)
Aya Model: An Instruction Finetuned Open-Access Multilingual Language Model
por: Üstün, Ahmet, et al.
Publicado: (2024)
por: Üstün, Ahmet, et al.
Publicado: (2024)
The Multilingual Alignment Prism: Aligning Global and Local Preferences to Reduce Harm
por: Aakanksha, et al.
Publicado: (2024)
por: Aakanksha, et al.
Publicado: (2024)
Tiny Aya: Bridging Scale and Multilingual Depth
por: Salamanca, Alejandro R., et al.
Publicado: (2026)
por: Salamanca, Alejandro R., et al.
Publicado: (2026)
Multilingual Arbitrage: Optimizing Data Pools to Accelerate Multilingual Progress
por: Odumakinde, Ayomide, et al.
Publicado: (2024)
por: Odumakinde, Ayomide, et al.
Publicado: (2024)
A Post-trainer's Guide to Multilingual Training Data: Uncovering Cross-lingual Transfer Dynamics
por: Shimabucoro, Luisa, et al.
Publicado: (2025)
por: Shimabucoro, Luisa, et al.
Publicado: (2025)
Diversify and Conquer: Diversity-Centric Data Selection with Iterative Refinement
por: Yu, Simon, et al.
Publicado: (2024)
por: Yu, Simon, et al.
Publicado: (2024)
NeoBabel: A Multilingual Open Tower for Visual Generation
por: Derakhshani, Mohammad Mahdi, et al.
Publicado: (2025)
por: Derakhshani, Mohammad Mahdi, et al.
Publicado: (2025)
Unlocking Reasoning Capability on Machine Translation in Large Language Models
por: Rajaee, Sara, et al.
Publicado: (2026)
por: Rajaee, Sara, et al.
Publicado: (2026)
Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning
por: Aakanksha, et al.
Publicado: (2024)
por: Aakanksha, et al.
Publicado: (2024)
Verification Limits Code LLM Training
por: Gureja, Srishti, et al.
Publicado: (2025)
por: Gureja, Srishti, et al.
Publicado: (2025)
Ta'keed: The First Generative Fact-Checking System for Arabic Claims
por: Althabiti, Saud, et al.
Publicado: (2024)
por: Althabiti, Saud, et al.
Publicado: (2024)
Automatic WordNet Construction Using Markov Chain Monte Carlo
por: Marzieh Fadaee
Publicado: (2013)
por: Marzieh Fadaee
Publicado: (2013)
To Code, or Not To Code? Exploring Impact of Code in Pre-training
por: Aryabumi, Viraat, et al.
Publicado: (2024)
por: Aryabumi, Viraat, et al.
Publicado: (2024)
The Multilingual Divide and Its Impact on Global AI Safety
por: Peppin, Aidan, et al.
Publicado: (2025)
por: Peppin, Aidan, et al.
Publicado: (2025)
From Tools to Teammates: Evaluating LLMs in Multi-Session Coding Interactions
por: Rakotonirina, Nathanaël Carraz, et al.
Publicado: (2025)
por: Rakotonirina, Nathanaël Carraz, et al.
Publicado: (2025)
One Tokenizer To Rule Them All: Emergent Language Plasticity via Multilingual Tokenizers
por: Abagyan, Diana, et al.
Publicado: (2025)
por: Abagyan, Diana, et al.
Publicado: (2025)
Aya 23: Open Weight Releases to Further Multilingual Progress
por: Aryabumi, Viraat, et al.
Publicado: (2024)
por: Aryabumi, Viraat, et al.
Publicado: (2024)
Best-of-N Jailbreaking
por: Hughes, John, et al.
Publicado: (2024)
por: Hughes, John, et al.
Publicado: (2024)
Fast Best-of-N Decoding via Speculative Rejection
por: Sun, Hanshi, et al.
Publicado: (2024)
por: Sun, Hanshi, et al.
Publicado: (2024)
SLURG: Investigating the Feasibility of Generating Synthetic Online Fallacious Discourse
por: Blanco, Cal, et al.
Publicado: (2025)
por: Blanco, Cal, et al.
Publicado: (2025)
Evaluation of Best-of-N Sampling Strategies for Language Model Alignment
por: Ichihara, Yuki, et al.
Publicado: (2025)
por: Ichihara, Yuki, et al.
Publicado: (2025)
Majority of the Bests: Improving Best-of-N via Bootstrapping
por: Rakhsha, Amin, et al.
Publicado: (2025)
por: Rakhsha, Amin, et al.
Publicado: (2025)
Variational Best-of-N Alignment
por: Amini, Afra, et al.
Publicado: (2024)
por: Amini, Afra, et al.
Publicado: (2024)
Novel Pathogenic Variant in Exon 31 of the TSC2 Gene Associated With a Severe Phenotype of Tuberous Sclerosis
por: Tabitha D'souza, et al.
Publicado: (2025)
por: Tabitha D'souza, et al.
Publicado: (2025)
Synthetic Tabular Data Generation for Imbalanced Classification: The Surprising Effectiveness of an Overlap Class
por: D'souza, Annie, et al.
Publicado: (2024)
por: D'souza, Annie, et al.
Publicado: (2024)
Critical Learning Periods: Leveraging Early Training Dynamics for Efficient Data Pruning
por: Chimoto, Everlyn Asiko, et al.
Publicado: (2024)
por: Chimoto, Everlyn Asiko, et al.
Publicado: (2024)
PairJudge RM: Perform Best-of-N Sampling with Knockout Tournament
por: Liu, Yantao, et al.
Publicado: (2025)
por: Liu, Yantao, et al.
Publicado: (2025)
N-gram-like Language Models Predict Reading Time Best
por: Michaelov, James A., et al.
Publicado: (2026)
por: Michaelov, James A., et al.
Publicado: (2026)
Beyond Majority Voting: Efficient Best-Of-N with Radial Consensus Score
por: Nguyen, Manh, et al.
Publicado: (2026)
por: Nguyen, Manh, et al.
Publicado: (2026)
From Test-Taking to Test-Making: Examining LLM Authoring of Commonsense Assessment Items
por: Roemmele, Melissa, et al.
Publicado: (2024)
por: Roemmele, Melissa, et al.
Publicado: (2024)
AdaBoN: Adaptive Best-of-N Alignment
por: Raman, Vinod, et al.
Publicado: (2025)
por: Raman, Vinod, et al.
Publicado: (2025)
Structured Pruning for Diverse Best-of-N Reasoning Optimization
por: Nguyen, Hieu Trung, et al.
Publicado: (2025)
por: Nguyen, Hieu Trung, et al.
Publicado: (2025)
GenSelect: A Generative Approach to Best-of-N
por: Toshniwal, Shubham, et al.
Publicado: (2025)
por: Toshniwal, Shubham, et al.
Publicado: (2025)
Learning Generative Selection for Best-of-N
por: Toshniwal, Shubham, et al.
Publicado: (2026)
por: Toshniwal, Shubham, et al.
Publicado: (2026)
Ejemplares similares
-
When Life Gives You Samples: The Benefits of Scaling up Inference Compute for Multilingual LLMs
por: Khairi, Ammar, et al.
Publicado: (2025) -
Treasure Hunt: Real-time Targeting of the Long Tail using Training-Time Markers
por: D'souza, Daniel, et al.
Publicado: (2025) -
Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
por: Kreutzer, Julia, et al.
Publicado: (2025) -
LLM See, LLM Do: Guiding Data Generation to Target Non-Differentiable Objectives
por: Shimabucoro, Luísa, et al.
Publicado: (2024) -
The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It
por: Yong, Zheng-Xin, et al.
Publicado: (2025)