Optimal Stopping vs Best-of-$N$ for Inference Time Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Kalayci, Yusuf, Raman, Vinod, Dughmi, Shaddin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Generate-Rank-Verify: Inference-Time Search with Costly Verification
by: Dughmi, Shaddin, et al.
Published: (2026)
by: Dughmi, Shaddin, et al.
Published: (2026)
Is Transductive Learning Equivalent to PAC Learning?
by: Dughmi, Shaddin, et al.
Published: (2024)
by: Dughmi, Shaddin, et al.
Published: (2024)
A Theory of Time-Sensitive Language Generation: Sparse Hallucination Beats Mode Collapse
by: Ganju, Atul, et al.
Published: (2026)
by: Ganju, Atul, et al.
Published: (2026)
AdaBoN: Adaptive Best-of-N Alignment
by: Raman, Vinod, et al.
Published: (2025)
by: Raman, Vinod, et al.
Published: (2025)
PAC Learning is just Bipartite Matching (Sort of)
by: Dughmi, Shaddin
Published: (2025)
by: Dughmi, Shaddin
Published: (2025)
Near-Optimal Sparsifiers for Stochastic Knapsack and Assignment Problems
by: Dughmi, Shaddin, et al.
Published: (2025)
by: Dughmi, Shaddin, et al.
Published: (2025)
Relatively Smart: A New Approach for Instance-Optimal Learning
by: Dughmi, Shaddin, et al.
Published: (2026)
by: Dughmi, Shaddin, et al.
Published: (2026)
Limitations of Stochastic Selection with Pairwise Independent Priors
by: Dughmi, Shaddin, et al.
Published: (2023)
by: Dughmi, Shaddin, et al.
Published: (2023)
Local Regularizers Are Not Transductive Learners
by: Jafar, Sky, et al.
Published: (2025)
by: Jafar, Sky, et al.
Published: (2025)
Regularization and Optimal Multiclass Learning
by: Asilis, Julian, et al.
Published: (2023)
by: Asilis, Julian, et al.
Published: (2023)
Representative Language Generation
by: Peale, Charlotte, et al.
Published: (2025)
by: Peale, Charlotte, et al.
Published: (2025)
Structured Pruning for Diverse Best-of-N Reasoning Optimization
by: Nguyen, Hieu Trung, et al.
Published: (2025)
by: Nguyen, Hieu Trung, et al.
Published: (2025)
TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling
by: Qiu, Jiahao, et al.
Published: (2024)
by: Qiu, Jiahao, et al.
Published: (2024)
LYNX: Learning Dynamic Exits for Confidence-Controlled Reasoning
by: Akgül, Ömer Faruk, et al.
Published: (2025)
by: Akgül, Ömer Faruk, et al.
Published: (2025)
Learning to Choose or Choosing to Learn: Best-of-N vs. Supervised Fine-Tuning for Bit String Generation
by: Somerstep, Seamus, et al.
Published: (2025)
by: Somerstep, Seamus, et al.
Published: (2025)
Transductive Learning Is Compact
by: Asilis, Julian, et al.
Published: (2024)
by: Asilis, Julian, et al.
Published: (2024)
Efficient Multi-Agent Delegated Search
by: Bechtel, Curtis, et al.
Published: (2024)
by: Bechtel, Curtis, et al.
Published: (2024)
Inference-Aware Fine-Tuning for Best-of-N Sampling in Large Language Models
by: Chow, Yinlam, et al.
Published: (2024)
by: Chow, Yinlam, et al.
Published: (2024)
Inference Scaling vs Reasoning: An Empirical Analysis of Compute-Optimal LLM Problem-Solving
by: AbdElhameed, Marwan, et al.
Published: (2024)
by: AbdElhameed, Marwan, et al.
Published: (2024)
Best-of-N Jailbreaking
by: Hughes, John, et al.
Published: (2024)
by: Hughes, John, et al.
Published: (2024)
Just on Time: Token-Level Early Stopping for Diffusion Language Models
by: Kohut, Zahar, et al.
Published: (2026)
by: Kohut, Zahar, et al.
Published: (2026)
GenSelect: A Generative Approach to Best-of-N
by: Toshniwal, Shubham, et al.
Published: (2025)
by: Toshniwal, Shubham, et al.
Published: (2025)
FLOP-Efficient Training: Early Stopping Based on Test-Time Compute Awareness
by: Amer, Hossam, et al.
Published: (2026)
by: Amer, Hossam, et al.
Published: (2026)
Proper Learnability and the Role of Unlabeled Data
by: Asilis, Julian, et al.
Published: (2025)
by: Asilis, Julian, et al.
Published: (2025)
Variational Best-of-N Alignment
by: Amini, Afra, et al.
Published: (2024)
by: Amini, Afra, et al.
Published: (2024)
Majority of the Bests: Improving Best-of-N via Bootstrapping
by: Rakhsha, Amin, et al.
Published: (2025)
by: Rakhsha, Amin, et al.
Published: (2025)
Tracking the Best Expert Privately
by: Saha, Aadirupa, et al.
Published: (2025)
by: Saha, Aadirupa, et al.
Published: (2025)
Learning Generative Selection for Best-of-N
by: Toshniwal, Shubham, et al.
Published: (2026)
by: Toshniwal, Shubham, et al.
Published: (2026)
MarkovScale: Towards Optimal Sequential Scaling at Inference Time
by: Wang, Youkang, et al.
Published: (2026)
by: Wang, Youkang, et al.
Published: (2026)
TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning
by: Nagle, Alliot, et al.
Published: (2026)
by: Nagle, Alliot, et al.
Published: (2026)
AI-rithmetic
by: Bie, Alex, et al.
Published: (2026)
by: Bie, Alex, et al.
Published: (2026)
BOND: Aligning LLMs with Best-of-N Distillation
by: Sessa, Pier Giuseppe, et al.
Published: (2024)
by: Sessa, Pier Giuseppe, et al.
Published: (2024)
BoNBoN Alignment for Large Language Models and the Sweetness of Best-of-n Sampling
by: Gui, Lin, et al.
Published: (2024)
by: Gui, Lin, et al.
Published: (2024)
Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding
by: Oba, Daisuke, et al.
Published: (2026)
by: Oba, Daisuke, et al.
Published: (2026)
Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws
by: Sardana, Nikhil, et al.
Published: (2023)
by: Sardana, Nikhil, et al.
Published: (2023)
The Solution for the AIGC Inference Performance Optimization Competition
by: Pan, Sishun, et al.
Published: (2024)
by: Pan, Sishun, et al.
Published: (2024)
Stop-Think-AutoRegress: Language Modeling with Latent Diffusion Planning
by: Lovelace, Justin, et al.
Published: (2026)
by: Lovelace, Justin, et al.
Published: (2026)
Efficient Prompt Optimization Through the Lens of Best Arm Identification
by: Shi, Chengshuai, et al.
Published: (2024)
by: Shi, Chengshuai, et al.
Published: (2024)
Correct and Optimal: the Regular Expression Inference Challenge
by: Valizadeh, Mojtaba, et al.
Published: (2023)
by: Valizadeh, Mojtaba, et al.
Published: (2023)
Compute Optimal Scaling of Skills: Knowledge vs Reasoning
by: Roberts, Nicholas, et al.
Published: (2025)
by: Roberts, Nicholas, et al.
Published: (2025)
Similar Items
-
Adaptive Generate-Rank-Verify: Inference-Time Search with Costly Verification
by: Dughmi, Shaddin, et al.
Published: (2026) -
Is Transductive Learning Equivalent to PAC Learning?
by: Dughmi, Shaddin, et al.
Published: (2024) -
A Theory of Time-Sensitive Language Generation: Sparse Hallucination Beats Mode Collapse
by: Ganju, Atul, et al.
Published: (2026) -
AdaBoN: Adaptive Best-of-N Alignment
by: Raman, Vinod, et al.
Published: (2025) -
PAC Learning is just Bipartite Matching (Sort of)
by: Dughmi, Shaddin
Published: (2025)