Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Jingkai, Ma, Will, Zhou, Zhengyuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BEACON: Bayesian Optimal Stopping for Efficient LLM Sampling
von: Wan, Guangya, et al.
Veröffentlicht: (2025)
von: Wan, Guangya, et al.
Veröffentlicht: (2025)
Leveraging Self-Consistency for Data-Efficient Amortized Bayesian Inference
von: Schmitt, Marvin, et al.
Veröffentlicht: (2023)
von: Schmitt, Marvin, et al.
Veröffentlicht: (2023)
Consistency Training Helps Stop Sycophancy and Jailbreaks
von: Irpan, Alex, et al.
Veröffentlicht: (2025)
von: Irpan, Alex, et al.
Veröffentlicht: (2025)
Bayesian Inference of Training Dataset Membership
von: Huang, Yongchao
Veröffentlicht: (2025)
von: Huang, Yongchao
Veröffentlicht: (2025)
Functional Critics Are Essential for Actor-Critic: From Off-Policy Stability to Efficient Exploration
von: Bai, Qinxun, et al.
Veröffentlicht: (2025)
von: Bai, Qinxun, et al.
Veröffentlicht: (2025)
Optimal Singular Damage: Efficient LLM Inference in Low Storage Regimes
von: Alipour, Mohammadsajad, et al.
Veröffentlicht: (2025)
von: Alipour, Mohammadsajad, et al.
Veröffentlicht: (2025)
Optimal Self-Consistency for Efficient Reasoning with Large Language Models
von: Feng, Austin, et al.
Veröffentlicht: (2025)
von: Feng, Austin, et al.
Veröffentlicht: (2025)
Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning
von: Zhou, Zhi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhi, et al.
Veröffentlicht: (2025)
ES-dLLM: Efficient Inference for Diffusion Large Language Models by Early-Skipping
von: Zhu, Zijian, et al.
Veröffentlicht: (2026)
von: Zhu, Zijian, et al.
Veröffentlicht: (2026)
Robust and Efficient Zeroth-Order LLM Fine-Tuning via Adaptive Bayesian Subspace Optimizer
von: Feng, Jian, et al.
Veröffentlicht: (2026)
von: Feng, Jian, et al.
Veröffentlicht: (2026)
eFedLLM: Efficient LLM Inference Based on Federated Learning
von: Ding, Shengwen, et al.
Veröffentlicht: (2024)
von: Ding, Shengwen, et al.
Veröffentlicht: (2024)
A Theoretical Study on Bridging Internal Probability and Self-Consistency for LLM Reasoning
von: Zhou, Zhi, et al.
Veröffentlicht: (2025)
von: Zhou, Zhi, et al.
Veröffentlicht: (2025)
Efficient Refusal Ablation in LLM through Optimal Transport
von: Nanfack, Geraldin, et al.
Veröffentlicht: (2026)
von: Nanfack, Geraldin, et al.
Veröffentlicht: (2026)
Efficient Reinforcement Learning from Human Feedback via Bayesian Preference Inference
von: Cercola, Matteo, et al.
Veröffentlicht: (2025)
von: Cercola, Matteo, et al.
Veröffentlicht: (2025)
Single-Trajectory Distributionally Robust Reinforcement Learning
von: Liang, Zhipeng, et al.
Veröffentlicht: (2023)
von: Liang, Zhipeng, et al.
Veröffentlicht: (2023)
How does Bayesian Sampling help Membership Inference Attacks?
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
von: Liu, Zhenlong, et al.
Veröffentlicht: (2025)
When to Stop Reusing: Dynamic Gradient Gating for Sample-Efficient RLVR
von: Miao, Yuchun, et al.
Veröffentlicht: (2026)
von: Miao, Yuchun, et al.
Veröffentlicht: (2026)
LLM Swiss Round: Aggregating Multi-Benchmark Performance via Competitive Swiss-System Dynamics
von: Liu, Jiashuo, et al.
Veröffentlicht: (2025)
von: Liu, Jiashuo, et al.
Veröffentlicht: (2025)
Bayesian Inverse Problems Meet Flow Matching: Efficient and Flexible Inference via Transformers
von: Sherki, Daniil, et al.
Veröffentlicht: (2025)
von: Sherki, Daniil, et al.
Veröffentlicht: (2025)
Learning to Answer from Correct Demonstrations
von: Joshi, Nirmit, et al.
Veröffentlicht: (2025)
von: Joshi, Nirmit, et al.
Veröffentlicht: (2025)
HeadInfer: Memory-Efficient LLM Inference by Head-wise Offloading
von: Luo, Cheng, et al.
Veröffentlicht: (2025)
von: Luo, Cheng, et al.
Veröffentlicht: (2025)
Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
von: Chen, Tiejin, et al.
Veröffentlicht: (2026)
CITE: Anytime-Valid Statistical Inference in LLM Self-Consistency
von: Ota, Hirofumi, et al.
Veröffentlicht: (2026)
von: Ota, Hirofumi, et al.
Veröffentlicht: (2026)
PETS: A Principled Framework Towards Optimal Trajectory Allocation for Efficient Test-Time Self-Consistency
von: Liu, Zhangyi, et al.
Veröffentlicht: (2026)
von: Liu, Zhangyi, et al.
Veröffentlicht: (2026)
Learning the Optimal Stopping for Early Classification within Finite Horizons via Sequential Probability Ratio Test
von: Ebihara, Akinori F., et al.
Veröffentlicht: (2025)
von: Ebihara, Akinori F., et al.
Veröffentlicht: (2025)
Combining Bayesian Inference and Reinforcement Learning for Agent Decision Making: A Review
von: Zhou, Chengmin, et al.
Veröffentlicht: (2025)
von: Zhou, Chengmin, et al.
Veröffentlicht: (2025)
Collapsed Inference for Bayesian Deep Learning
von: Zeng, Zhe, et al.
Veröffentlicht: (2023)
von: Zeng, Zhe, et al.
Veröffentlicht: (2023)
Geometric Scaling of Bayesian Inference in LLMs
von: Agarwal, Naman, et al.
Veröffentlicht: (2025)
von: Agarwal, Naman, et al.
Veröffentlicht: (2025)
On Sequential Bayesian Inference for Continual Learning
von: Kessler, Samuel, et al.
Veröffentlicht: (2023)
von: Kessler, Samuel, et al.
Veröffentlicht: (2023)
ConSol: Sequential Probability Ratio Testing to Find Consistent LLM Reasoning Paths Efficiently
von: Lee, Jaeyeon, et al.
Veröffentlicht: (2025)
von: Lee, Jaeyeon, et al.
Veröffentlicht: (2025)
Consistency Models for Scalable and Fast Simulation-Based Inference
von: Schmitt, Marvin, et al.
Veröffentlicht: (2023)
von: Schmitt, Marvin, et al.
Veröffentlicht: (2023)
ESPO: Early-Stopping Proximal Policy Optimization
von: Li, Zihang, et al.
Veröffentlicht: (2026)
von: Li, Zihang, et al.
Veröffentlicht: (2026)
RAMP: Reinforcement Adaptive Mixed Precision Quantization for Efficient On Device LLM Inference
von: Gautam, Arpit Singh, et al.
Veröffentlicht: (2026)
von: Gautam, Arpit Singh, et al.
Veröffentlicht: (2026)
DSAC: Distributional Soft Actor-Critic for Risk-Sensitive Reinforcement Learning
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
von: Ma, Xiaoteng, et al.
Veröffentlicht: (2020)
From Stochastic Answers to Verifiable Reasoning: Interpretable Decision-Making with LLM-Generated Code
von: Mahesh, Anirudh Jaidev, et al.
Veröffentlicht: (2026)
von: Mahesh, Anirudh Jaidev, et al.
Veröffentlicht: (2026)
Anytime-Valid Answer Sufficiency Certificates for LLM Generation via Sequential Information Lift
von: Akter, Sanjeda, et al.
Veröffentlicht: (2025)
von: Akter, Sanjeda, et al.
Veröffentlicht: (2025)
EntropyStop: Unsupervised Deep Outlier Detection with Loss Entropy
von: Huang, Yihong, et al.
Veröffentlicht: (2024)
von: Huang, Yihong, et al.
Veröffentlicht: (2024)
Optimal Policy Minimum Bayesian Risk
von: Astudillo, Ramón Fernandez, et al.
Veröffentlicht: (2025)
von: Astudillo, Ramón Fernandez, et al.
Veröffentlicht: (2025)
Calibrated Test-Time Guidance for Bayesian Inference
von: Geyfman, Daniel, et al.
Veröffentlicht: (2026)
von: Geyfman, Daniel, et al.
Veröffentlicht: (2026)
Bayesian Inference for Correlated Human Experts and Classifiers
von: Kelly, Markelle, et al.
Veröffentlicht: (2025)
von: Kelly, Markelle, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BEACON: Bayesian Optimal Stopping for Efficient LLM Sampling
von: Wan, Guangya, et al.
Veröffentlicht: (2025) -
Leveraging Self-Consistency for Data-Efficient Amortized Bayesian Inference
von: Schmitt, Marvin, et al.
Veröffentlicht: (2023) -
Consistency Training Helps Stop Sycophancy and Jailbreaks
von: Irpan, Alex, et al.
Veröffentlicht: (2025) -
Bayesian Inference of Training Dataset Membership
von: Huang, Yongchao
Veröffentlicht: (2025) -
Functional Critics Are Essential for Actor-Critic: From Off-Policy Stability to Efficient Exploration
von: Bai, Qinxun, et al.
Veröffentlicht: (2025)