Saved in:
| Main Authors: | Wan, Guangya, Xu, Zixin Stephen, Zorc, Sasa, Baucells, Manel, Hu, Mengxuan, Wang, Hao, Li, Sheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2510.15945 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers
by: Huang, Jingkai, et al.
Published: (2026)
by: Huang, Jingkai, et al.
Published: (2026)
Large Language Models for Causal Discovery: Current Landscape and Future Directions
by: Wan, Guangya, et al.
Published: (2024)
by: Wan, Guangya, et al.
Published: (2024)
Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling
by: Wan, Guangya, et al.
Published: (2024)
by: Wan, Guangya, et al.
Published: (2024)
Task-Driven Causal Feature Distillation: Towards Trustworthy Risk Prediction
by: Chu, Zhixuan, et al.
Published: (2023)
by: Chu, Zhixuan, et al.
Published: (2023)
Batch Bayesian Active Learning with Partial Batch Label Sampling
by: Hu, Kangping, et al.
Published: (2025)
by: Hu, Kangping, et al.
Published: (2025)
When to Stop Reusing: Dynamic Gradient Gating for Sample-Efficient RLVR
by: Miao, Yuchun, et al.
Published: (2026)
by: Miao, Yuchun, et al.
Published: (2026)
Self-Improving Tabular Language Models via Iterative Reward-Guided Post-Training
by: Long, Yunbo, et al.
Published: (2026)
by: Long, Yunbo, et al.
Published: (2026)
IsoCompute Playbook: Optimally Scaling Sampling Compute for LLM RL
by: Cheng, Zhoujun, et al.
Published: (2026)
by: Cheng, Zhoujun, et al.
Published: (2026)
BEACON: Behavioral Malware Classification with Large Language Model Embeddings and Deep Learning
by: Perera, Wadduwage Shanika, et al.
Published: (2025)
by: Perera, Wadduwage Shanika, et al.
Published: (2025)
Efficient Network Automatic Relevance Determination
by: Zhang, Hongwei, et al.
Published: (2025)
by: Zhang, Hongwei, et al.
Published: (2025)
Ratio-Variance Regularized Policy Optimization for Efficient LLM Fine-tuning
by: Luo, Yu, et al.
Published: (2026)
by: Luo, Yu, et al.
Published: (2026)
Efficient Detection of LLM-generated Texts with a Bayesian Surrogate Model
by: Miao, Yibo, et al.
Published: (2023)
by: Miao, Yibo, et al.
Published: (2023)
Learning Optimal and Sample-Efficient Decision Policies with Guarantees
by: Shao, Daqian
Published: (2026)
by: Shao, Daqian
Published: (2026)
Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility
by: Wang, Mengxuan, et al.
Published: (2026)
by: Wang, Mengxuan, et al.
Published: (2026)
Heterogeneity-Aware Client Sampling for Optimal and Efficient Federated Learning
by: Weng, Shudi, et al.
Published: (2025)
by: Weng, Shudi, et al.
Published: (2025)
Sample-Efficient Bayesian Optimization with Transfer Learning for Heterogeneous Search Spaces
by: Deshwal, Aryan, et al.
Published: (2024)
by: Deshwal, Aryan, et al.
Published: (2024)
Are LLMs Ready for Neural-integrated Mechanistic Modeling? A Benchmark and Agentic Framework
by: Guan, Zihan, et al.
Published: (2026)
by: Guan, Zihan, et al.
Published: (2026)
Efficient Refusal Ablation in LLM through Optimal Transport
by: Nanfack, Geraldin, et al.
Published: (2026)
by: Nanfack, Geraldin, et al.
Published: (2026)
A Novel Approach to Balance Convenience and Nutrition in Meals With Long-Term Group Recommendations and Reasoning on Multimodal Recipes and its Implementation in BEACON
by: Nagpal, Vansh, et al.
Published: (2024)
by: Nagpal, Vansh, et al.
Published: (2024)
S2O: Early Stopping for Sparse Attention via Online Permutation
by: Zhang, Yu, et al.
Published: (2026)
by: Zhang, Yu, et al.
Published: (2026)
ESPO: Early-Stopping Proximal Policy Optimization
by: Li, Zihang, et al.
Published: (2026)
by: Li, Zihang, et al.
Published: (2026)
Intrusion Prevention through Optimal Stopping
by: Hammar, Kim, et al.
Published: (2021)
by: Hammar, Kim, et al.
Published: (2021)
Ten Words Only Still Help: Improving Black-Box AI-Generated Text Detection via Proxy-Guided Efficient Re-Sampling
by: Shi, Yuhui, et al.
Published: (2024)
by: Shi, Yuhui, et al.
Published: (2024)
Learning the Optimal Stopping for Early Classification within Finite Horizons via Sequential Probability Ratio Test
by: Ebihara, Akinori F., et al.
Published: (2025)
by: Ebihara, Akinori F., et al.
Published: (2025)
FSX: Message Flow Sensitivity Enhanced Structural Explainer for Graph Neural Networks
by: Feng, Bizu, et al.
Published: (2026)
by: Feng, Bizu, et al.
Published: (2026)
eFedLLM: Efficient LLM Inference Based on Federated Learning
by: Ding, Shengwen, et al.
Published: (2024)
by: Ding, Shengwen, et al.
Published: (2024)
FwdLLM: Efficient FedLLM using Forward Gradient
by: Xu, Mengwei, et al.
Published: (2023)
by: Xu, Mengwei, et al.
Published: (2023)
Efficient Quantification of Multimodal Interaction at Sample Level
by: Yang, Zequn, et al.
Published: (2025)
by: Yang, Zequn, et al.
Published: (2025)
MASPO: Unifying Gradient Utilization, Probability Mass, and Signal Reliability for Robust and Sample-Efficient LLM Reasoning
by: Fu, Xiaoliang, et al.
Published: (2026)
by: Fu, Xiaoliang, et al.
Published: (2026)
Almost Minimax Optimal Best Arm Identification in Piecewise Stationary Linear Bandits
by: Hou, Yunlong, et al.
Published: (2024)
by: Hou, Yunlong, et al.
Published: (2024)
Reason for Future, Act for Now: A Principled Framework for Autonomous LLM Agents with Provable Sample Efficiency
by: Liu, Zhihan, et al.
Published: (2023)
by: Liu, Zhihan, et al.
Published: (2023)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
by: Melo, Luckeciano C., et al.
Published: (2025)
by: Melo, Luckeciano C., et al.
Published: (2025)
PerFlow: Physics-Embedded Rectified Flow for Efficient Reconstruction and Uncertainty Quantification of Spatiotemporal Dynamics
by: Zhou, Hao, et al.
Published: (2026)
by: Zhou, Hao, et al.
Published: (2026)
APLOT: Robust Reward Modeling via Adaptive Preference Learning with Optimal Transport
by: Li, Zhuo, et al.
Published: (2025)
by: Li, Zhuo, et al.
Published: (2025)
TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning
by: Nagle, Alliot, et al.
Published: (2026)
by: Nagle, Alliot, et al.
Published: (2026)
Flexible Bayesian Last Layer Models Using Implicit Priors and Diffusion Posterior Sampling
by: Xu, Jian, et al.
Published: (2024)
by: Xu, Jian, et al.
Published: (2024)
Optimal Policy Minimum Bayesian Risk
by: Astudillo, Ramón Fernandez, et al.
Published: (2025)
by: Astudillo, Ramón Fernandez, et al.
Published: (2025)
Machine Unlearning in Contrastive Learning
by: Wang, Zixin, et al.
Published: (2024)
by: Wang, Zixin, et al.
Published: (2024)
Optimal Transport for LLM Reward Modeling from Noisy Preference
by: Pan, Licheng, et al.
Published: (2026)
by: Pan, Licheng, et al.
Published: (2026)
Scaling Up Bayesian DAG Sampling
by: Nikzad, Daniele, et al.
Published: (2025)
by: Nikzad, Daniele, et al.
Published: (2025)
Similar Items
-
Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers
by: Huang, Jingkai, et al.
Published: (2026) -
Large Language Models for Causal Discovery: Current Landscape and Future Directions
by: Wan, Guangya, et al.
Published: (2024) -
Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling
by: Wan, Guangya, et al.
Published: (2024) -
Task-Driven Causal Feature Distillation: Towards Trustworthy Risk Prediction
by: Chu, Zhixuan, et al.
Published: (2023) -
Batch Bayesian Active Learning with Partial Batch Label Sampling
by: Hu, Kangping, et al.
Published: (2025)