Efficient and Principled Scientific Discovery through Bayesian Optimization: A Tutorial
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Zhongwei, Tutunov, Rasul, Maraval, Alexandre Max, Xie, Zikai, Tan, Zhenzhi, Wang, Jiankang, Cao, Bin, Li, Zijing, Xu, Liangliang, Yang, Qi, Jiang, Jun, Luo, Sanzhong, Guo, Zhenxiao, Zhang, Tongyi, Bou-Ammar, Haitham, Wang, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Model-Based and Sample-Efficient AI-Assisted Math Discovery in Sphere Packing
by: Tutunov, Rasul, et al.
Published: (2025)
by: Tutunov, Rasul, et al.
Published: (2025)
Why Can Large Language Models Generate Correct Chain-of-Thoughts?
by: Tutunov, Rasul, et al.
Published: (2023)
by: Tutunov, Rasul, et al.
Published: (2023)
Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers
by: Ji, Xiaotong, et al.
Published: (2026)
by: Ji, Xiaotong, et al.
Published: (2026)
Scalable Power Sampling: Unlocking Efficient, Training-Free Reasoning for LLMs via Distribution Sharpening
by: Ji, Xiaotong, et al.
Published: (2026)
by: Ji, Xiaotong, et al.
Published: (2026)
Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving
by: Zimmer, Matthieu, et al.
Published: (2025)
by: Zimmer, Matthieu, et al.
Published: (2025)
The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling
by: Nguyen, Tu, et al.
Published: (2026)
by: Nguyen, Tu, et al.
Published: (2026)
Risk-Controlled Lean-as-Judge for Natural-Language Mathematical Reasoning
by: Bourigault, Pauline, et al.
Published: (2026)
by: Bourigault, Pauline, et al.
Published: (2026)
The $\mathbf{Y}$-Combinator for LLMs: Solving Long-Context Rot with $λ$-Calculus
by: Roy, Amartya, et al.
Published: (2026)
by: Roy, Amartya, et al.
Published: (2026)
Many of Your DPOs are Secretly One: Attempting Unification Through Mutual Information
by: Tutnov, Rasul, et al.
Published: (2025)
by: Tutnov, Rasul, et al.
Published: (2025)
Bottlenecked Transformers: Periodic KV Cache Consolidation for Generalised Reasoning
by: Oomerjee, Adnan, et al.
Published: (2025)
by: Oomerjee, Adnan, et al.
Published: (2025)
Contextual Causal Bayesian Optimisation
by: Arsenyan, Vahan, et al.
Published: (2023)
by: Arsenyan, Vahan, et al.
Published: (2023)
Bayesian Reward Models for LLM Alignment
by: Yang, Adam X., et al.
Published: (2024)
by: Yang, Adam X., et al.
Published: (2024)
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
by: Christopoulou, Fenia, et al.
Published: (2024)
by: Christopoulou, Fenia, et al.
Published: (2024)
Why the Brain Consolidates: Predictive Forgetting for Optimal Generalisation
by: Fountas, Zafeirios, et al.
Published: (2026)
by: Fountas, Zafeirios, et al.
Published: (2026)
Mixture of Attentions For Speculative Decoding
by: Zimmer, Matthieu, et al.
Published: (2024)
by: Zimmer, Matthieu, et al.
Published: (2024)
Subjective Depth and Timescale Transformers: Learning Where and When to Compute
by: Wieser, Frederico, et al.
Published: (2025)
by: Wieser, Frederico, et al.
Published: (2025)
Al-Khwarizmi: Discovering Physical Laws with Foundation Models
by: Mower, Christopher E., et al.
Published: (2025)
by: Mower, Christopher E., et al.
Published: (2025)
Synthetic Applicability Domain (SynAD): Navigating Chemical Space for Reliable AI‐Driven Reaction Prediction
by: Zhenzhi Tan, et al.
Published: (2026)
by: Zhenzhi Tan, et al.
Published: (2026)
Synthetic Applicability Domain (SynAD): Navigating Chemical Space for Reliable AI‐Driven Reaction Prediction
by: Zhenzhi Tan, et al.
Published: (2026)
by: Zhenzhi Tan, et al.
Published: (2026)
Untangling Component Imbalance in Hybrid Linear Attention Conversion Methods
by: Benfeghoul, Martin, et al.
Published: (2025)
by: Benfeghoul, Martin, et al.
Published: (2025)
Data-driven Interpretable Hybrid Robot Dynamics
by: Mower, Christopher E., et al.
Published: (2025)
by: Mower, Christopher E., et al.
Published: (2025)
On Almost Surely Safe Alignment of Large Language Models at Inference-Time
by: Ji, Xiaotong, et al.
Published: (2025)
by: Ji, Xiaotong, et al.
Published: (2025)
SuRe: Surprise-Driven Prioritised Replay for Continual LLM Learning
by: Hazard, Hugo, et al.
Published: (2025)
by: Hazard, Hugo, et al.
Published: (2025)
Efficient Reinforcement Learning with Large Language Model Priors
by: Yan, Xue, et al.
Published: (2024)
by: Yan, Xue, et al.
Published: (2024)
Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications
by: Liu, Puze, et al.
Published: (2024)
by: Liu, Puze, et al.
Published: (2024)
Rethinking Large Language Model Distillation: A Constrained Markov Decision Process Perspective
by: Zimmer, Matthieu, et al.
Published: (2025)
by: Zimmer, Matthieu, et al.
Published: (2025)
Human-inspired Episodic Memory for Infinite Context LLMs
by: Fountas, Zafeirios, et al.
Published: (2024)
by: Fountas, Zafeirios, et al.
Published: (2024)
La religion de Constantin
by: Pierre Maraval
Published: (2013)
by: Pierre Maraval
Published: (2013)
ZSL-RPPO: Zero-Shot Learning for Quadrupedal Locomotion in Challenging Terrains using Recurrent Proximal Policy Optimization
by: Zhao, Yao, et al.
Published: (2024)
by: Zhao, Yao, et al.
Published: (2024)
Kolb-Based Experiential Learning for Generalist Agents with Human-Level Kaggle Data Science Performance
by: Grosnit, Antoine, et al.
Published: (2024)
by: Grosnit, Antoine, et al.
Published: (2024)
ShortCircuit: AlphaZero-Driven Circuit Design
by: Tsaras, Dimitrios, et al.
Published: (2024)
by: Tsaras, Dimitrios, et al.
Published: (2024)
A Tutorial on LLM Reasoning: Relevant Methods behind ChatGPT o1
by: Wang, Jun
Published: (2025)
by: Wang, Jun
Published: (2025)
Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates
by: Diwan, Anish, et al.
Published: (2026)
by: Diwan, Anish, et al.
Published: (2026)
A Brain-like Synergistic Core in LLMs Drives Behaviour and Learning
by: Urbina-Rodriguez, Pedro, et al.
Published: (2026)
by: Urbina-Rodriguez, Pedro, et al.
Published: (2026)
Phase‐Transfer Catalysis for Electrochemical Chlorination and Nitration of Arenes
by: Yingdong Duan, et al.
Published: (2024)
by: Yingdong Duan, et al.
Published: (2024)
Principle-Evolvable Scientific Discovery via Uncertainty Minimization
by: Pu, Yingming, et al.
Published: (2026)
by: Pu, Yingming, et al.
Published: (2026)
Multi-Task GRPO: Reliable LLM Reasoning Across Tasks
by: Ramesh, Shyam Sundhar, et al.
Published: (2026)
by: Ramesh, Shyam Sundhar, et al.
Published: (2026)
Willmore surfaces in 4-dimensional conformal manifolds
by: Wang, Changping, et al.
Published: (2023)
by: Wang, Changping, et al.
Published: (2023)
Research Progress, Challenges, and Prospects of Primary Liver Cancer Organoids
by: Zhenxiao Wang, et al.
Published: (2026)
by: Zhenxiao Wang, et al.
Published: (2026)
A Pragmatist Robot: Learning to Plan Tasks by Experiencing the Real World
by: Qu, Kaixian, et al.
Published: (2025)
by: Qu, Kaixian, et al.
Published: (2025)
Similar Items
-
Model-Based and Sample-Efficient AI-Assisted Math Discovery in Sphere Packing
by: Tutunov, Rasul, et al.
Published: (2025) -
Why Can Large Language Models Generate Correct Chain-of-Thoughts?
by: Tutunov, Rasul, et al.
Published: (2023) -
Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers
by: Ji, Xiaotong, et al.
Published: (2026) -
Scalable Power Sampling: Unlocking Efficient, Training-Free Reasoning for LLMs via Distribution Sharpening
by: Ji, Xiaotong, et al.
Published: (2026) -
Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving
by: Zimmer, Matthieu, et al.
Published: (2025)