Efficient and Principled Scientific Discovery through Bayesian Optimization: A Tutorial
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Zhongwei, Tutunov, Rasul, Maraval, Alexandre Max, Xie, Zikai, Tan, Zhenzhi, Wang, Jiankang, Cao, Bin, Li, Zijing, Xu, Liangliang, Yang, Qi, Jiang, Jun, Luo, Sanzhong, Guo, Zhenxiao, Zhang, Tongyi, Bou-Ammar, Haitham, Wang, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Model-Based and Sample-Efficient AI-Assisted Math Discovery in Sphere Packing
von: Tutunov, Rasul, et al.
Veröffentlicht: (2025)
von: Tutunov, Rasul, et al.
Veröffentlicht: (2025)
Why Can Large Language Models Generate Correct Chain-of-Thoughts?
von: Tutunov, Rasul, et al.
Veröffentlicht: (2023)
von: Tutunov, Rasul, et al.
Veröffentlicht: (2023)
Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers
von: Ji, Xiaotong, et al.
Veröffentlicht: (2026)
von: Ji, Xiaotong, et al.
Veröffentlicht: (2026)
Scalable Power Sampling: Unlocking Efficient, Training-Free Reasoning for LLMs via Distribution Sharpening
von: Ji, Xiaotong, et al.
Veröffentlicht: (2026)
von: Ji, Xiaotong, et al.
Veröffentlicht: (2026)
Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2025)
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2025)
The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling
von: Nguyen, Tu, et al.
Veröffentlicht: (2026)
von: Nguyen, Tu, et al.
Veröffentlicht: (2026)
Risk-Controlled Lean-as-Judge for Natural-Language Mathematical Reasoning
von: Bourigault, Pauline, et al.
Veröffentlicht: (2026)
von: Bourigault, Pauline, et al.
Veröffentlicht: (2026)
The $\mathbf{Y}$-Combinator for LLMs: Solving Long-Context Rot with $λ$-Calculus
von: Roy, Amartya, et al.
Veröffentlicht: (2026)
von: Roy, Amartya, et al.
Veröffentlicht: (2026)
Many of Your DPOs are Secretly One: Attempting Unification Through Mutual Information
von: Tutnov, Rasul, et al.
Veröffentlicht: (2025)
von: Tutnov, Rasul, et al.
Veröffentlicht: (2025)
Bottlenecked Transformers: Periodic KV Cache Consolidation for Generalised Reasoning
von: Oomerjee, Adnan, et al.
Veröffentlicht: (2025)
von: Oomerjee, Adnan, et al.
Veröffentlicht: (2025)
Contextual Causal Bayesian Optimisation
von: Arsenyan, Vahan, et al.
Veröffentlicht: (2023)
von: Arsenyan, Vahan, et al.
Veröffentlicht: (2023)
Bayesian Reward Models for LLM Alignment
von: Yang, Adam X., et al.
Veröffentlicht: (2024)
von: Yang, Adam X., et al.
Veröffentlicht: (2024)
SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks
von: Christopoulou, Fenia, et al.
Veröffentlicht: (2024)
von: Christopoulou, Fenia, et al.
Veröffentlicht: (2024)
Why the Brain Consolidates: Predictive Forgetting for Optimal Generalisation
von: Fountas, Zafeirios, et al.
Veröffentlicht: (2026)
von: Fountas, Zafeirios, et al.
Veröffentlicht: (2026)
Mixture of Attentions For Speculative Decoding
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2024)
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2024)
Subjective Depth and Timescale Transformers: Learning Where and When to Compute
von: Wieser, Frederico, et al.
Veröffentlicht: (2025)
von: Wieser, Frederico, et al.
Veröffentlicht: (2025)
Al-Khwarizmi: Discovering Physical Laws with Foundation Models
von: Mower, Christopher E., et al.
Veröffentlicht: (2025)
von: Mower, Christopher E., et al.
Veröffentlicht: (2025)
Synthetic Applicability Domain (SynAD): Navigating Chemical Space for Reliable AI‐Driven Reaction Prediction
von: Zhenzhi Tan, et al.
Veröffentlicht: (2026)
von: Zhenzhi Tan, et al.
Veröffentlicht: (2026)
Synthetic Applicability Domain (SynAD): Navigating Chemical Space for Reliable AI‐Driven Reaction Prediction
von: Zhenzhi Tan, et al.
Veröffentlicht: (2026)
von: Zhenzhi Tan, et al.
Veröffentlicht: (2026)
Untangling Component Imbalance in Hybrid Linear Attention Conversion Methods
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2025)
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2025)
Data-driven Interpretable Hybrid Robot Dynamics
von: Mower, Christopher E., et al.
Veröffentlicht: (2025)
von: Mower, Christopher E., et al.
Veröffentlicht: (2025)
On Almost Surely Safe Alignment of Large Language Models at Inference-Time
von: Ji, Xiaotong, et al.
Veröffentlicht: (2025)
von: Ji, Xiaotong, et al.
Veröffentlicht: (2025)
SuRe: Surprise-Driven Prioritised Replay for Continual LLM Learning
von: Hazard, Hugo, et al.
Veröffentlicht: (2025)
von: Hazard, Hugo, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Learning with Large Language Model Priors
von: Yan, Xue, et al.
Veröffentlicht: (2024)
von: Yan, Xue, et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications
von: Liu, Puze, et al.
Veröffentlicht: (2024)
von: Liu, Puze, et al.
Veröffentlicht: (2024)
Rethinking Large Language Model Distillation: A Constrained Markov Decision Process Perspective
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2025)
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2025)
Human-inspired Episodic Memory for Infinite Context LLMs
von: Fountas, Zafeirios, et al.
Veröffentlicht: (2024)
von: Fountas, Zafeirios, et al.
Veröffentlicht: (2024)
La religion de Constantin
von: Pierre Maraval
Veröffentlicht: (2013)
von: Pierre Maraval
Veröffentlicht: (2013)
ZSL-RPPO: Zero-Shot Learning for Quadrupedal Locomotion in Challenging Terrains using Recurrent Proximal Policy Optimization
von: Zhao, Yao, et al.
Veröffentlicht: (2024)
von: Zhao, Yao, et al.
Veröffentlicht: (2024)
Kolb-Based Experiential Learning for Generalist Agents with Human-Level Kaggle Data Science Performance
von: Grosnit, Antoine, et al.
Veröffentlicht: (2024)
von: Grosnit, Antoine, et al.
Veröffentlicht: (2024)
ShortCircuit: AlphaZero-Driven Circuit Design
von: Tsaras, Dimitrios, et al.
Veröffentlicht: (2024)
von: Tsaras, Dimitrios, et al.
Veröffentlicht: (2024)
A Tutorial on LLM Reasoning: Relevant Methods behind ChatGPT o1
von: Wang, Jun
Veröffentlicht: (2025)
von: Wang, Jun
Veröffentlicht: (2025)
Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates
von: Diwan, Anish, et al.
Veröffentlicht: (2026)
von: Diwan, Anish, et al.
Veröffentlicht: (2026)
A Brain-like Synergistic Core in LLMs Drives Behaviour and Learning
von: Urbina-Rodriguez, Pedro, et al.
Veröffentlicht: (2026)
von: Urbina-Rodriguez, Pedro, et al.
Veröffentlicht: (2026)
Phase‐Transfer Catalysis for Electrochemical Chlorination and Nitration of Arenes
von: Yingdong Duan, et al.
Veröffentlicht: (2024)
von: Yingdong Duan, et al.
Veröffentlicht: (2024)
Principle-Evolvable Scientific Discovery via Uncertainty Minimization
von: Pu, Yingming, et al.
Veröffentlicht: (2026)
von: Pu, Yingming, et al.
Veröffentlicht: (2026)
Multi-Task GRPO: Reliable LLM Reasoning Across Tasks
von: Ramesh, Shyam Sundhar, et al.
Veröffentlicht: (2026)
von: Ramesh, Shyam Sundhar, et al.
Veröffentlicht: (2026)
Willmore surfaces in 4-dimensional conformal manifolds
von: Wang, Changping, et al.
Veröffentlicht: (2023)
von: Wang, Changping, et al.
Veröffentlicht: (2023)
Research Progress, Challenges, and Prospects of Primary Liver Cancer Organoids
von: Zhenxiao Wang, et al.
Veröffentlicht: (2026)
von: Zhenxiao Wang, et al.
Veröffentlicht: (2026)
A Pragmatist Robot: Learning to Plan Tasks by Experiencing the Real World
von: Qu, Kaixian, et al.
Veröffentlicht: (2025)
von: Qu, Kaixian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Model-Based and Sample-Efficient AI-Assisted Math Discovery in Sphere Packing
von: Tutunov, Rasul, et al.
Veröffentlicht: (2025) -
Why Can Large Language Models Generate Correct Chain-of-Thoughts?
von: Tutunov, Rasul, et al.
Veröffentlicht: (2023) -
Decoding as Optimisation on the Probability Simplex: From Top-K to Top-P (Nucleus) to Best-of-K Samplers
von: Ji, Xiaotong, et al.
Veröffentlicht: (2026) -
Scalable Power Sampling: Unlocking Efficient, Training-Free Reasoning for LLMs via Distribution Sharpening
von: Ji, Xiaotong, et al.
Veröffentlicht: (2026) -
Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2025)