Bayesian Optimistic Optimisation with Exponentially Decaying Regret
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tran-The, Hung, Gupta, Sunil, Rana, Santu, Venkatesh, Svetha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Sub-linear Regret Bounds for Bayesian Optimisation in Unknown Search Spaces
von: Tran-The, Hung, et al.
Veröffentlicht: (2020)
von: Tran-The, Hung, et al.
Veröffentlicht: (2020)
Regret Bounds for Expected Improvement Algorithms in Gaussian Process Bandit Optimization
von: Tran-The, Hung, et al.
Veröffentlicht: (2022)
von: Tran-The, Hung, et al.
Veröffentlicht: (2022)
Trading Convergence Rate with Computational Budget in High Dimensional Bayesian Optimization
von: Tran-The, Hung, et al.
Veröffentlicht: (2019)
von: Tran-The, Hung, et al.
Veröffentlicht: (2019)
Novel Kernel Models and Exact Representor Theory for Neural Networks Beyond the Over-Parameterized Regime
von: Shilton, Alistair, et al.
Veröffentlicht: (2024)
von: Shilton, Alistair, et al.
Veröffentlicht: (2024)
Enhanced Bayesian Optimization via Preferential Modeling of Abstract Properties
von: A V, Arun Kumar, et al.
Veröffentlicht: (2024)
von: A V, Arun Kumar, et al.
Veröffentlicht: (2024)
Efficient Symmetry-Aware Materials Generation via Hierarchical Generative Flow Networks
von: Nguyen, Tri Minh, et al.
Veröffentlicht: (2024)
von: Nguyen, Tri Minh, et al.
Veröffentlicht: (2024)
Score-based Integrated Gradient for Root Cause Explanations of Outliers
von: Nguyen, Phuoc, et al.
Veröffentlicht: (2026)
von: Nguyen, Phuoc, et al.
Veröffentlicht: (2026)
Stable Hadamard Memory: Revitalizing Memory-Augmented Agents for Reinforcement Learning
von: Le, Hung, et al.
Veröffentlicht: (2024)
von: Le, Hung, et al.
Veröffentlicht: (2024)
Continual Fine-Tuning of Large Language Models via Program Memory
von: Le, Hung, et al.
Veröffentlicht: (2026)
von: Le, Hung, et al.
Veröffentlicht: (2026)
Revisiting the Dataset Bias Problem from a Statistical Perspective
von: Do, Kien, et al.
Veröffentlicht: (2024)
von: Do, Kien, et al.
Veröffentlicht: (2024)
Composite Concept Extraction through Backdooring
von: Ghosh, Banibrata, et al.
Veröffentlicht: (2024)
von: Ghosh, Banibrata, et al.
Veröffentlicht: (2024)
Adaptive Acquisition Selection for Bayesian Optimization with Large Language Models
von: Ngo, Giang, et al.
Veröffentlicht: (2026)
von: Ngo, Giang, et al.
Veröffentlicht: (2026)
Federated Domain Generalization with Latent Space Inversion
von: Palakkadavath, Ragja, et al.
Veröffentlicht: (2025)
von: Palakkadavath, Ragja, et al.
Veröffentlicht: (2025)
Generating Realistic Tabular Data with Large Language Models
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
Enhancing Length Extrapolation in Sequential Models with Pointer-Augmented Neural Memory
von: Le, Hung, et al.
Veröffentlicht: (2024)
von: Le, Hung, et al.
Veröffentlicht: (2024)
Finding the Trigger: Causal Abductive Reasoning on Video Events
von: Le, Thao Minh, et al.
Veröffentlicht: (2025)
von: Le, Thao Minh, et al.
Veröffentlicht: (2025)
Variable-Agnostic Causal Exploration for Reinforcement Learning
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2024)
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2024)
Beyond Surprise: Improving Exploration Through Surprise Novelty
von: Le, Hung, et al.
Veröffentlicht: (2023)
von: Le, Hung, et al.
Veröffentlicht: (2023)
Reasoning Under 1 Billion: Memory-Augmented Reinforcement Learning for Large Language Models
von: Le, Hung, et al.
Veröffentlicht: (2025)
von: Le, Hung, et al.
Veröffentlicht: (2025)
SPaCe: Unlocking Sample-Efficient Large Language Models Training With Self-Pace Curriculum Learning
von: Do, Dai, et al.
Veröffentlicht: (2025)
von: Do, Dai, et al.
Veröffentlicht: (2025)
High Dimensional Bayesian Optimization using Lasso Variable Selection
von: Hoang, Vu Viet, et al.
Veröffentlicht: (2025)
von: Hoang, Vu Viet, et al.
Veröffentlicht: (2025)
Uncertainty-Guided Checkpoint Selection for Reinforcement Finetuning of Large Language Models
von: Nguyen, Manh, et al.
Veröffentlicht: (2025)
von: Nguyen, Manh, et al.
Veröffentlicht: (2025)
Bayesian Optimization for Unknown Cost-Varying Variable Subsets with No-Regret Costs
von: Hoang, Vu Viet, et al.
Veröffentlicht: (2024)
von: Hoang, Vu Viet, et al.
Veröffentlicht: (2024)
Large Language Models for Imbalanced Classification: Diversity makes the difference
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Tail Distribution of Regret in Optimistic Reinforcement Learning
von: Khodadadian, Sajad, et al.
Veröffentlicht: (2025)
von: Khodadadian, Sajad, et al.
Veröffentlicht: (2025)
ChargeFlow: Flow-Matching Refinement of Charge-Conditioned Electron Densities
von: Nguyen, Tri Minh, et al.
Veröffentlicht: (2026)
von: Nguyen, Tri Minh, et al.
Veröffentlicht: (2026)
Multi-Reference Preference Optimization for Large Language Models
von: Le, Hung, et al.
Veröffentlicht: (2024)
von: Le, Hung, et al.
Veröffentlicht: (2024)
Bayesian Optimisation with Unknown Hyperparameters: Regret Bounds Logarithmically Closer to Optimal
von: Ziomek, Juliusz, et al.
Veröffentlicht: (2024)
von: Ziomek, Juliusz, et al.
Veröffentlicht: (2024)
TRUST: Test-time Resource Utilization for Superior Trustworthiness
von: Harikumar, Haripriya, et al.
Veröffentlicht: (2025)
von: Harikumar, Haripriya, et al.
Veröffentlicht: (2025)
PINN-BO: A Black-box Optimization Algorithm using Physics-Informed Neural Networks
von: Phan-Trong, Dat, et al.
Veröffentlicht: (2024)
von: Phan-Trong, Dat, et al.
Veröffentlicht: (2024)
Optimistic Regret Bounds for Online Learning in Adversarial Markov Decision Processes
von: Moon, Sang Bin, et al.
Veröffentlicht: (2024)
von: Moon, Sang Bin, et al.
Veröffentlicht: (2024)
Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees
von: Ganguly, Sourav, et al.
Veröffentlicht: (2026)
von: Ganguly, Sourav, et al.
Veröffentlicht: (2026)
Policy Learning for Off-Dynamics RL with Deficient Support
von: Van, Linh Le Pham, et al.
Veröffentlicht: (2024)
von: Van, Linh Le Pham, et al.
Veröffentlicht: (2024)
Optimistic Thompson Sampling for No-Regret Learning in Unknown Games
von: Li, Yingru, et al.
Veröffentlicht: (2024)
von: Li, Yingru, et al.
Veröffentlicht: (2024)
Minimizing Weighted Counterfactual Regret with Optimistic Online Mirror Descent
von: Xu, Hang, et al.
Veröffentlicht: (2024)
von: Xu, Hang, et al.
Veröffentlicht: (2024)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
von: Vakili, Sattar, et al.
Veröffentlicht: (2024)
von: Vakili, Sattar, et al.
Veröffentlicht: (2024)
Causal Discovery via Bayesian Optimization
von: Duong, Bao, et al.
Veröffentlicht: (2025)
von: Duong, Bao, et al.
Veröffentlicht: (2025)
Leveraging Human Feedback for Semantically-Relevant Skill Discovery
von: Hussonnois, Maxence, et al.
Veröffentlicht: (2026)
von: Hussonnois, Maxence, et al.
Veröffentlicht: (2026)
Probabilities Are All You Need: A Probability-Only Approach to Uncertainty Estimation in Large Language Models
von: Nguyen, Manh, et al.
Veröffentlicht: (2025)
von: Nguyen, Manh, et al.
Veröffentlicht: (2025)
Retrieval-augmented Decoding for Improving Truthfulness in Open-ended Generation
von: Nguyen, Manh, et al.
Veröffentlicht: (2025)
von: Nguyen, Manh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Sub-linear Regret Bounds for Bayesian Optimisation in Unknown Search Spaces
von: Tran-The, Hung, et al.
Veröffentlicht: (2020) -
Regret Bounds for Expected Improvement Algorithms in Gaussian Process Bandit Optimization
von: Tran-The, Hung, et al.
Veröffentlicht: (2022) -
Trading Convergence Rate with Computational Budget in High Dimensional Bayesian Optimization
von: Tran-The, Hung, et al.
Veröffentlicht: (2019) -
Novel Kernel Models and Exact Representor Theory for Neural Networks Beyond the Over-Parameterized Regime
von: Shilton, Alistair, et al.
Veröffentlicht: (2024) -
Enhanced Bayesian Optimization via Preferential Modeling of Abstract Properties
von: A V, Arun Kumar, et al.
Veröffentlicht: (2024)