MINTS: Minimalist Thompson Sampling
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Wang, Kaizheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Minimalist Bayesian Framework for Stochastic Optimization
von: Wang, Kaizheng
Veröffentlicht: (2025)
von: Wang, Kaizheng
Veröffentlicht: (2025)
Memory-Efficient LLM Pretraining via Minimalist Optimizer Design
von: Glentis, Athanasios, et al.
Veröffentlicht: (2025)
von: Glentis, Athanasios, et al.
Veröffentlicht: (2025)
Optimism Stabilizes Thompson Sampling for Adaptive Inference
von: Yan, Shunxing, et al.
Veröffentlicht: (2026)
von: Yan, Shunxing, et al.
Veröffentlicht: (2026)
Analysis of Thompson Sampling for Controlling Unknown Linear Diffusion Processes
von: Faradonbeh, Mohamad Kazem Shirani, et al.
Veröffentlicht: (2022)
von: Faradonbeh, Mohamad Kazem Shirani, et al.
Veröffentlicht: (2022)
A Stability Principle for Learning under Non-Stationarity
von: Huang, Chengpiao, et al.
Veröffentlicht: (2023)
von: Huang, Chengpiao, et al.
Veröffentlicht: (2023)
A Similarity Measure Between Functions with Applications to Statistical Learning and Optimization
von: Huang, Chengpiao, et al.
Veröffentlicht: (2025)
von: Huang, Chengpiao, et al.
Veröffentlicht: (2025)
Optimal Labeler Assignment and Sampling for Active Learning in the Presence of Imperfect Labels
von: Ahadi, Pouya, et al.
Veröffentlicht: (2025)
von: Ahadi, Pouya, et al.
Veröffentlicht: (2025)
Conformal Prediction-Driven Adaptive Sampling for Digital Water Twins
von: Homaei, Mohammadhossein, et al.
Veröffentlicht: (2025)
von: Homaei, Mohammadhossein, et al.
Veröffentlicht: (2025)
Bucketized Active Sampling for Learning ACOPF
von: Klamkin, Michael, et al.
Veröffentlicht: (2022)
von: Klamkin, Michael, et al.
Veröffentlicht: (2022)
Cactus: Accelerating Auto-Regressive Decoding with Constrained Acceptance Speculative Sampling
von: Hao, Yongchang, et al.
Veröffentlicht: (2026)
von: Hao, Yongchang, et al.
Veröffentlicht: (2026)
Mean-Field Path-Integral Diffusion: From Samples to Interacting Agents
von: Chertkov, Michael
Veröffentlicht: (2026)
von: Chertkov, Michael
Veröffentlicht: (2026)
Asymptotic and Finite Sample Analysis of Nonexpansive Stochastic Approximations with Markovian Noise
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
von: Blaser, Ethan, et al.
Veröffentlicht: (2024)
Provably Safe Generative Sampling with Constricting Barrier Functions
von: Gadginmath, Darshan, et al.
Veröffentlicht: (2026)
von: Gadginmath, Darshan, et al.
Veröffentlicht: (2026)
Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model
von: Mortensen, Oliver, et al.
Veröffentlicht: (2025)
von: Mortensen, Oliver, et al.
Veröffentlicht: (2025)
ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule
von: Huang, Yilie, et al.
Veröffentlicht: (2026)
von: Huang, Yilie, et al.
Veröffentlicht: (2026)
T-SKM-Net: Trainable Neural Network Framework for Linear Constraint Satisfaction via Sampling Kaczmarz-Motzkin Method
von: Zhu, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhu, Haoyu, et al.
Veröffentlicht: (2025)
Achieving Tighter Finite-Time Rates for Heterogeneous Federated Stochastic Approximation under Markovian Sampling
von: Zhu, Feng, et al.
Veröffentlicht: (2025)
von: Zhu, Feng, et al.
Veröffentlicht: (2025)
Gaussian Process Thompson Sampling via Rootfinding
von: Adebiyi, Taiwo A., et al.
Veröffentlicht: (2024)
von: Adebiyi, Taiwo A., et al.
Veröffentlicht: (2024)
Epsilon-Greedy Thompson Sampling to Bayesian Optimization
von: Do, Bach, et al.
Veröffentlicht: (2024)
von: Do, Bach, et al.
Veröffentlicht: (2024)
Feel-Good Thompson Sampling for Contextual Dueling Bandits
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
Variance-Aware Feel-Good Thompson Sampling for Contextual Bandits
von: Li, Xuheng, et al.
Veröffentlicht: (2025)
von: Li, Xuheng, et al.
Veröffentlicht: (2025)
Localized exploration in contextual dynamic pricing achieves dimension-free regret
von: Chai, Jinhang, et al.
Veröffentlicht: (2024)
von: Chai, Jinhang, et al.
Veröffentlicht: (2024)
Contextual Distributionally Robust Optimization with Causal and Continuous Structure: An Interpretable and Tractable Approach
von: Zhang, Fenglin, et al.
Veröffentlicht: (2026)
von: Zhang, Fenglin, et al.
Veröffentlicht: (2026)
Data Uniformity Improves Training Efficiency and More, with a Convergence Framework Beyond the NTK Regime
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
von: Wang, Yuqing, et al.
Veröffentlicht: (2025)
The Sharpness Disparity Principle in Transformers for Accelerating Language Model Pre-Training
von: Wang, Jinbo, et al.
Veröffentlicht: (2025)
von: Wang, Jinbo, et al.
Veröffentlicht: (2025)
Benchmarking PtO and PnO Methods in the Predictive Combinatorial Optimization Regime
von: Geng, Haoyu, et al.
Veröffentlicht: (2023)
von: Geng, Haoyu, et al.
Veröffentlicht: (2023)
Optimizer-Model Consistency: Full Finetuning with the Same Optimizer as Pretraining Forgets Less
von: Liu, Yuxing, et al.
Veröffentlicht: (2026)
von: Liu, Yuxing, et al.
Veröffentlicht: (2026)
Learning the Riccati solution operator for time-varying LQR via Deep Operator Networks
von: Chen, Jun, et al.
Veröffentlicht: (2026)
von: Chen, Jun, et al.
Veröffentlicht: (2026)
OTAD: An Optimal Transport-Induced Robust Model for Agnostic Adversarial Attack
von: Gai, Kuo, et al.
Veröffentlicht: (2024)
von: Gai, Kuo, et al.
Veröffentlicht: (2024)
LLM Serving Optimization with Variable Prefill and Decode Lengths
von: Wang, Meixuan, et al.
Veröffentlicht: (2025)
von: Wang, Meixuan, et al.
Veröffentlicht: (2025)
ORLoopBench: Solver-in-the-Loop Benchmarks for Self-Correction and Behavioral Rationality in Operations Research
von: Ao, Ruicheng, et al.
Veröffentlicht: (2026)
von: Ao, Ruicheng, et al.
Veröffentlicht: (2026)
OptiRepair: Closed-Loop Diagnosis and Repair of Supply Chain Optimization Models with LLM Agents
von: Ao, Ruicheng, et al.
Veröffentlicht: (2026)
von: Ao, Ruicheng, et al.
Veröffentlicht: (2026)
Implicit Regularization of Gradient Flow on One-Layer Softmax Attention
von: Sheen, Heejune, et al.
Veröffentlicht: (2024)
von: Sheen, Heejune, et al.
Veröffentlicht: (2024)
Primal-Dual Spectral Representation for Off-policy Evaluation
von: Hu, Yang, et al.
Veröffentlicht: (2024)
von: Hu, Yang, et al.
Veröffentlicht: (2024)
How Well Can Transformers Emulate In-context Newton's Method?
von: Giannou, Angeliki, et al.
Veröffentlicht: (2024)
von: Giannou, Angeliki, et al.
Veröffentlicht: (2024)
ARO: A New Lens On Matrix Optimization For Large Models
von: Gong, Wenbo, et al.
Veröffentlicht: (2026)
von: Gong, Wenbo, et al.
Veröffentlicht: (2026)
Provable Acceleration of Nesterov's Accelerated Gradient Method over Heavy Ball Method in Training Over-Parameterized Neural Networks
von: Liu, Xin, et al.
Veröffentlicht: (2022)
von: Liu, Xin, et al.
Veröffentlicht: (2022)
From Soliloquy to Agora: Memory-Enhanced LLM Agents with Decentralized Debate for Optimization Modeling
von: Lin, Jianghao, et al.
Veröffentlicht: (2026)
von: Lin, Jianghao, et al.
Veröffentlicht: (2026)
DR-SAC: Distributionally Robust Soft Actor-Critic for Reinforcement Learning under Uncertainty
von: Cui, Mingxuan, et al.
Veröffentlicht: (2025)
von: Cui, Mingxuan, et al.
Veröffentlicht: (2025)
One-Layer Transformer Provably Learns One-Nearest Neighbor In Context
von: Li, Zihao, et al.
Veröffentlicht: (2024)
von: Li, Zihao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Minimalist Bayesian Framework for Stochastic Optimization
von: Wang, Kaizheng
Veröffentlicht: (2025) -
Memory-Efficient LLM Pretraining via Minimalist Optimizer Design
von: Glentis, Athanasios, et al.
Veröffentlicht: (2025) -
Optimism Stabilizes Thompson Sampling for Adaptive Inference
von: Yan, Shunxing, et al.
Veröffentlicht: (2026) -
Analysis of Thompson Sampling for Controlling Unknown Linear Diffusion Processes
von: Faradonbeh, Mohamad Kazem Shirani, et al.
Veröffentlicht: (2022) -
A Stability Principle for Learning under Non-Stationarity
von: Huang, Chengpiao, et al.
Veröffentlicht: (2023)