Symmetric Linear Bandits with Hidden Symmetry
Fuente:
arXiv
Salvato in:
| Autori principali: | Tran, Nam Phuong, Ta, The Anh, Mandal, Debmalya, Tran-Thanh, Long |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning the Expected Core of Strictly Convex Stochastic Cooperative Games
di: Tran, Nam Phuong, et al.
Pubblicazione: (2024)
di: Tran, Nam Phuong, et al.
Pubblicazione: (2024)
Sparse Offline Reinforcement Learning with Corruption Robustness
di: Tran, Nam Phuong, et al.
Pubblicazione: (2025)
di: Tran, Nam Phuong, et al.
Pubblicazione: (2025)
Pruning at Initialisation through the lens of Graphon Limit: Convergence, Expressivity, and Generalisation
di: Pham, Hoang, et al.
Pubblicazione: (2026)
di: Pham, Hoang, et al.
Pubblicazione: (2026)
The Graphon Limit Hypothesis: Understanding Neural Network Pruning via Infinite Width Analysis
di: Pham, Hoang, et al.
Pubblicazione: (2025)
di: Pham, Hoang, et al.
Pubblicazione: (2025)
AF-KAN: Activation Function-Based Kolmogorov-Arnold Networks for Efficient Representation Learning
di: Ta, Hoang-Thang, et al.
Pubblicazione: (2025)
di: Ta, Hoang-Thang, et al.
Pubblicazione: (2025)
Performative Reinforcement Learning with Linear Markov Decision Process
di: Mandal, Debmalya, et al.
Pubblicazione: (2024)
di: Mandal, Debmalya, et al.
Pubblicazione: (2024)
Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization
di: Nguyen, Tuan, et al.
Pubblicazione: (2025)
di: Nguyen, Tuan, et al.
Pubblicazione: (2025)
A Novel Approach in Solving Stochastic Generalized Linear Regression via Nonconvex Programming
di: Anh, Vu Duc, et al.
Pubblicazione: (2024)
di: Anh, Vu Duc, et al.
Pubblicazione: (2024)
Combinations of Fast Activation and Trigonometric Functions in Kolmogorov-Arnold Networks
di: Ta, Hoang-Thang, et al.
Pubblicazione: (2025)
di: Ta, Hoang-Thang, et al.
Pubblicazione: (2025)
Flatness-aware Sequential Learning Generates Resilient Backdoors
di: Pham, Hoang, et al.
Pubblicazione: (2024)
di: Pham, Hoang, et al.
Pubblicazione: (2024)
PRKAN: Parameter-Reduced Kolmogorov-Arnold Networks
di: Ta, Hoang-Thang, et al.
Pubblicazione: (2025)
di: Ta, Hoang-Thang, et al.
Pubblicazione: (2025)
On Corruption-Robustness in Performative Reinforcement Learning
di: Pollatos, Vasilis, et al.
Pubblicazione: (2025)
di: Pollatos, Vasilis, et al.
Pubblicazione: (2025)
Identifying the Best Arm in the Presence of Global Environment Shifts
di: Srisawad, Phurinut, et al.
Pubblicazione: (2024)
di: Srisawad, Phurinut, et al.
Pubblicazione: (2024)
CASUAL: Conditional Support Alignment for Domain Adaptation with Label Shift
di: Nguyen, Anh T, et al.
Pubblicazione: (2023)
di: Nguyen, Anh T, et al.
Pubblicazione: (2023)
Dual-Model Defense: Safeguarding Diffusion Models from Membership Inference Attacks through Disjoint Data Splitting
di: Tran, Bao Q., et al.
Pubblicazione: (2024)
di: Tran, Bao Q., et al.
Pubblicazione: (2024)
Selective Sinkhorn Routing for Improved Sparse Mixture of Experts
di: Nguyen, Duc Anh, et al.
Pubblicazione: (2025)
di: Nguyen, Duc Anh, et al.
Pubblicazione: (2025)
Fourier-Mixed Window Attention: Accelerating Informer for Long Sequence Time-Series Forecasting
di: Tran, Nhat Thanh, et al.
Pubblicazione: (2023)
di: Tran, Nhat Thanh, et al.
Pubblicazione: (2023)
NeuFACO: Neural Focused Ant Colony Optimization for Traveling Salesman Problem
di: Tran, Dat Thanh, et al.
Pubblicazione: (2025)
di: Tran, Dat Thanh, et al.
Pubblicazione: (2025)
High-Dimensional Bayesian Optimization via Random Projection of Manifold Subspaces
di: Nguyen, Quoc-Anh Hoang, et al.
Pubblicazione: (2024)
di: Nguyen, Quoc-Anh Hoang, et al.
Pubblicazione: (2024)
Attack On Prompt: Backdoor Attack in Prompt-Based Continual Learning
di: Nguyen, Trang, et al.
Pubblicazione: (2024)
di: Nguyen, Trang, et al.
Pubblicazione: (2024)
A Hyper-Transformer model for Controllable Pareto Front Learning with Split Feasibility Constraints
di: Tuan, Tran Anh, et al.
Pubblicazione: (2024)
di: Tuan, Tran Anh, et al.
Pubblicazione: (2024)
Performative Reinforcement Learning in Gradually Shifting Environments
di: Rank, Ben, et al.
Pubblicazione: (2024)
di: Rank, Ben, et al.
Pubblicazione: (2024)
GRAPHLCP: Structure-Aware Localized Conformal Prediction on Graphs
di: Baghershahi, Peyman, et al.
Pubblicazione: (2026)
di: Baghershahi, Peyman, et al.
Pubblicazione: (2026)
Building a temperature forecasting model for the city with the regression neural network (RNN)
di: Tran, Nguyen Phuc, et al.
Pubblicazione: (2024)
di: Tran, Nguyen Phuc, et al.
Pubblicazione: (2024)
Distributionally Robust Reinforcement Learning with Human Feedback
di: Mandal, Debmalya, et al.
Pubblicazione: (2025)
di: Mandal, Debmalya, et al.
Pubblicazione: (2025)
FWin transformer for dengue prediction under climate and ocean influence
di: Tran, Nhat Thanh, et al.
Pubblicazione: (2024)
di: Tran, Nhat Thanh, et al.
Pubblicazione: (2024)
Variational Neural Networks
di: Oleksiienko, Illia, et al.
Pubblicazione: (2022)
di: Oleksiienko, Illia, et al.
Pubblicazione: (2022)
Robust SDE Parameter Estimation Under Missing Time Information Setting
di: Van Tran, Long, et al.
Pubblicazione: (2026)
di: Van Tran, Long, et al.
Pubblicazione: (2026)
GROOT: Effective Design of Biological Sequences with Limited Experimental Data
di: Tran, Thanh V. T., et al.
Pubblicazione: (2024)
di: Tran, Thanh V. T., et al.
Pubblicazione: (2024)
Revisiting LARS for Large Batch Training Generalization of Neural Networks
di: Do, Khoi, et al.
Pubblicazione: (2023)
di: Do, Khoi, et al.
Pubblicazione: (2023)
Regret Bounds for Expected Improvement Algorithms in Gaussian Process Bandit Optimization
di: Tran-The, Hung, et al.
Pubblicazione: (2022)
di: Tran-The, Hung, et al.
Pubblicazione: (2022)
BRIDGE: Budget-aware Reasoning via Intermediate Distillation with Guided Examples
di: Le, Xuan-An, et al.
Pubblicazione: (2025)
di: Le, Xuan-An, et al.
Pubblicazione: (2025)
Agnostic Sharpness-Aware Minimization
di: Nguyen, Van-Anh, et al.
Pubblicazione: (2024)
di: Nguyen, Van-Anh, et al.
Pubblicazione: (2024)
Strategyproof Reinforcement Learning from Human Feedback
di: Buening, Thomas Kleine, et al.
Pubblicazione: (2025)
di: Buening, Thomas Kleine, et al.
Pubblicazione: (2025)
Effective Context Modeling Framework for Emotion Recognition in Conversations
di: Van, Cuong Tran, et al.
Pubblicazione: (2024)
di: Van, Cuong Tran, et al.
Pubblicazione: (2024)
Compute the edge p-Laplacian centrality for air traffic network
di: Tran, Loc Hoang, et al.
Pubblicazione: (2025)
di: Tran, Loc Hoang, et al.
Pubblicazione: (2025)
Ensemble Learning for Vietnamese Scene Text Spotting in Urban Environments
di: Nguyen, Hieu, et al.
Pubblicazione: (2024)
di: Nguyen, Hieu, et al.
Pubblicazione: (2024)
Surprisingly Popular Voting for Concentric Rank-Order Models
di: Hosseini, Hadi, et al.
Pubblicazione: (2024)
di: Hosseini, Hadi, et al.
Pubblicazione: (2024)
Early Prediction of Alzheimer's and Related Dementias: A Machine Learning Approach Utilizing Social Determinants of Health Data
di: Kindo, Bereket, et al.
Pubblicazione: (2025)
di: Kindo, Bereket, et al.
Pubblicazione: (2025)
HR-Bandit: Human-AI Collaborated Linear Recourse Bandit
di: Cao, Junyu, et al.
Pubblicazione: (2024)
di: Cao, Junyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Learning the Expected Core of Strictly Convex Stochastic Cooperative Games
di: Tran, Nam Phuong, et al.
Pubblicazione: (2024) -
Sparse Offline Reinforcement Learning with Corruption Robustness
di: Tran, Nam Phuong, et al.
Pubblicazione: (2025) -
Pruning at Initialisation through the lens of Graphon Limit: Convergence, Expressivity, and Generalisation
di: Pham, Hoang, et al.
Pubblicazione: (2026) -
The Graphon Limit Hypothesis: Understanding Neural Network Pruning via Infinite Width Analysis
di: Pham, Hoang, et al.
Pubblicazione: (2025) -
AF-KAN: Activation Function-Based Kolmogorov-Arnold Networks for Efficient Representation Learning
di: Ta, Hoang-Thang, et al.
Pubblicazione: (2025)