Transfer Learning in Latent Contextual Bandits with Covariate Shift Through Causal Transportability
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Deng, Mingwei, Kyrki, Ville, Baumann, Dominik |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Distributionally Robust Policy Evaluation under General Covariate Shift in Contextual Bandits
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
von: Guo, Yihong, et al.
Veröffentlicht: (2024)
A Lightweight Crowd Model for Robot Social Navigation
von: Eskeri, Maryam Kazemi, et al.
Veröffentlicht: (2025)
von: Eskeri, Maryam Kazemi, et al.
Veröffentlicht: (2025)
Bayesian Floor Field: Transferring people flow predictions across environments
von: Verdoja, Francesco, et al.
Veröffentlicht: (2022)
von: Verdoja, Francesco, et al.
Veröffentlicht: (2022)
Exploring Contextual Representation and Multi-Modality for End-to-End Autonomous Driving
von: Azam, Shoaib, et al.
Veröffentlicht: (2022)
von: Azam, Shoaib, et al.
Veröffentlicht: (2022)
The Role of Higher-Order Cognitive Models in Active Learning
von: Keurulainen, Oskar, et al.
Veröffentlicht: (2024)
von: Keurulainen, Oskar, et al.
Veröffentlicht: (2024)
COSBO: Conservative Offline Simulation-Based Policy Optimization
von: Kargar, Eshagh, et al.
Veröffentlicht: (2024)
von: Kargar, Eshagh, et al.
Veröffentlicht: (2024)
Transfer Learning for Contextual Multi-armed Bandits
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
von: Cai, Changxiao, et al.
Veröffentlicht: (2022)
Automated Feature Selection for Inverse Reinforcement Learning
von: Baimukashev, Daulet, et al.
Veröffentlicht: (2024)
von: Baimukashev, Daulet, et al.
Veröffentlicht: (2024)
On Transportability for Structural Causal Bandits
von: Park, Min Woo, et al.
Veröffentlicht: (2025)
von: Park, Min Woo, et al.
Veröffentlicht: (2025)
DDGC: Generative Deep Dexterous Grasping in Clutter
von: Lundell, Jens, et al.
Veröffentlicht: (2021)
von: Lundell, Jens, et al.
Veröffentlicht: (2021)
Causal Contextual Bandits with Adaptive Context
von: Madhavan, Rahul, et al.
Veröffentlicht: (2024)
von: Madhavan, Rahul, et al.
Veröffentlicht: (2024)
Transfer Learning for Meta-analysis Under Covariate Shift
von: Wang, Zilong, et al.
Veröffentlicht: (2026)
von: Wang, Zilong, et al.
Veröffentlicht: (2026)
Combining Bayesian Inference and Reinforcement Learning for Agent Decision Making: A Review
von: Zhou, Chengmin, et al.
Veröffentlicht: (2025)
von: Zhou, Chengmin, et al.
Veröffentlicht: (2025)
Leveraging Offline Data in Linear Latent Contextual Bandits
von: Kausik, Chinmaya, et al.
Veröffentlicht: (2024)
von: Kausik, Chinmaya, et al.
Veröffentlicht: (2024)
Decentralized Contextual Bandits with Network Adaptivity
von: Deng, Chuyun, et al.
Veröffentlicht: (2025)
von: Deng, Chuyun, et al.
Veröffentlicht: (2025)
Causal Covariate Shift Correction using Fisher information penalty
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
Online Learning of Human Constraints from Feedback in Shared Autonomy
von: Zhu, Shibei, et al.
Veröffentlicht: (2024)
von: Zhu, Shibei, et al.
Veröffentlicht: (2024)
Active Learning for Stochastic Contextual Linear Bandits
von: Brunskill, Emma, et al.
Veröffentlicht: (2026)
von: Brunskill, Emma, et al.
Veröffentlicht: (2026)
Beyond Reweighting: On the Predictive Role of Covariate Shift in Effect Generalization
von: Jin, Ying, et al.
Veröffentlicht: (2024)
von: Jin, Ying, et al.
Veröffentlicht: (2024)
Beam Selection in ISAC using Contextual Bandit with Multi-modal Transformer and Transfer Learning
von: Farzanullah, Mohammad, et al.
Veröffentlicht: (2025)
von: Farzanullah, Mohammad, et al.
Veröffentlicht: (2025)
Learning When to Trust in Contextual Bandits
von: Ghasemi, Majid, et al.
Veröffentlicht: (2026)
von: Ghasemi, Majid, et al.
Veröffentlicht: (2026)
Local Learning for Covariate Selection in Nonparametric Causal Effect Estimation with Latent Variables
von: Li, Zheng, et al.
Veröffentlicht: (2024)
von: Li, Zheng, et al.
Veröffentlicht: (2024)
Latent Covariate Shift: Unlocking Partial Identifiability for Multi-Source Domain Adaptation
von: Liu, Yuhang, et al.
Veröffentlicht: (2022)
von: Liu, Yuhang, et al.
Veröffentlicht: (2022)
Sparse Nonparametric Contextual Bandits
von: Flynn, Hamish, et al.
Veröffentlicht: (2025)
von: Flynn, Hamish, et al.
Veröffentlicht: (2025)
Pessimistic Risk-Aware Policy Learning in Contextual Bandits
von: Wan, Yilong, et al.
Veröffentlicht: (2026)
von: Wan, Yilong, et al.
Veröffentlicht: (2026)
Graph Learning Is Suboptimal in Causal Bandits
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
von: Shahverdikondori, Mohammad, et al.
Veröffentlicht: (2025)
Causal Feature Selection Method for Contextual Multi-Armed Bandits in Recommender System
von: Zhao, Zhenyu, et al.
Veröffentlicht: (2024)
von: Zhao, Zhenyu, et al.
Veröffentlicht: (2024)
CTRL Your Shift: Clustered Transfer Residual Learning for Many Small Datasets
von: Jain, Gauri, et al.
Veröffentlicht: (2025)
von: Jain, Gauri, et al.
Veröffentlicht: (2025)
Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
von: Goyal, Tanmay, et al.
Veröffentlicht: (2025)
Contextual Optimization under Covariate Shift: A Robust Approach by Intersecting Wasserstein Balls
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
von: Wang, Tianyu, et al.
Veröffentlicht: (2024)
Linear Contextual Bandits with Interference
von: Xu, Yang, et al.
Veröffentlicht: (2024)
von: Xu, Yang, et al.
Veröffentlicht: (2024)
Contextual Decision-Making with Knapsacks Beyond the Worst Case
von: Chen, Zhaohua, et al.
Veröffentlicht: (2022)
von: Chen, Zhaohua, et al.
Veröffentlicht: (2022)
Latent Preference Bandits
von: Mwai, Newton, et al.
Veröffentlicht: (2025)
von: Mwai, Newton, et al.
Veröffentlicht: (2025)
Latent Order Bandits
von: Carlsson, Emil, et al.
Veröffentlicht: (2026)
von: Carlsson, Emil, et al.
Veröffentlicht: (2026)
Learning with Incomplete Context: Linear Contextual Bandits with Pretrained Imputation
von: Yan, Hao, et al.
Veröffentlicht: (2025)
von: Yan, Hao, et al.
Veröffentlicht: (2025)
Direction-Aware Offline-to-Online Learning in Linear Contextual Bandits
von: Han, Zean, et al.
Veröffentlicht: (2026)
von: Han, Zean, et al.
Veröffentlicht: (2026)
Wasserstein Distributionally Robust Policy Evaluation and Learning for Contextual Bandits
von: Shen, Yi, et al.
Veröffentlicht: (2023)
von: Shen, Yi, et al.
Veröffentlicht: (2023)
Contextual Bandits for Resource-Constrained Devices using Probabilistic Learning
von: Angioli, Marco, et al.
Veröffentlicht: (2026)
von: Angioli, Marco, et al.
Veröffentlicht: (2026)
Contextual Linear Bandits with Delay as Payoff
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
von: Zhang, Mengxiao, et al.
Veröffentlicht: (2025)
Group-Sensitive Offline Contextual Bandits
von: Guo, Yihong, et al.
Veröffentlicht: (2025)
von: Guo, Yihong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Distributionally Robust Policy Evaluation under General Covariate Shift in Contextual Bandits
von: Guo, Yihong, et al.
Veröffentlicht: (2024) -
A Lightweight Crowd Model for Robot Social Navigation
von: Eskeri, Maryam Kazemi, et al.
Veröffentlicht: (2025) -
Bayesian Floor Field: Transferring people flow predictions across environments
von: Verdoja, Francesco, et al.
Veröffentlicht: (2022) -
Exploring Contextual Representation and Multi-Modality for End-to-End Autonomous Driving
von: Azam, Shoaib, et al.
Veröffentlicht: (2022) -
The Role of Higher-Order Cognitive Models in Active Learning
von: Keurulainen, Oskar, et al.
Veröffentlicht: (2024)