Model-Based Transfer Learning for Contextual Reinforcement Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Cho, Jung-Hoon, Jayawardana, Vindula, Li, Sirui, Wu, Cathy |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Multi-residual Mixture of Experts Learning for Cooperative Control in Multi-vehicle Systems
par: Jayawardana, Vindula, et autres
Publié: (2025)
par: Jayawardana, Vindula, et autres
Publié: (2025)
Generalizing Cooperative Eco-driving via Multi-residual Task Learning
par: Jayawardana, Vindula, et autres
Publié: (2024)
par: Jayawardana, Vindula, et autres
Publié: (2024)
Structure Detection for Contextual Reinforcement Learning
par: Zhou, Tianyue, et autres
Publié: (2026)
par: Zhou, Tianyue, et autres
Publié: (2026)
IntersectionZoo: Eco-driving for Benchmarking Multi-Agent Contextual Reinforcement Learning
par: Jayawardana, Vindula, et autres
Publié: (2024)
par: Jayawardana, Vindula, et autres
Publié: (2024)
Temporal Transfer Learning for Traffic Optimization with Coarse-grained Advisory Autonomy
par: Cho, Jung-Hoon, et autres
Publié: (2023)
par: Cho, Jung-Hoon, et autres
Publié: (2023)
NeuralMOVES: A lightweight and microscopic vehicle emission estimation model based on reverse engineering and surrogate learning
par: Ramirez-Sanchez, Edgar, et autres
Publié: (2025)
par: Ramirez-Sanchez, Edgar, et autres
Publié: (2025)
Route Recommendations for Traffic Management Under Learned Partial Driver Compliance
par: Bang, Heeseung, et autres
Publié: (2025)
par: Bang, Heeseung, et autres
Publié: (2025)
Expert with Clustering: Hierarchical Online Preference Learning Framework
par: Zhou, Tianyue, et autres
Publié: (2024)
par: Zhou, Tianyue, et autres
Publié: (2024)
The Nah Bandit: Modeling User Non-compliance in Recommendation Systems
par: Zhou, Tianyue, et autres
Publié: (2024)
par: Zhou, Tianyue, et autres
Publié: (2024)
Learning to Segment for Vehicle Routing Problems
par: Ouyang, Wenbin, et autres
Publié: (2025)
par: Ouyang, Wenbin, et autres
Publié: (2025)
Mitigating Metropolitan Carbon Emissions with Dynamic Eco-driving at Scale
par: Jayawardana, Vindula, et autres
Publié: (2024)
par: Jayawardana, Vindula, et autres
Publié: (2024)
Learning-Guided Rolling Horizon Optimization for Long-Horizon Flexible Job-Shop Scheduling
par: Li, Sirui, et autres
Publié: (2025)
par: Li, Sirui, et autres
Publié: (2025)
Towards Foundation Models for Mixed Integer Linear Programming
par: Li, Sirui, et autres
Publié: (2024)
par: Li, Sirui, et autres
Publié: (2024)
Transfer Learning for Contextual Multi-armed Bandits
par: Cai, Changxiao, et autres
Publié: (2022)
par: Cai, Changxiao, et autres
Publié: (2022)
Cooperative Advisory Residual Policies for Congestion Mitigation
par: Hasan, Aamir, et autres
Publié: (2024)
par: Hasan, Aamir, et autres
Publié: (2024)
Contextual Latent World Models for Offline Meta Reinforcement Learning
par: Nakheai, Mohammadreza, et autres
Publié: (2026)
par: Nakheai, Mohammadreza, et autres
Publié: (2026)
Contextual Intelligence The Next Leap for Reinforcement Learning
par: Biedenkapp, André
Publié: (2026)
par: Biedenkapp, André
Publié: (2026)
Contextual Pre-planning on Reward Machine Abstractions for Enhanced Transfer in Deep Reinforcement Learning
par: Azran, Guy, et autres
Publié: (2023)
par: Azran, Guy, et autres
Publié: (2023)
Energy-Based Transfer for Reinforcement Learning
par: Deng, Zeyun, et autres
Publié: (2025)
par: Deng, Zeyun, et autres
Publié: (2025)
Learning to Compress Time-to-Control: A Reinforcement Learning Framework for Chronic Disease Management
par: Singh, Prabhjot, et autres
Publié: (2026)
par: Singh, Prabhjot, et autres
Publié: (2026)
Transfer Learning for Nonparametric Contextual Dynamic Pricing
par: Wang, Fan, et autres
Publié: (2025)
par: Wang, Fan, et autres
Publié: (2025)
Learning Where It Matters: Geometric Anchoring for Robust Preference Alignment
par: Cho, Youngjae, et autres
Publié: (2026)
par: Cho, Youngjae, et autres
Publié: (2026)
Self Paced Gaussian Contextual Reinforcement Learning
par: Ardakani, Mohsen Sahraei, et autres
Publié: (2026)
par: Ardakani, Mohsen Sahraei, et autres
Publié: (2026)
Transfer of Reinforcement Learning-Based Controllers from Model- to Hardware-in-the-Loop
par: Picerno, Mario, et autres
Publié: (2023)
par: Picerno, Mario, et autres
Publié: (2023)
Learning and Transferring Sparse Contextual Bigrams with Linear Transformers
par: Ren, Yunwei, et autres
Publié: (2024)
par: Ren, Yunwei, et autres
Publié: (2024)
Tighter Regret Bounds for Contextual Action-Set Reinforcement Learning
par: Chen, Zijun, et autres
Publié: (2026)
par: Chen, Zijun, et autres
Publié: (2026)
What is a typical signalized intersection in a city? A pipeline for intersection data imputation from OpenStreetMap
par: Qu, Ao, et autres
Publié: (2024)
par: Qu, Ao, et autres
Publié: (2024)
Pessimistic Risk-Aware Policy Learning in Contextual Bandits
par: Wan, Yilong, et autres
Publié: (2026)
par: Wan, Yilong, et autres
Publié: (2026)
Self-Abstraction Learning for Effective and Stable Training of Deep Neural Networks
par: Cho, Wonyong, et autres
Publié: (2026)
par: Cho, Wonyong, et autres
Publié: (2026)
Learning Personalized Ad Impact via Contextual Reinforcement Learning under Delayed Rewards
par: Cheng, Yuwei, et autres
Publié: (2025)
par: Cheng, Yuwei, et autres
Publié: (2025)
STAS: Spatial-Temporal Return Decomposition for Multi-agent Reinforcement Learning
par: Chen, Sirui, et autres
Publié: (2023)
par: Chen, Sirui, et autres
Publié: (2023)
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
par: Lu, Xiaodong, et autres
Publié: (2026)
par: Lu, Xiaodong, et autres
Publié: (2026)
Low-Rank Contextual Reinforcement Learning from Heterogeneous Human Feedback
par: Lee, Seong Jin, et autres
Publié: (2024)
par: Lee, Seong Jin, et autres
Publié: (2024)
Transfer Learning in Latent Contextual Bandits with Covariate Shift Through Causal Transportability
par: Deng, Mingwei, et autres
Publié: (2025)
par: Deng, Mingwei, et autres
Publié: (2025)
Privacy-Preserving Reinforcement Learning from Human Feedback via Decoupled Reward Modeling
par: Cho, Young Hyun, et autres
Publié: (2026)
par: Cho, Young Hyun, et autres
Publié: (2026)
HCRMP: A LLM-Hinted Contextual Reinforcement Learning Framework for Autonomous Driving
par: Chen, Zhiwen, et autres
Publié: (2025)
par: Chen, Zhiwen, et autres
Publié: (2025)
Contextual Bilevel Reinforcement Learning for Incentive Alignment
par: Thoma, Vinzenz, et autres
Publié: (2024)
par: Thoma, Vinzenz, et autres
Publié: (2024)
Multi-Agent Reinforcement Learning for Assessing False-Data Injection Attacks on Transportation Networks
par: Eghtesad, Taha, et autres
Publié: (2023)
par: Eghtesad, Taha, et autres
Publié: (2023)
On Rollouts in Model-Based Reinforcement Learning
par: Frauenknecht, Bernd, et autres
Publié: (2025)
par: Frauenknecht, Bernd, et autres
Publié: (2025)
Model-Based Reinforcement Learning for Atari
par: Kaiser, Lukasz, et autres
Publié: (2019)
par: Kaiser, Lukasz, et autres
Publié: (2019)
Documents similaires
-
Multi-residual Mixture of Experts Learning for Cooperative Control in Multi-vehicle Systems
par: Jayawardana, Vindula, et autres
Publié: (2025) -
Generalizing Cooperative Eco-driving via Multi-residual Task Learning
par: Jayawardana, Vindula, et autres
Publié: (2024) -
Structure Detection for Contextual Reinforcement Learning
par: Zhou, Tianyue, et autres
Publié: (2026) -
IntersectionZoo: Eco-driving for Benchmarking Multi-Agent Contextual Reinforcement Learning
par: Jayawardana, Vindula, et autres
Publié: (2024) -
Temporal Transfer Learning for Traffic Optimization with Coarse-grained Advisory Autonomy
par: Cho, Jung-Hoon, et autres
Publié: (2023)