An advantage based policy transfer algorithm for reinforcement learning with measures of transferability
Fuente:
arXiv
Saved in:
| Main Authors: | Alam, Md Ferdous, Naghizadeh, Parinaz, Hoelzle, David |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Task diversity produces systematic transfer but inhibits continual reinforcement learning
by: Seth, Purab, et al.
Published: (2026)
by: Seth, Purab, et al.
Published: (2026)
Bellman operator convergence enhancements in reinforcement learning algorithms
by: Kadurha, David Krame, et al.
Published: (2025)
by: Kadurha, David Krame, et al.
Published: (2025)
Representation Learning for Sequential Volumetric Design Tasks
by: Alam, Md Ferdous, et al.
Published: (2023)
by: Alam, Md Ferdous, et al.
Published: (2023)
Anticipating Gaming to Incentivize Improvement: Guiding Agents in (Fair) Strategic Classification
by: Alhanouti, Sura, et al.
Published: (2025)
by: Alhanouti, Sura, et al.
Published: (2025)
Counterfactual experience augmented off-policy reinforcement learning
by: Lee, Sunbowen, et al.
Published: (2025)
by: Lee, Sunbowen, et al.
Published: (2025)
Physics-informed transfer learning for SHM via feature selection
by: Poole, J., et al.
Published: (2025)
by: Poole, J., et al.
Published: (2025)
Deep reinforcement learning with time-scale invariant memory
by: Kabir, Md Rysul, et al.
Published: (2024)
by: Kabir, Md Rysul, et al.
Published: (2024)
On-site estimation of battery electrochemical parameters via transfer learning based physics-informed neural network approach
by: Yeregui, Josu, et al.
Published: (2025)
by: Yeregui, Josu, et al.
Published: (2025)
Convergence of a model-free entropy-regularized inverse reinforcement learning algorithm
by: Renard, Titouan, et al.
Published: (2024)
by: Renard, Titouan, et al.
Published: (2024)
DCD: Decomposition-based Causal Discovery from Autocorrelated and Non-Stationary Temporal Data
by: Ferdous, Muhammad Hasan, et al.
Published: (2026)
by: Ferdous, Muhammad Hasan, et al.
Published: (2026)
Curriculum reinforcement learning with measurable task representation learning
by: Wen, Yongyan, et al.
Published: (2026)
by: Wen, Yongyan, et al.
Published: (2026)
Trust Region Preference Approximation: A simple and stable reinforcement learning algorithm for LLM reasoning
by: Su, Xuerui, et al.
Published: (2025)
by: Su, Xuerui, et al.
Published: (2025)
Attractor learning for spatiotemporally chaotic dynamical systems using echo state networks with transfer learning
by: Alam, Mohammad Shah, et al.
Published: (2025)
by: Alam, Mohammad Shah, et al.
Published: (2025)
Causal prompting model-based offline reinforcement learning
by: Yu, Xuehui, et al.
Published: (2024)
by: Yu, Xuehui, et al.
Published: (2024)
From high-frequency sensors to noon reports: Using transfer learning for shaft power prediction in maritime
by: Sharma, Akriti, et al.
Published: (2025)
by: Sharma, Akriti, et al.
Published: (2025)
On the transferability of Sparse Autoencoders for interpreting compressed models
by: Gupte, Suchit, et al.
Published: (2025)
by: Gupte, Suchit, et al.
Published: (2025)
Xeno-learning: knowledge transfer across species in deep learning-based spectral image analysis
by: Sellner, Jan, et al.
Published: (2024)
by: Sellner, Jan, et al.
Published: (2024)
Exploring CausalWorld: Enhancing robotic manipulation via knowledge transfer and curriculum learning
by: Wang, Xinrui, et al.
Published: (2024)
by: Wang, Xinrui, et al.
Published: (2024)
Learning for Dynamic Combinatorial Optimization without Training Data
by: Liao, Yiqiao, et al.
Published: (2025)
by: Liao, Yiqiao, et al.
Published: (2025)
Adaptive Bounded Exploration and Intermediate Actions for Data Debiasing
by: Yang, Yifan, et al.
Published: (2025)
by: Yang, Yifan, et al.
Published: (2025)
Test-Time Adaptation for Unsupervised Combinatorial Optimization
by: Liao, Yiqiao, et al.
Published: (2026)
by: Liao, Yiqiao, et al.
Published: (2026)
Generalization Error Bounds for Learning under Censored Feedback
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
Economic span selection of bridge based on deep reinforcement learning
by: Zhang, Leye, et al.
Published: (2024)
by: Zhang, Leye, et al.
Published: (2024)
Predicting trucking accidents with truck drivers 'safety climate perception across companies: A transfer learning approach
by: Sun, Kailai, et al.
Published: (2024)
by: Sun, Kailai, et al.
Published: (2024)
Survival and grade of the glioma prediction using transfer learning
by: Rubio, Santiago Valbuena, et al.
Published: (2024)
by: Rubio, Santiago Valbuena, et al.
Published: (2024)
Normalization and effective learning rates in reinforcement learning
by: Lyle, Clare, et al.
Published: (2024)
by: Lyle, Clare, et al.
Published: (2024)
Scalable and reliable deep transfer learning for intelligent fault detection via multi-scale neural processes embedded with knowledge
by: Li, Zhongzhi, et al.
Published: (2024)
by: Li, Zhongzhi, et al.
Published: (2024)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Policy-shaped prediction: avoiding distractions in model-based reinforcement learning
by: Hutson, Miles, et al.
Published: (2024)
by: Hutson, Miles, et al.
Published: (2024)
Occam's model: Selecting simpler representations for better transferability estimation
by: Singh, Prabhant, et al.
Published: (2025)
by: Singh, Prabhant, et al.
Published: (2025)
Continual learning under domain transfer with sparse synaptic bursting
by: Beaulieu, Shawn L., et al.
Published: (2021)
by: Beaulieu, Shawn L., et al.
Published: (2021)
X-ray transferable polyrepresentation learning
by: Hryniewska-Guzik, Weronika, et al.
Published: (2025)
by: Hryniewska-Guzik, Weronika, et al.
Published: (2025)
Automating proton PBS treatment planning for head and neck cancers using policy gradient-based deep reinforcement learning
by: Wang, Qingqing, et al.
Published: (2024)
by: Wang, Qingqing, et al.
Published: (2024)
CHARME: A chain-based reinforcement learning approach for the minor embedding problem
by: Ngo, Hoang M., et al.
Published: (2024)
by: Ngo, Hoang M., et al.
Published: (2024)
Friends in Unexpected Places: Enhancing Local Fairness in Federated Learning through Clustering
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
Soft $Q(λ)$: A multi-step off-policy method for entropy regularised reinforcement learning using eligibility traces
by: Mahajan, Pranav, et al.
Published: (2026)
by: Mahajan, Pranav, et al.
Published: (2026)
Simulation-based reinforcement learning for real-world autonomous driving
by: Osiński, Błażej, et al.
Published: (2019)
by: Osiński, Błażej, et al.
Published: (2019)
$μ$pscaling small models: Principled warm starts and hyperparameter transfer
by: Ma, Yuxin, et al.
Published: (2026)
by: Ma, Yuxin, et al.
Published: (2026)
Properties that allow or prohibit transferability of adversarial attacks among quantized networks
by: Shrestha, Abhishek, et al.
Published: (2024)
by: Shrestha, Abhishek, et al.
Published: (2024)
BenchRL-QAS: Benchmarking reinforcement learning algorithms for quantum architecture search
by: Ikhtiarudin, Azhar, et al.
Published: (2025)
by: Ikhtiarudin, Azhar, et al.
Published: (2025)
Similar Items
-
Task diversity produces systematic transfer but inhibits continual reinforcement learning
by: Seth, Purab, et al.
Published: (2026) -
Bellman operator convergence enhancements in reinforcement learning algorithms
by: Kadurha, David Krame, et al.
Published: (2025) -
Representation Learning for Sequential Volumetric Design Tasks
by: Alam, Md Ferdous, et al.
Published: (2023) -
Anticipating Gaming to Incentivize Improvement: Guiding Agents in (Fair) Strategic Classification
by: Alhanouti, Sura, et al.
Published: (2025) -
Counterfactual experience augmented off-policy reinforcement learning
by: Lee, Sunbowen, et al.
Published: (2025)