Task diversity produces systematic transfer but inhibits continual reinforcement learning
Fuente:
arXiv
Saved in:
| Main Authors: | Seth, Purab, Shah, Neil, Jha, Kunal, Gershman, Samuel J., Kleiman-Weiner, Max, Carvalho, Wilka |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
NiceWebRL: a Python library for human subject experiments with reinforcement learning environments
by: Carvalho, Wilka, et al.
Published: (2025)
by: Carvalho, Wilka, et al.
Published: (2025)
Modeling Others' Minds as Code
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Predictive representations: building blocks of intelligence
by: Carvalho, Wilka, et al.
Published: (2024)
by: Carvalho, Wilka, et al.
Published: (2024)
Value Internalization: Learning and Generalizing from Social Reward
by: Rong, Frieda, et al.
Published: (2024)
by: Rong, Frieda, et al.
Published: (2024)
Estimating the Empowerment of Language Model Agents
by: Song, Jinyeop, et al.
Published: (2025)
by: Song, Jinyeop, et al.
Published: (2025)
Evaluating LLMs in Open-Source Games
by: Sistla, Swadesh, et al.
Published: (2025)
by: Sistla, Swadesh, et al.
Published: (2025)
Fast weight programming and linear transformers: from machine learning to neurobiology
by: Irie, Kazuki, et al.
Published: (2025)
by: Irie, Kazuki, et al.
Published: (2025)
Soup to go: mitigating forgetting during continual learning with model averaging
by: Kleiman, Anat, et al.
Published: (2025)
by: Kleiman, Anat, et al.
Published: (2025)
An advantage based policy transfer algorithm for reinforcement learning with measures of transferability
by: Alam, Md Ferdous, et al.
Published: (2023)
by: Alam, Md Ferdous, et al.
Published: (2023)
When Empowerment Disempowers
by: Yang, Claire, et al.
Published: (2025)
by: Yang, Claire, et al.
Published: (2025)
Generative Value Conflicts Reveal LLM Priorities
by: Liu, Andy, et al.
Published: (2025)
by: Liu, Andy, et al.
Published: (2025)
The Lock-in Hypothesis: Stagnation by Algorithm
by: Qiu, Tianyi Alex, et al.
Published: (2025)
by: Qiu, Tianyi Alex, et al.
Published: (2025)
Naturalistic Computational Cognitive Science: Towards generalizable models and theories that capture the full range of natural behavior
by: Carvalho, Wilka, et al.
Published: (2025)
by: Carvalho, Wilka, et al.
Published: (2025)
Artificial intelligence for science: The easy and hard problems
by: Battleday, Ruairidh M., et al.
Published: (2024)
by: Battleday, Ruairidh M., et al.
Published: (2024)
The impact of behavioral diversity in multi-agent reinforcement learning
by: Bettini, Matteo, et al.
Published: (2024)
by: Bettini, Matteo, et al.
Published: (2024)
Key-value memory in the brain
by: Gershman, Samuel J., et al.
Published: (2025)
by: Gershman, Samuel J., et al.
Published: (2025)
Deep reinforcement learning for weakly coupled MDP's with continuous actions
by: Robledo, Francisco, et al.
Published: (2024)
by: Robledo, Francisco, et al.
Published: (2024)
Scores as Actions: a framework of fine-tuning diffusion models by continuous-time reinforcement learning
by: Zhao, Hanyang, et al.
Published: (2024)
by: Zhao, Hanyang, et al.
Published: (2024)
Application of linear regression and quasi-Newton methods to the deep reinforcement learning in continuous action cases
by: Komatsu, Hisato
Published: (2025)
by: Komatsu, Hisato
Published: (2025)
Preemptive Solving of Future Problems: Multitask Preplay in Humans and Machines
by: Carvalho, Wilka, et al.
Published: (2025)
by: Carvalho, Wilka, et al.
Published: (2025)
Subjective functions
by: Gershman, Samuel J.
Published: (2025)
by: Gershman, Samuel J.
Published: (2025)
On the consistency of hyper-parameter selection in value-based deep reinforcement learning
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
In value-based deep reinforcement learning, a pruned network is a good network
by: Obando-Ceron, Johan, et al.
Published: (2024)
by: Obando-Ceron, Johan, et al.
Published: (2024)
Multi-Objective Constraint Inference using Inverse reinforcement learning
by: Shah, Syed Ihtesham Hussain, et al.
Published: (2026)
by: Shah, Syed Ihtesham Hussain, et al.
Published: (2026)
Normalization and effective learning rates in reinforcement learning
by: Lyle, Clare, et al.
Published: (2024)
by: Lyle, Clare, et al.
Published: (2024)
Found-RL: foundation model-enhanced reinforcement learning for autonomous driving
by: Qu, Yansong, et al.
Published: (2026)
by: Qu, Yansong, et al.
Published: (2026)
Curriculum reinforcement learning with measurable task representation learning
by: Wen, Yongyan, et al.
Published: (2026)
by: Wen, Yongyan, et al.
Published: (2026)
A step toward a reinforcement learning de novo genome assembler
by: Padovani, Kleber, et al.
Published: (2021)
by: Padovani, Kleber, et al.
Published: (2021)
Sycophantic Chatbots Cause Delusional Spiraling, Even in Ideal Bayesians
by: Chandra, Kartik, et al.
Published: (2026)
by: Chandra, Kartik, et al.
Published: (2026)
CLadder: Assessing Causal Reasoning in Language Models
by: Jin, Zhijing, et al.
Published: (2023)
by: Jin, Zhijing, et al.
Published: (2023)
Counterfactual experience augmented off-policy reinforcement learning
by: Lee, Sunbowen, et al.
Published: (2025)
by: Lee, Sunbowen, et al.
Published: (2025)
Causal prompting model-based offline reinforcement learning
by: Yu, Xuehui, et al.
Published: (2024)
by: Yu, Xuehui, et al.
Published: (2024)
Deep reinforcement learning with time-scale invariant memory
by: Kabir, Md Rysul, et al.
Published: (2024)
by: Kabir, Md Rysul, et al.
Published: (2024)
Delayed homomorphic reinforcement learning for environments with delayed feedback
by: Lee, Jongsoo, et al.
Published: (2026)
by: Lee, Jongsoo, et al.
Published: (2026)
Offline reinforcement learning for job-shop scheduling problems
by: Echeverria, Imanol, et al.
Published: (2024)
by: Echeverria, Imanol, et al.
Published: (2024)
Bellman operator convergence enhancements in reinforcement learning algorithms
by: Kadurha, David Krame, et al.
Published: (2025)
by: Kadurha, David Krame, et al.
Published: (2025)
Local vs Global continual learning
by: Lanzillotta, Giulia, et al.
Published: (2024)
by: Lanzillotta, Giulia, et al.
Published: (2024)
Leveraging weights signals -- Predicting and improving generalizability in reinforcement learning
by: Moulin, Olivier, et al.
Published: (2025)
by: Moulin, Olivier, et al.
Published: (2025)
Dynamic feature selection in medical predictive monitoring by reinforcement learning
by: Chen, Yutong, et al.
Published: (2024)
by: Chen, Yutong, et al.
Published: (2024)
Similar Items
-
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
by: Jha, Kunal, et al.
Published: (2025) -
NiceWebRL: a Python library for human subject experiments with reinforcement learning environments
by: Carvalho, Wilka, et al.
Published: (2025) -
Modeling Others' Minds as Code
by: Jha, Kunal, et al.
Published: (2025) -
Predictive representations: building blocks of intelligence
by: Carvalho, Wilka, et al.
Published: (2024) -
Value Internalization: Learning and Generalizing from Social Reward
by: Rong, Frieda, et al.
Published: (2024)